Forward Deployed Site Reliability Engineer - US GovernmentActive$125K–$185K

The opportunity

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more.

What you'll do

  • Maintaining availability of physical Linux servers that power the Palantir: platform in air-gapped production environments

  • Design, deploy, and operate infrastructure to support customer & product: requirements via modern orchestration & monitoring platforms

  • Collaborate closely with product teams on requirements & SLOs for deploying: software into air-gapped environments

  • Identifying, troubleshooting, and solving network & systems issues

  • Scripting to automate away routine operational tasks

  • Provide technical troubleshooting support for production issues, ensuring: timely resolution and minimal impact on operations. Participate in a support on-call schedule

What they're looking for

  • Confidence in troubleshooting complex systems issues independently using: stack traces and observability & systems tools
  • Comfort with configuration management, load balancing, monitoring & alerting: infrastructure, and container orchestration on small hardware form factors.
  • Demonstrated ability to continuously learn and work independently, making: decisions with minimal supervision while working in secure facilities
  • Experience with containers (Docker/Podman) and orchestration (OpenShift/Kubernetes) at scale is a plus