Skip to content
PalantirPalantirWashington, DC

Forward Deployed Site Reliability Engineer

Forward Deployed Site Reliability Engineer responsible for building, operating, and maintaining scalable infrastructure in air-gapped on-prem environments for US Government customers. Requires 4+ years Linux admin experience, hardware/networking knowledge, scripting skills, 50% travel availability, and active Top Secret clearance.

Salary not listed
Hybrid4+ YOEDevOps / SRE

About the role

Core Responsibilities

  • Maintaining availability of physical Linux servers that power the Palantir platform in air-gapped production environments
  • Design, deploy, and operate infrastructure to support customer & product requirements via modern orchestration & monitoring platforms
  • Collaborate closely with product teams on requirements & SLOs for deploying software into air-gapped environments
  • Identifying, troubleshooting, and solving network & systems issues
  • Scripting to automate away routine operational tasks
  • Provide technical troubleshooting support for production issues, ensuring timely resolution and minimal impact on operations. Participate in a support on-call schedule

What We Value

  • Confidence in troubleshooting complex systems issues independently using stack traces and observability & systems tools
  • Comfort with configuration management, load balancing, monitoring & alerting infrastructure, and container orchestration on small hardware form factors
  • Demonstrated ability to continuously learn and work independently, making decisions with minimal supervision while working in secure facilities
  • Experience with containers (Docker/Podman) and orchestration (OpenShift/Kubernetes) at scale is a plus
  • Preferred Certifications: DOD 8570 IAT Level II or greater (CISSP, Sec+), Unix/Linux Computing Environment (e.g Linux+, RHCE)

What We Require

  • Available for 50% travel (domestic and international)
  • 4+ years of experience with Linux system administration (RHEL or equivalent preferred)
  • Experience with hardware environments, including setup, configuration, and management of physical servers and networking equipment
  • Familiarity with monitoring systems using tools like Prometheus and writing health checks
  • Proficiency with at least one programming or scripting language, such as Java, Go, Python, JavaScript, Bash, or similar languages
  • Strong engineering background, preferred in fields such as Computer Science, Mathematics, Software Engineering, Physics, and Data Science
  • Active US Security Clearance at or above the Top Secret level

Skills

LinuxrhelKubernetesopenshiftDockerpodmanPrometheusPythonGoJavaBashJavaScript

Similar roles

DevOps / SRE jobs
Runway

API Deployment Manager

RunwayNew York, NY +2

Senior individual contributor managing relationships with strategic API partners, driving post-sale success including onboarding, adoption, and expansion. Conducts technical discovery, prototypes solutions using Generative AI APIs, and resolves escalations for Fortune 100 and media clients. Requires 5+ years in technical customer-facing roles.

145k – 270k/yrRemote5+ YOEDevOps / SRE
Anthropic

Software Engineer, Infrastructure, Interpretability

AnthropicSan Francisco, CA +1

Build secure, scalable infrastructure, data systems, compute tooling, and developer experiences for Anthropic’s Interpretability research team. The role partners closely with researchers, security, and platform teams and requires strong programming and infrastructure experience.

320k – 485k/yrHybridDevOps / SRE
Baseten

Software Engineer - Continuous Delivery

BasetenSan Francisco, CA +1

Build and operate continuous delivery infrastructure for Kubernetes deployments across global regions, including progressive rollouts, automated health evaluation, and rollback systems. The role requires strong Go or Python skills, large-scale Kubernetes experience, and familiarity with GitOps tooling.

165k – 330k/yrHybridDevOps / SRE
Bestow

Platform Engineer II

BestowUnited States

The Platform Engineer II automates cloud infrastructure, builds developer tooling, and improves the reliability and manageability of software platforms. The role requires cloud, infrastructure-as-code, scripting, CI/CD, containerization, and systems administration experience, along with participation in on-call operations.

115k – 130k/yrRemoteDevOps / SRE
Fusion Health

DevOps Engineer

Fusion HealthWoodbridge, NJ

Owns secure, scalable Azure infrastructure for healthcare applications, including cloud migrations, Terraform-based automation, CI/CD pipelines, monitoring, and compliance. Requires 3–5+ years of Azure experience and strong DevOps and cloud-security expertise.

120k – 140k/yrHybrid5+ YOEDevOps / SRE