Forward Deployed Site Reliability Engineer responsible for building, operating, and maintaining scalable infrastructure in air-gapped on-prem environments for US Government customers. Requires 4+ years Linux admin experience, hardware/networking knowledge, scripting skills, 50% travel availability, and active Top Secret clearance.
Salary not listed
Hybrid4+ YOEDevOps / SRE
About the role
Core Responsibilities
Maintaining availability of physical Linux servers that power the Palantir platform in air-gapped production environments
Design, deploy, and operate infrastructure to support customer & product requirements via modern orchestration & monitoring platforms
Collaborate closely with product teams on requirements & SLOs for deploying software into air-gapped environments
Identifying, troubleshooting, and solving network & systems issues
Scripting to automate away routine operational tasks
Provide technical troubleshooting support for production issues, ensuring timely resolution and minimal impact on operations. Participate in a support on-call schedule
What We Value
Confidence in troubleshooting complex systems issues independently using stack traces and observability & systems tools
Comfort with configuration management, load balancing, monitoring & alerting infrastructure, and container orchestration on small hardware form factors
Demonstrated ability to continuously learn and work independently, making decisions with minimal supervision while working in secure facilities
Experience with containers (Docker/Podman) and orchestration (OpenShift/Kubernetes) at scale is a plus
Preferred Certifications: DOD 8570 IAT Level II or greater (CISSP, Sec+), Unix/Linux Computing Environment (e.g Linux+, RHCE)
What We Require
Available for 50% travel (domestic and international)
4+ years of experience with Linux system administration (RHEL or equivalent preferred)
Experience with hardware environments, including setup, configuration, and management of physical servers and networking equipment
Familiarity with monitoring systems using tools like Prometheus and writing health checks
Proficiency with at least one programming or scripting language, such as Java, Go, Python, JavaScript, Bash, or similar languages
Strong engineering background, preferred in fields such as Computer Science, Mathematics, Software Engineering, Physics, and Data Science
Active US Security Clearance at or above the Top Secret level
Senior individual contributor managing relationships with strategic API partners, driving post-sale success including onboarding, adoption, and expansion. Conducts technical discovery, prototypes solutions using Generative AI APIs, and resolves escalations for Fortune 100 and media clients. Requires 5+ years in technical customer-facing roles.
Build secure, scalable infrastructure, data systems, compute tooling, and developer experiences for Anthropic’s Interpretability research team. The role partners closely with researchers, security, and platform teams and requires strong programming and infrastructure experience.
320k – 485k/yrHybridDevOps / SRE
Software Engineer - Continuous Delivery
BasetenSan Francisco, CA +1
Build and operate continuous delivery infrastructure for Kubernetes deployments across global regions, including progressive rollouts, automated health evaluation, and rollback systems. The role requires strong Go or Python skills, large-scale Kubernetes experience, and familiarity with GitOps tooling.
165k – 330k/yrHybridDevOps / SRE
Platform Engineer II
BestowUnited States
The Platform Engineer II automates cloud infrastructure, builds developer tooling, and improves the reliability and manageability of software platforms. The role requires cloud, infrastructure-as-code, scripting, CI/CD, containerization, and systems administration experience, along with participation in on-call operations.
115k – 130k/yrRemoteDevOps / SRE
DevOps Engineer
Fusion HealthWoodbridge, NJ
Owns secure, scalable Azure infrastructure for healthcare applications, including cloud migrations, Terraform-based automation, CI/CD pipelines, monitoring, and compliance. Requires 3–5+ years of Azure experience and strong DevOps and cloud-security expertise.