DevOps Engineer
Designs and builds internal compute infrastructure platforms using Palantir products and open-source tools. Requires 3+ years in software development on core infrastructure, expertise in Go/Python/Rust, containers, Kubernetes, and cloud providers.
About the job
Core Responsibilities
- Apply modern engineering practices to improve the maintainability, reliability, and utility of Palantir's internal compute infrastructure.
- Drive build-vs-buy decisions and partner with vendors to integrate external technologies into our platform.
- Collaborate with other teams to understand emerging needs and deliver solutions that help users succeed.
- Participate in the on-call rotation for high-severity incidents affecting critical systems.
- Research and evaluate new technologies to identify where our infrastructure can improve.
What We Value
- Systems programming experience with strong proficiency in Go, Python, or Rust.
- Deep familiarity with containers (Docker) and orchestration (Kubernetes).
- Experience working with a cloud provider (AWS/Azure/GCE), or sysadmin/SRE experience in data centers.
- Up to date with modern industry practices and open-source advancements.
- Solid understanding of distributed systems, APIs, cloud platforms, and onprem infrastructure.
- Hands-on experience with CI/CD pipelines, DevOps practices, and system reliability principles.
What We Require
- 3+ years of professional software development experience on core infrastructure with emphasis on operational excellence.
- 2+ years of experience contributing to the system design or architecture (architecture, design patterns, reliability and scaling) of new and existing systems.
Skills
Go, Python, Rust, Docker, Kubernetes, AWS, Azure, GCP, CI/CD, Distributed Systems
Similar jobs
DevOps / SRE jobsDesigns and operates foundational developer-infrastructure services for CI, builds, deployments, and testing. The role requires senior-level systems engineering, end-to-end service ownership, and cross-functional technical leadership.
Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.
Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.
Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.