Forward Deployed Infrastructure Engineer - UK Government
Operates, deploys, and improves reliable Palantir infrastructure and services for government environments. The role emphasizes production troubleshooting, automation, scalable systems, cross-functional collaboration, and security clearance eligibility.
About the job
Responsibilities
- Support and operate Palantir software, including monitoring and alerting, configuration management, and upgrades.
- Deploy Palantir products across production environments and migrate to updated infrastructure types.
- Debug, improve, and optimize services and infrastructure for long-term reliability and scalability.
- Automate workflows, processes, and runbooks to reduce manual operations.
- Provide technical troubleshooting support for production issues and participate in an on-call schedule.
- Develop solutions in Palantir Foundry and Apollo to address infrastructure challenges.
Requirements
- Strong engineering background, preferably in Computer Science, Mathematics, Software Engineering, Physics, or Data Science.
- Strong coding proficiency in Java, Go, Python, JavaScript, or similar languages.
- Experience troubleshooting complex systems independently using observability tools and service logs.
- Ability to identify and automate manual tasks.
- Familiarity with large-scale production systems and technologies such as load balancing, monitoring, distributed systems, and configuration management.
- Ability to work autonomously in a rapidly changing environment and collaborate effectively with multifunctional teams.
- Excellent communication and interpersonal skills.
- Security clearance or the ability to obtain security clearance.
Skills
Java, Go, Python, JavaScript, Bash, Monitoring, Alerting, Configuration Management, Distributed Systems, Load Balancing, Observability, Palantir Foundry, Palantir Apollo, Automation, Service Logs
Similar jobs
DevOps / SRE jobsBuild and maintain Cloudflare’s deployment platform, enabling progressive rollouts, health-mediated releases, and automated workflows at scale. The role requires at least four years of software development experience, backend and frontend experience, and comfort with rapid delivery and on-call support.
Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.
Site Reliability Engineers build and operate reliable, scalable production infrastructure across GitLab’s Infrastructure Platforms teams. The role requires strong software engineering and operations fundamentals, Kubernetes and infrastructure-as-code experience, cloud expertise, and comfort with automation, observability, and incident response.
Build and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.
Infrastructure engineer responsible for building and operating highly available cloud systems, automating operations, and improving reliability across a large-scale AI platform. Requires 5+ years of infrastructure or DevOps experience, production Kubernetes, cloud infrastructure, Terraform, and Python or Go.