Skip to content
PalantirPalantir

Edge Infrastructure Engineer

Operates and scales reliable edge and bare-metal infrastructure for Palantir’s production platform across isolated, on-premises, and cloud environments. The role requires at least five years managing large-scale systems, strong Linux, Kubernetes, networking, and server-hardware expertise, plus French proficiency and significant travel.

About the job

Responsibilities

  • Ensure the availability of cloud and physical Kubernetes servers powering the Palantir platform in isolated production environments.
  • Design, deploy, and operate infrastructure that meets customer and product requirements using modern orchestration and monitoring platforms.
  • Collaborate closely with product teams on requirements and service-level objectives (SLOs) for deploying software in isolated environments.
  • Identify, troubleshoot, and resolve network and systems issues.
  • Write scripts to automate routine operational tasks.
  • Advise customers on hardware procurement, capacity planning, firewalls, and related infrastructure needs.
  • Participate in the design, deployment, operation, and scaling of new on-premises and edge environments.
  • Share responsibility for diagnosing, resolving, and preventing production issues.

Requirements

  • Minimum 5 years of experience managing medium- to large-scale systems.
  • Ability to collaborate and communicate effectively with teams distributed across a wide geographic area.
  • High level of independence and resilience, with the ability to work remotely with limited supervision and direction.
  • Understanding of database systems such as Oracle and PostgreSQL and how to manage them; database administration expertise is not required.
  • Knowledge of Kubernetes, Cassandra, Hadoop, and distributed systems.
  • Previous experience with server hardware, such as Dell and HP, and network engineering in production data-center environments.
  • Strong hands-on experience with Linux operating systems, Linux administration, and troubleshooting, including related tools such as Logical Volume Manager (LVM) commands.
  • French language proficiency.
  • Willingness to travel 20–40% of the time.

Benefits

  • Support for employee health, well-being, and professional growth.
  • Hybrid work options may be available depending on the team and business needs.

Skills

Kubernetes, Linux, Linux Administration, Lvm, Oracle, Postgres, Cassandra, Hadoop, Distributed Systems, Server Hardware, Network Engineering, Firewalls, Capacity Planning, Monitoring, Automation

Granica

Granica

Remote

Software Engineer, Infrastructure
No salary listedRemote5+ YOEDevOps / SRE

Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.

Supabase

Supabase

Remote

Platform Engineer - Compute Capacity
No salary listedRemote5+ YOEDevOps / SRE

Platform engineer responsible for forecasting and automating compute capacity across regions, including reservations, fleet reconciliation, observability, and cost optimization. Requires 5+ years in infrastructure, SRE, platform, or capacity engineering plus production software and AWS EC2 experience.

Alpaca

Alpaca

Remote

Production Support Engineer
No salary listedRemote4+ YOEDevOps / SRE

Provides hands-on L2 technical escalation support for enterprise customers in the APAC region, troubleshooting distributed systems and APIs while leading root-cause analysis, support process improvements, and technical documentation. Requires 4+ years of support or escalation engineering experience.

PostHog

PostHog

Remote

ClickHouse Operations Engineer
No salary listedRemoteDevOps / SRE

Automate, manage, and optimize large-scale ClickHouse clusters handling trillions of events and 100+ PB data. Build provisioning systems with Terraform, Ansible, Kubernetes; focus on performance, scaling, and bleeding-edge features.

Lightning AI

Lightning AI

Remote

Senior Network Engineer
$150k+/yrRemote5+ YOEDevOps / SRE

The Senior Network Engineer will design, automate, and operate large-scale, high-performance network infrastructure for AI data centers and GPU clusters. The role requires 5+ years of data center networking experience, expertise in spine-leaf fabrics and routing protocols, and familiarity with HPC or GPU-dense environments.