Own reliability and scalability of a horizontal identity platform, including core cloud and FedRAMP environments. Build observability, automate operations, lead incident response, and ensure new features are reliable from the start. Requires production SRE experience at scale with Kubernetes, IaC, and strong programming skills.
180k – 250k/yr
Hybrid5+ YOEDevOps / SRE
About the role
What You'll Do
Own the reliability and scalability of our platform—design, build, and operate the infrastructure that keeps C1 running for customers who depend on us. Work on both our core cloud environment, as well as our FedRAMP environment.
Build observability that drives action—create monitoring, alerting, and tooling that helps teams understand system behavior and respond to incidents quickly.
Drive operational excellence across engineering—partner with product teams to ensure new features are built with reliability in mind from the start.
Automate relentlessly—if you're doing something twice, build a system to do it for you. We believe in infrastructure as code and eliminating toil.
Respond to and learn from incidents—lead incident response, conduct blameless postmortems, and drive systemic improvements.
Plan and execute infrastructure projects with incremental deliverables—you'll assess technical risks, communicate tradeoffs, and ship iteratively.
What We're Looking For
You Likely Have
A track record of building and operating production systems at scale—you've kept real systems running for real customers.
Deep experience with cloud infrastructure (AWS, GCP, or similar) and infrastructure-as-code (Terraform, Pulumi, or similar).
Strong programming skills in Go, Python, or similar—you write tools and automation, not just scripts.
Experience with Kubernetes and container orchestration in production.
Strong systems thinking—you understand how components interact and where failures cascade.
High agency—you figure out what needs to be built, not just how to build what you're told. You move fast and unblock yourself.
Deep understanding of observability—you know how to instrument systems, build dashboards, and create alerts that actually matter.
Experience with AI-assisted development (Claude Code, Cursor, Copilot, or similar)—you're already using these tools and excited about what's next.
Clear, persuasive communication—you can explain complex systems to diverse audiences and drive alignment during incidents.
Ego in check—you care about getting it right, not being right.
You Might Also Have
Experience with AI/ML infrastructure or serving LLMs in production.
Background in security-focused environments or compliance frameworks (SOC 2, FedRAMP, etc.).
Experience building developer platforms or internal tooling.
Familiarity with identity systems and protocols (SCIM, SAML, OAuth, LDAP).
Contributions to open source projects or engineering communities.
Compensation & Benefits
Salary range: $180k – $250k
Meaningful equity
Full medical, dental, and vision coverage
In-office in Portland
Solid benefits across the board
Skills
KubernetesTerraformAWSGCPGoPythonObservabilityInfrastructure As CodeIncident ResponseFedRAMP
Infrastructure Engineer building and securing Kubernetes-based platforms, cloud-native deployments, and air-gapped appliances for military planning software. Requires 5+ years production infrastructure experience, deep Kubernetes and cloud expertise, security fundamentals, and full-stack engineering skills in languages like Go or Python.
180k – 290k/yrRemote5+ YOEDevOps / SRE
Software Engineer
xAIPalo Alto, CA
Build and optimize large-scale distributed systems powering xAI's massive supercomputing clusters for AI training. Requires strong systems programming in Rust/C++ and deep Kubernetes/Linux expertise.
180k – 440k/yrOn-site5+ YOEDevOps / SRE
AI Infrastructure Engineer, Sandbox Platform
Scale AISan Francisco, CA +2
Build and evolve a secure, high-performance agent sandboxing platform for code execution. Combine deep systems expertise in isolation/virtualization with strong focus on developer experience, APIs, and internal partnerships. Requires 4+ years in high-performance systems software.
180k – 225k/yrHybrid4+ YOEDevOps / SRE
Platform Engineer
UnusualNew York, NY
Build and scale data ingestion, storage, search, and deployment infrastructure across commercial, FedRAMP High, and IL5 environments. Own platform used by engineers and AI agents for government contracting data, with customer-facing work on integrations and security.
180k – 250k/yrHybrid5+ YOEDevOps / SRE
Platform Engineer
GovsignalsNew York, NY
Platform Engineer building and scaling data ingestion, storage, search, and deployment infrastructure across commercial, FedRAMP High, and IL5 environments. Requires 5+ years in backend/platform/infra roles with deep experience in Postgres, Kubernetes, and TypeScript/Python.