Skip to content
OpenAIOpenAI

Systems Integration Engineer, Build Systems | Consumer Devices

Build and operate scalable build systems, CI pipelines, and developer infrastructure for consumer-device software. The role requires 5+ years of engineering experience, expertise with Bazel or comparable build systems, and experience improving CI reliability and performance at scale.

About the job

Responsibilities

  • Own and evolve Bazel- and Yocto-based build and test workflows in a polyrepo environment.
  • Design and maintain Starlark rules, macros, toolchains, and integrations for hermetic, reproducible builds.
  • Improve Buildkite CI performance and reliability, including queue time, build time, cache hit rates, retries, and flaky-test isolation.
  • Reduce unnecessary CI work through affected-target detection, dependency-graph analysis, test selection, caching, batching, and smarter scheduling.
  • Unify local and CI workflows to improve build-failure reproduction and debugging.
  • Operate build infrastructure across Docker/OCI images, Kubernetes runners, cloud resources, and remote caching/execution systems.
  • Instrument build and CI systems with metrics, logs, traces, dashboards, and analytics.
  • Partner with engineering users to diagnose build issues, onboard projects, and remove systemic bottlenecks.
  • Apply AI to CI failure analysis, flaky-test debugging, pull-request triage, automated remediation, and developer-facing tooling.
  • Participate in an on-call rotation for critical developer and factory infrastructure.

Requirements

  • 5+ years of engineering experience, including significant experience building developer infrastructure and tooling.
  • Hands-on experience with Bazel, Buck, Gradle, or similar build systems.
  • Understanding of hermetic builds, dependency graphs, caching, sandboxing, and remote execution.
  • Experience building CI systems at scale, particularly where build time, queue time, test flakiness, and developer trust affect engineering velocity.
  • Ability to own production software in environments with strong SLA requirements.
  • Ability to debug distributed build and CI failures across source control, dependency management, containers, runners, remote caches, test frameworks, and service infrastructure.
  • Experience operating in a polyrepo environment with source code from multiple parties.

Preferred Qualifications

  • Strong focus on developer experience and reducing operational toil.
  • Interest in applying AI to developer infrastructure while maintaining quality, reliability, and safety.

Technologies

  • Bazel
  • Starlark
  • Yocto
  • Buildkite
  • Docker
  • OCI images
  • Kubernetes
  • Python
  • Go
  • TypeScript
  • Rust
  • C++
  • Terraform
  • Remote caching
  • Remote execution

Skills

Bazel, Starlark, Yocto, Buildkite, Docker, Kubernetes, Python, Go, TypeScript, Rust, C++, Terraform, Remote Caching, Remote Execution, CI/CD

OpenAI

OpenAI

San Francisco, CA

Network Engineer
$293k+/yrHybridDevOps / SRE

Designs, operates, and improves secure enterprise networks spanning offices, campuses, cloud environments, and connectivity services. The role combines architecture, production operations, troubleshooting, observability, security, and infrastructure automation.

Anthropic

Anthropic

San Francisco, CA

DevOps / AgentOps Engineer, GTM Systems
$320k+/yrHybridDevOps / SRE

Build and operate an AI-first CI/CD and agent-operations platform for Salesforce and custom GTM applications. The role focuses on governed releases, approval workflows, observability, rollback, sandboxing, and SOX-compliant auditability.

Anthropic

Anthropic

San Francisco, CA
Software Engineer, Infrastructure, Interpretability
$320k+/yrHybridDevOps / SRE

Build secure, scalable infrastructure, data systems, compute tooling, and developer experiences for Anthropic’s Interpretability research team. The role partners closely with researchers, security, and platform teams and requires strong programming and infrastructure experience.

Firecrawl

Firecrawl

San Francisco, CA

Cloud DevOps Engineer
$240k+/yrHybrid5+ YOEDevOps / SRE

Build and operate the Kubernetes-based cloud and on-premises infrastructure powering large-scale crawling, search, and ML workloads. The role requires 5+ years in DevOps, platform engineering, or cloud infrastructure, with strong Kubernetes, cloud, Docker, Terraform, and distributed-systems experience.

Fireworks AI

Fireworks AI

San Mateo, CA
Member of Technical Staff - Reliability Engineering
$240k+/yrHybrid5+ YOEDevOps / SRE

Owns reliability standards, incident management, observability, failure testing, and automation for a high-throughput AI infrastructure platform. The role requires deep Linux, networking, software, cloud-native, and distributed-systems experience, along with the ability to influence teams across the organization.