Software Engineer, Productivity - Networking
Enhances developer productivity for OpenAI's networking team by improving build systems, CI/CD pipelines, test harnesses, and workflows for C++ and Python codebases in multi-server environments. Requires experience with developer tools and infrastructure automation.
About the job
Responsibilities
- Improve development workflows for engineers building and operating OpenAI's networking systems
- Design and improve continuous deployment, release, and validation pipelines
- Build and maintain test harnesses for multi-server, networked, and hardware-backed environments
- Improve iteration speed across C++, Python, and build-system-heavy codebases
- Partner with engineers to identify friction in CI, testing, debugging, and deployment workflows
- Drive testing and reliability strategy for infrastructure components that support large-scale training and inference workloads
- Work closely with centralized developer experience teams while staying deeply embedded with the networking engineers
Requirements
- Motivated by helping other engineers move faster and with more confidence
- Experience with CI/CD, release pipelines, testing infrastructure, or build systems
- Comfortable moving between C++, Python, and build systems such as CMake, Bazel, or Blaze
- Enjoy building test harnesses, automation, and workflow improvements for complex systems
- Excited to learn about networking domain to make the team more effective
- Instinct to fix friction like slow builds, flaky tests, brittle release processes, painful debugging, unclear validation
- Pragmatic, balance high standards with forward progress
- Like going end-to-end: understanding engineers, workflows, codebase, and operational realities
- Self-directed and comfortable operating with ambiguity in high-context infrastructure environment
Skills
C++, Python, CI/CD, Bazel, Cmake, Blaze, Testing Infrastructure, Release Pipelines, Build Systems, Test Harnesses
Similar jobs
DevOps / SRE jobsBuild and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.
Build and operate the Kubernetes-based cloud and on-premises infrastructure powering large-scale crawling, search, and ML workloads. The role requires 5+ years in DevOps, platform engineering, or cloud infrastructure, with strong Kubernetes, cloud, Docker, Terraform, and distributed-systems experience.
Owns reliability standards, incident management, observability, failure testing, and automation for a high-throughput AI infrastructure platform. The role requires deep Linux, networking, software, cloud-native, and distributed-systems experience, along with the ability to influence teams across the organization.
Build and own production-grade AI agent infrastructure across multiple clouds, with responsibility for Kubernetes, Terraform, observability, security, reliability, and automation. Requires 5+ years of cloud infrastructure experience and strong CI/CD, networking, and production operations expertise.
Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.