Senior Software Engineer - Pipeline Infrastructure & Integration
Senior engineer owning safety-critical software pipelines and infrastructure, from static and dynamic analysis through CI enforcement, dashboards, and reliability tooling. Requires an advanced technical degree, 7+ years working with large codebases, and expertise in Bazel, Python, backend infrastructure, and C++.
About the job
Responsibilities
- Own the software safety and stability pipeline ecosystem, including Clang-Tidy, AddressSanitizer, ThreadSanitizer, Harden, and Infer, across weekly, nightly, and continuous integration pipelines.
- Advance pipeline prototypes into production-grade, CI-enforced systems across branches.
- Build dashboards and KPIs from issue data to monitor pipeline health and software quality trends.
- Investigate cross-platform and nondeterministic failures, and develop instrumentation and tooling for root-cause analysis.
- Support checker rollouts, testing strategy, and infrastructure scaling.
- Define delivery milestones, automate tooling, report business metrics, participate in code reviews and incident postmortems, and maintain process documentation.
Requirements
- Master's or PhD degree in computer science, electrical engineering, robotics, aerospace, or a related field.
- 7+ years of experience working on large codebases.
- Experience with Bazel and Python.
- Experience building backend infrastructure, including databases, web servers, and shell scripts.
- Experience interfacing with C++ and Python software.
- Ability to create system design diagrams and documentation for customer communication and long-term maintenance.
- Ability to develop scripts for automation, metrics generation, and traceability.
- Deep experience with enterprise-scale deployment practices and related issues.
Nice-to-haves
- Experience with PipeDream, Airflow, and Looker.
- Proficiency in React, TypeScript, and C++, including C++11, C++14, C++17, memory management, threading, and debugging patterns.
Skills
Bazel, Python, C++, React, TypeScript, Shell Scripting, Databases, Web Servers, CI/CD, Clang-Tidy, Addresssanitizer, Threadsanitizer, Airflow, Looker
Similar jobs
DevOps / SRE jobsOwn and evolve a broad infrastructure platform spanning cloud, Kubernetes, deployment, reliability, security, and GPU-backed AI systems. The role requires 8+ years operating production distributed systems, strong incident and architecture experience, and practical cloud infrastructure expertise.
Build and evolve the developer platform that enables reliable, efficient software delivery across the company. The role requires 5+ years of software engineering experience, strong programming and system-design fundamentals, and expertise in build systems, CI/CD, testing, and deployment automation.
Own and improve the CI/CD, testing, and deployment infrastructure that enables fast, safe, observable releases at scale. The role requires strong distributed-systems expertise, hands-on Kubernetes and infrastructure-as-code experience, and a track record of measurable cross-team improvements.
Build and improve cloud infrastructure, developer workflows, and internal tooling that make software development, testing, and releases more efficient and reliable. The role requires cloud architecture knowledge, CI/CD experience, Terraform and Bazel proficiency, and software development skills in Go, Python, or C++.
Build and operate scalable control-plane and data-plane infrastructure for distributed AI workloads, including Ray cluster orchestration, scheduling, observability, and accelerator integration. Requires a bachelor's degree or equivalent experience, 3+ years of production coding, cloud-native expertise, Kubernetes, and Go/Python proficiency.