Skip to content

Software Engineer, Developer Productivity, AI Tools

Build and standardize AI-powered coding tools, agents, and dev environments to accelerate internal software development while maintaining security and quality. Requires experience with productivity tooling for large codebases, container/CI platforms, and AI model APIs.

About the job

What You’ll Do

  • Enable our researchers and engineers to leverage AI to improve coding productivity without compromising code quality
  • Standardize AI coding tools, such as Claude Code, Cursor, and Codex. Help configure, harden, and maintain the best tools, integrating org-wide configurations with individual preferences.
  • Build secure, reproducible agent sandboxes for remote dev & CI testing.
  • Set up golden-path dev environments and guardrails for secrets/PII.
  • Help individual contributors develop their personalized AI-enabled workflow.
  • Track tool usage, reliability, and cost.

Skills and Qualifications

Minimum qualifications:

  • Bachelor’s degree or equivalent industry experience in computer science, engineering, or similar.
  • Experience developing productivity tools and best practices for large codebases.
  • Ability to communicate clearly and work with researchers to build and manage a variety of internal tools.

Preferred qualifications:

  • Hands-on experience with container platforms (e.g. Docker/Kubernetes), modern CI (GitHub Actions/Buildkite), and package management tools (uv).
  • Practical experience with AI coding tools and model APIs (e.g. OSS via vLLM / SGLang / TGI).
  • Solid Linux/networking fundamentals; comfort with secrets management and safe egress.
  • Proficiency in systems programming languages (e.g. Rust) and scripting languages (e.g. Python).

Skills

Docker, Kubernetes, GitHub Actions, Buildkite, Uv, vLLM, Sglang, Tgi, Linux, Rust, Python, Ai Coding Tools, Secrets Management

Thinking Machines Lab

Thinking Machines Lab

San Francisco, CA

Site Reliability Engineer (SRE)
$350k+/yrOn-siteDevOps / SRE

Site Reliability Engineer drives end-to-end reliability for AI fine-tuning platform Tinker, including SLOs, monitoring, incident response, and multi-tenant GPU scheduling. Requires distributed systems experience, software proficiency for reliability, and production incident handling.

Anthropic

Anthropic

San Francisco, CA

DevOps / AgentOps Engineer, GTM Systems
$320k+/yrHybridDevOps / SRE

Build and operate an AI-first CI/CD and agent-operations platform for Salesforce and custom GTM applications. The role focuses on governed releases, approval workflows, observability, rollback, sandboxing, and SOX-compliant auditability.

Anthropic

Anthropic

San Francisco, CA
Software Engineer, Infrastructure, Interpretability
$320k+/yrHybridDevOps / SRE

Build secure, scalable infrastructure, data systems, compute tooling, and developer experiences for Anthropic’s Interpretability research team. The role partners closely with researchers, security, and platform teams and requires strong programming and infrastructure experience.

OpenAI

OpenAI

San Francisco, CA

Systems Integration Engineer, Build Systems | Consumer Devices
$293k+/yrHybrid5+ YOEDevOps / SRE

Build and operate scalable build systems, CI pipelines, and developer infrastructure for consumer-device software. The role requires 5+ years of engineering experience, expertise with Bazel or comparable build systems, and experience improving CI reliability and performance at scale.

OpenAI

OpenAI

San Francisco, CA

Network Engineer
$293k+/yrHybridDevOps / SRE

Designs, operates, and improves secure enterprise networks spanning offices, campuses, cloud environments, and connectivity services. The role combines architecture, production operations, troubleshooting, observability, security, and infrastructure automation.