Skip to content
AbridgeAbridge

Senior / Staff Software Engineer, Agentic Engineering

Build and own CI/CD systems, agentic AI tooling, and developer platforms that power engineering velocity at a fast-growing healthcare AI company. Requires strong experience with modern build systems, Kubernetes, and AI-assisted development workflows.

About the job

What You’ll Do

  • Build and evolve developer infrastructure: Own CI/CD systems, build tooling, and developer platforms that improve velocity and reliability across engineering. Drive measurable improvements in build/test times, deployment frequency, and onboarding experience.
  • Design and implement agentic workflows: Build MCP servers, autonomous agents, and AI-assisted tooling that embed AI deeply into the development lifecycle. Help teams integrate these tools effectively into their day-to-day work.
  • Own the developer experience: Improve the end-to-end developer journey — from local development and testing to deployment and observability. Identify and eliminate friction, treating internal engineers as your customers.
  • Drive technical excellence: Bring strong systems thinking to architecture and tooling decisions that improve developer ergonomics, feedback loops, and development velocity.
  • Evaluate and adopt AI tooling: Stay current on AI-powered development tools and guide adoption across the engineering org to deliver measurable productivity gains.

What You’ll Bring

  • Strong experience building and evolving CI/CD systems and build tooling (e.g., Bazel, Buck, GitHub Actions, Buildkite, CircleCI).
  • Hands-on experience with agentic frameworks, MCP servers, LLM tooling, or AI-assisted development workflows.
  • Experience improving developer experience across local development, CI/CD pipelines, internal tooling, and developer workflows.
  • Familiarity with production systems in public cloud environments (AWS, GCP, or similar).
  • Experience with containerized environments, Kubernetes, and modern infrastructure tooling (e.g., ArgoCD, Istio, Atmos).
  • Strong track record of leveraging AI tools and workflows to materially improve engineering velocity and operational efficiency.
  • Strong communication skills and ability to collaborate across teams and functions.

Bonus Points if...

  • You've built or contributed to an internal developer platform (IDP) at scale.
  • You have experience with multi-agent orchestration (LangGraph, LangChain, or similar).
  • You've worked in regulated industries or healthcare environments.

How we take care of Abridgers

  • Generous Time Off: 14 paid holidays, flexible PTO for salaried employees, and accrued time off for hourly employees.
  • Comprehensive Health Plans: Medical, Dental, and Vision coverage for all full-time employees and their families.
  • Generous HSA Contribution: If you choose a High Deductible Health Plan, Abridge makes monthly contributions to your HSA.
  • Paid Parental Leave: Generous paid parental leave for all full-time employees.
  • Family Forming Benefits: Resources and financial support to help you build your family.
  • 401(k) Matching: Contribution matching to help invest in your future.
  • Personal Device Allowance: Tax free funds for personal device usage.
  • Pre-tax Benefits: Access to Flexible Spending Accounts (FSA) and Commuter Benefits.
  • Lifestyle Wallet: Monthly contributions for fitness, professional development, coworking, and more.
  • Mental Health Support: Dedicated access to therapy and coaching to help you reach your goals.
  • Sabbatical Leave: Paid Sabbatical Leave after 5 years of employment.

Skills

CI/CD, Bazel, Buck, GitHub Actions, Buildkite, CircleCI, Kubernetes, AWS, GCP, Argo CD, Istio, Atmos, LangGraph, LangChain, Llm Tooling

Shield AI

Shield AI

San Mateo, CA
Sr. Staff Lead Site Reliability Engineer
$220k+/yrOn-site7+ YOEDevOps / SRE

Leads the establishment and maturation of SRE practices across cloud infrastructure and platform services, improving observability, resilience, incident response, and operational tooling. Requires 7+ years of experience, major-cloud infrastructure expertise, infrastructure as code, distributed systems, and strong technical leadership.

Skydio

Skydio

San Mateo, CA
Staff Site Reliability Engineer
$240k+/yrRemote8+ YOEDevOps / SRE

Owns and scales production cloud infrastructure across Kubernetes/EKS, AWS, Terraform, CI/CD, networking, and observability. The role requires 8+ years of infrastructure experience, strong Kubernetes operations expertise, and depth in reliability or scaling challenges.

Coinbase

Coinbase

United States

Staff Software Engineer, Developer Infrastructure
$218k+/yrRemote8+ YOEDevOps / SRE

Leads development of Coinbase’s CI, build, and deployment infrastructure used by engineers across the organization. The role requires 8+ years building production distributed systems, strong Go or systems-language expertise, and demonstrated technical leadership across complex platform initiatives.

Coinbase

Coinbase

United States

Staff Infrastructure Engineer, Trading
$218k+/yrRemote8+ YOEDevOps / SRE

Own the infrastructure, deployment, and operational tooling for Coinbase’s latency-sensitive institutional trading platform across cloud and colocated environments. The role requires 8+ years of infrastructure, platform, or SRE experience, strong Linux and networking fundamentals, and experience operating regulated, low-latency systems.

Reddit

Reddit

San Francisco, CA

Staff Site Reliability Engineer - Site Experience
$217k+/yrOn-site8+ YOEDevOps / SRE

Leads reliability engineering for Reddit’s critical user-facing systems, improving availability, scalability, performance, automation, and incident response at internet scale. Requires 8+ years operating distributed systems and strong expertise in programming, observability, high availability, and production troubleshooting.