Skip to content
CrusoeCrusoe

Staff Software Engineer, Developer Experience

Staff-level engineer building developer tools, infrastructure, and automation to accelerate Crusoe engineering productivity. Requires Go, Kubernetes, CI/CD, and strong DevOps/SRE experience.

About the job

What You’ll Be Working On

Engineering Acceleration: Partner with the broader organization to build paved paths that engineers love to use. Our paved paths are fast, reliable and broadly adopted across all of engineering.

Toil Elimination: Ruthlessly identify and automate away the friction and repetitive tasks that slow down the development process, keeping Crusoe engineers in a state of high productivity.

Library & Environment Creation: Develop the libraries, tools, and pre-production environments necessary for vetting service APIs and complex microservice interactions.

Ecosystem Integration: Unify internal tooling and vendor services to automate workflows, build operational efficiency, and optimize security across the development stack.

Lifecycle Innovation: Innovate across every stage of the development lifecycle, including source code management, build systems, code review, CI/CD pipelines, platform runtimes, and telemetry.

Culture of Quality: Lead efforts to establish a culture of continuous quality delivery that scales seamlessly as our engineering headcount and infrastructure grow.

System Optimization: Work diligently to build efficient systems and processes that serve as a force multiplier for the impact of every engineer around you.

What You’ll Bring to the Team

  • Previous experience building developer tools and/or infrastructure for engineering teams
  • Expert ability to evaluate technical tradeoffs and understand how infrastructure decisions impact the daily productivity of the end-user (the developer)
  • Fluent knowledge of industry-standard AI tooling, build tooling, containerization, and open-source development frameworks
  • A demonstrated passion for building empathetic developer and operator workflows that prioritize human productivity
  • Professional experience managing or developing within Kubernetes clusters and a deep understanding of container orchestration
  • Proven experience in DevOps, Site Reliability Engineering (SRE), Release Engineering, or a similar productivity-focused discipline
  • Deep understanding of automated testing infrastructure and how to integrate it into a seamless CI/CD pipeline
  • Expertise in modern programming languages (specifically Go) and advanced proficiency in Git-based workflows (GitLab/GitHub)
  • A Bachelor’s or Master’s degree in Computer Science, Engineering, Mathematics, or a related analytical field (or equivalent professional experience)

Bonus Points

  • Experience informing long-term company objectives through technical insight and developer-centric advocacy
  • Experience building AI agent platforms for engineering teams
  • Hands-on experience with Linux image construction, package building, and kernel-level optimizations
  • Active involvement in the open-source community or a track record of staying current with recent industry advancements in developer productivity
  • A background in solving complex, multi-layered technical problems and then successfully automating the resulting solutions

Benefits

  • Competitive compensation
  • Restricted Stock Units
  • Paid time off & paid holidays
  • Comprehensive health, dental & vision insurance
  • Employer contributions to HSA account
  • Paid parental leave
  • Paid life insurance, short-term and long-term disability
  • Professional development & tuition reimbursement
  • Mental health & wellness support
  • Commuter benefits (parking & transit)
  • Cell phone stipend
  • 401(k) Retirement plan with company match up to 4% of salary
  • Volunteer time off

Skills

Go, Kubernetes, Git, GitLab, GitHub, CI/CD, Docker, Linux, SRE, DevOps

Airbnb

Airbnb

United States

Staff Software Engineer, Service Tools
$212k+/yrRemote9+ YOEDevOps / SRE

Leads technical direction for Airbnb’s service developer tooling platform, spanning AI-assisted development, JVM build infrastructure, testing, modernization, and observability. Requires 9+ years of industry experience, strong backend and distributed-systems expertise, and the ability to influence organizations and deliver multi-quarter infrastructure initiatives.

Temporal

Temporal

United States

Staff Software Engineer, Traffic
$212k+/yrRemote8+ YOEDevOps / SRE

Leads the design and development of scalable, secure network traffic systems and cloud infrastructure. The role requires 8+ years of coding experience, strong distributed-systems and concurrency expertise, and deep knowledge of networking and performance optimization.

Crusoe

Crusoe

San Francisco, CA
Staff Software Engineer
$215k+/yrOn-site7+ YOEDevOps / SRE

Build diagnostics, automation, observability, and repair tooling for Crusoe’s large-scale GPU fleet and data centers. The role requires software engineering expertise in distributed systems, reliability, cloud platforms, and at least one of Go, Python, Java, or Rust.

Reddit

Reddit

San Francisco, CA

Staff Site Reliability Engineer - Site Experience
$217k+/yrOn-site8+ YOEDevOps / SRE

Leads reliability engineering for Reddit’s critical user-facing systems, improving availability, scalability, performance, automation, and incident response at internet scale. Requires 8+ years operating distributed systems and strong expertise in programming, observability, high availability, and production troubleshooting.

Reddit

Reddit

San Francisco, CA

Staff Site Reliability Engineer, Ads
$217k+/yrRemote8+ YOEDevOps / SRE

Provides technical leadership for reliability, scalability, and operational excellence across Reddit’s advertising systems. The role requires 8+ years operating large-scale distributed systems, strong software engineering skills, and expertise in cloud-native architectures, observability, and incident response.