Skip to content
SprigSprig

Senior Platform Engineer

Own and modernize the build, CI, test automation, and ephemeral environment platform for a large TypeScript, React, and Go monorepo. The role requires 6+ years of large-scale build-system experience, strong Bazel or comparable tooling expertise, and deep knowledge of hermetic, reproducible development workflows.

About the job

Responsibilities

  • Lead migration from Node Bazel rules to a maintained successor and complete the JavaScript toolchain migration to Bazel modules.
  • Replace image and packaging workarounds with proper module linking and incremental compilation.
  • Establish and improve p50 and p95 pull-request build and test-time metrics through caching, incremental builds, and test parallelism.
  • Own remote cache fleet sizing, eviction behavior, and cost.
  • Build telemetry for per-target timing, cache hit rate, flake rate, and cost.
  • Own developer test workflows across unit, integration, browser end-to-end, and visual testing.
  • Make local and CI commands consistent, support single-file and affected-graph execution, and detect and quarantine flaky tests.
  • Improve lint, static analysis, and coverage as reliable quality gates.
  • Replace shared environments with short-lived, isolated environments created per branch or task.
  • Deliver hermetic, sandboxed, reproducible build and test environments that support safe parallel agent execution and automated review and presubmit signals.

Requirements

  • 6+ years owning a monorepo build system at scale.
  • Deep knowledge of hermeticity, sandboxing, action graphs, and cache invalidation.
  • Experience owning incremental builds, test sharding, self-hosted runner fleets, presubmit/postsubmit design, and merge queues.
  • Deep experience with production frontend bundler pipelines and shared UI packages.
  • Strong Node.js and TypeScript expertise, including CJS/ESM module resolution and packaging, with comfort working alongside Go.
  • Experience eliminating hidden network calls, machine-dependent toolchains, and undeclared inputs to ensure reproducible builds and prevent cache or sandbox poisoning.
  • Proven ownership of flake detection and quarantine, balanced sharding, merge signals, and measurable improvements in build times, CI spend, or artifact and bundle sizes.

Preferred Qualifications

  • Strong Bazel experience preferred; Buck2, Pants, Nx, or Turborepo also considered.
  • Experience with webpack, Vite, esbuild, or Turbopack.
  • Experience with Jest, Playwright, and Cypress.

Compensation & Benefits

  • Base salary: $180,000–$260,000 annually.
  • Competitive employee equity.
  • 401(k) program.
  • Medical, dental, and vision benefits.
  • FSA/HSA benefits.
  • $175/month commuter benefit.
  • Additional wellbeing benefits.
  • Flexible paid time off.
  • Paid parental leave.
  • Professional development stipend.
  • Hybrid office policy.
  • Lunch and dinner provided daily.
  • Company-sponsored social events.

Skills

Bazel, TypeScript, JavaScript, Node.js, Go, CI/CD, Test Sharding, webpack, Vite, Esbuild, Turbopack, Jest, Playwright, Cypress, Docker

Runpod

Runpod

United States

Senior HPC Storage Engineer
$180k+/yrRemote8+ YOEDevOps / SRE

Own the design, scaling, reliability, and automation of a multi-region storage platform supporting AI workloads. The role requires 8+ years of production infrastructure or storage engineering experience, distributed storage expertise, strong Linux and networking knowledge, and production programming skills.

tastytrade

tastytrade

Chicago, IL

Senior Site Reliability Engineer - Linux Systems & Application Observability
$180k+/yrHybrid5+ YOEDevOps / SRE

Senior Site Reliability Engineer responsible for building fault-tolerant infrastructure, scaling a Nomad-based service fabric, and strengthening observability for critical brokerage systems. The role requires production experience with distributed systems, Linux, networking, instrumentation, on-call operations, and reliability practices.

Camber

Camber

New York, NY

Senior Platform Software Engineer
$180k+/yrOn-site6+ YOEDevOps / SRE

Senior platform engineer responsible for reliable, secure, and scalable infrastructure, developer tooling, observability, and AI enablement. The role requires 6+ years in platform engineering, SRE, or DevOps, with strong AWS and incident leadership experience.

Onebrief

Onebrief

Colorado Springs, CO

Senior Site Reliability Engineer, Colorado Springs
$180k+/yrOn-site5+ YOEDevOps / SRE

Own reliability, scalability, security, observability, and incident response for mission-critical applications across Kubernetes, AWS, and on-premise DoD environments. Requires an active Top Secret clearance and at least five years of infrastructure-focused SRE, DevOps, or platform engineering experience.

Lightning AI

Lightning AI

New York, NY

Senior Infrastructure Software Engineer
$180k+/yrHybrid8+ YOEDevOps / SRE

Build and operate production software, APIs, and automation for large-scale bare-metal and GPU infrastructure. The role requires 8+ years of software or infrastructure engineering experience, strong Python and Linux skills, and expertise in provisioning, lifecycle management, and reliability.