Skip to content
RampRamp

Software Engineer, Production Engineering

Builds and operates scalable infrastructure for compute, storage, messaging, and observability to support high-volume financial transactions. Partners with product teams on architecture, reliability, and AI enablement with 2+ years software engineering experience in distributed systems and cloud (AWS preferred).

About the job

What You'll Do

  • Build and operate critical infrastructure across Ramp's compute, storage, messaging, and observability stack — owning the systems that handle real financial transactions at scale.
  • Drive architectural change — not just flag problems. When you surface a reliability or scalability issue, you own the path forward: you propose the solution, find the owners across engineering, and stay in until it's resolved.
  • Partner with product teams at the design phase — reviewing architectures, embedding golden paths, and making it easy to build correctly the first time.
  • Build Ramp's next level of scale — you'll be a hands-on contributor to the most consequential infrastructure shift happening right now: our move to a cellular architecture, enabling Ramp to scale, reach international markets, operate in highly regulated and constrained environments (e.g. FedRAMP), and deliver on enterprise-grade SLAs.
  • Enable AI-native engineering — as Ramp builds increasingly AI-powered products, PE is the team that makes sure the platform can support them. You'll proactively partner with product teams on AI infrastructure patterns, define the golden paths that turn one-off solutions into reusable foundations, and stay ahead of emerging challenges before they become blockers.
  • Build developer tooling and self-service infrastructure — so that other teams can answer their own questions (cost, performance, reliability) without involving PE.
  • Participate in on-call rotation — and more importantly, use every incident as a signal to eliminate the root cause, not just resolve the symptom.
  • Lead across the company — PE doesn't just review designs or show up when called. We proactively initiate cross-team architectural reviews, and we take full ownership of company-wide reliability and scalability initiatives: identifying the problem, proposing the solution, aligning the stakeholders, and staying in until it's done.

What We Look For

Experience profile:

  • 2+ years of software engineering experience shipping high-quality architectures for critical systems
  • Strong software engineering fundamentals — you write clean, well-tested, production-ready code
  • Hands-on experience with distributed systems at production scale
  • Experience with at least one major cloud provider (AWS preferred)
  • Familiarity with observability practices (SLOs, error budgets, alerting, dashboards)
  • Track record of leading technical projects end-to-end, including cross-team coordination
  • Comfortable using AI tooling and coding agents as part of your everyday engineering workflow — we expect our engineers to leverage these tools to move faster and think bigger

Bonus (not required):

  • Experience with cellular or multi-tenant architecture patterns
  • Prior work on workflow orchestration systems (Temporal)
  • Contributions to developer experience or internal platform tooling
  • Experience in fintech, payments, or regulated industries (FedRAMP, SOC 2)

Skills

AWS, Distributed Systems, Observability, SLOs, Terraform, Temporal, Cellular Architecture, Container Orchestration, Kubernetes, CI/CD

Otter

Otter

Mountain View, CA

Production Engineer
$155k+/yrHybrid2+ YOEDevOps / SRE

Production Engineer builds and operates large-scale systems, focusing on automation, monitoring, infrastructure management, and resilient operations. Requires 2+ years in SRE/DevOps, expertise in Linux, AWS, Kubernetes, and programming in Python or Golang.

Airtable

Airtable

San Francisco, CA
Software Engineer, Infrastructure (2-8 YOE)
$148k+/yrHybrid2+ YOEDevOps / SRE

Backend engineers build and scale Airtable's infrastructure across teams like Base, Compute, Data, Storage, and Traffic. Requires 2-8 years experience in distributed systems, databases; CS degree; hybrid work in SF, NYC, Seattle, or LA areas.

Greptile

Greptile

San Francisco, CA

Infrastructure Engineer
$190k+/yrOn-site1+ YOEDevOps / SRE

Build and operate robust infrastructure, support enterprise deployments, and improve on-premises delivery for a rapidly scaling AI code review platform. The role requires networking expertise, cloud and container experience, and at least one year of infrastructure or software engineering experience.

Fireworks AI

Fireworks AI

San Mateo, CA
Member of Technical Staff, Systems Infrastructure
$200k+/yrOn-siteDevOps / SRE

Build and operate large-scale scheduling, storage, caching, and networking infrastructure for AI training and inference. The role targets PhD researchers graduating by December 2026 with systems research depth and strong programming and performance-measurement skills.

Mercury

Mercury

San Francisco, CA
Software Engineer - Infrastructure
$116k+/yrRemote2+ YOEDevOps / SRE

Build Mercury’s secure, observable infrastructure platform across AWS, networking, containers, and developer tooling. The role requires strong Linux fundamentals, cloud-native experience, technical writing ability, and software development skills, with opportunities to support AI-agent infrastructure.