Skip to content
TailscaleTailscale

Infrastructure Engineer

Infrastructure Engineer builds and maintains internal engineering services, improves observability, CI/CD pipelines, and cloud infrastructure using tools like Kubernetes and AWS. Requires experience with distributed systems, infrastructure as code, and operating managed services in a remote environment.

About the job

Key Responsibilities

  • Work as part of a team of engineers to design, build, test, and document core software components.
  • Exhibit ownership over the running services that comprise Tailscale’s product by building for observability, participating in incident response, and fielding customer support escalations.
  • Analyze and improve efficiency, scalability, and stability of various system resources.

Example Deliverables

  • Improve observability through metrics, alerting, logging and telemetry integration.
  • Identify and build improvements for Continuous Deployment.
  • Utilize infrastructure as code to make changes in a cloud environment.
  • Collaborate with other engineering teams to build cross functional infrastructure improvements.
  • Automate upgrades for managed services and VMs.

What We Are Looking For

  • Experience with CI/CD, secrets management, infrastructure as code, and observability.
  • Experience with distributed systems.
  • Experience with operating managed services in a cloud environment (preferably AWS).
  • Experience with operating Kubernetes in production is a strong plus.
  • Familiarity with networks (IP addressing, routing, etc.).
  • Most of the non-front-end portions of the system are developed in the Go programming language. Experience with Go is a plus.
  • Ability to give and process constructive feedback, as well as work independently.
  • Flexibility to adjust to the dynamic nature of a startup.
  • Excellent written and verbal communication skills.

Skills

Go, Kubernetes, AWS, CI/CD, Infrastructure As Code, Observability, Distributed Systems, Secrets Management, Continuous Deployment

Roboflow

Roboflow

New York, NY
Infrastructure Engineer
$165k+/yrRemoteDevOps / SRE

Infrastructure Engineer responsible for securing, scaling, and operating cloud infrastructure and machine-learning platforms across a distributed startup. The role requires Kubernetes, infrastructure-as-code, cloud operations, CI/CD, programming, observability, and security experience.

Baseten

Baseten

San Francisco, CA
Software Engineer - Continuous Delivery
$165k+/yrHybridDevOps / SRE

Build and operate continuous delivery infrastructure for Kubernetes deployments across global regions, including progressive rollouts, automated health evaluation, and rollback systems. The role requires strong Go or Python skills, large-scale Kubernetes experience, and familiarity with GitOps tooling.

Hebbia

Hebbia

New York, NY
Software Engineer, Infrastructure
$160k+/yrOn-site5+ YOEDevOps / SRE

Build and operate Hebbia’s AWS infrastructure and developer platform entirely through code. The role focuses on multi-account architecture, CI/CD, container orchestration, cloud cost controls, security compliance, and scalable platform foundations, requiring 5+ years of production cloud infrastructure experience.

Ramp

Ramp

New York, NY
TLM, Production Engineering
$168k+/yrHybrid3+ YOEDevOps / SRE

Production Engineer responsible for building and operating scalable infrastructure, driving reliability and architecture initiatives, and enabling product teams through platform tooling. Requires software engineering experience, distributed-systems expertise, cloud experience, and cross-team technical leadership.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.