Skip to content
TwilioTwilio

Software Engineer L2

Software Engineer L2 responsible for evolving and maintaining Twilio's Compute infrastructure, including VM orchestration, AWS Auto Scaling Groups, hardened AMIs, secure container images, and automation of operational tasks in a remote-first environment.

About the job

Responsibilities

  • Collaborate with Tech Leaders, Architects and other Engineers to develop solutions for complex problems in distributed computing and infrastructure management.
  • Automate solutions for operational issues, such as monitoring, performance, planning, and disaster response.
  • Participate in an on-call rotation to support our business-critical infrastructure.
  • Ensure a high quality implementation by applying Infrastructure as Code industry standards.
  • Author and review design documents, runbooks, and other service documentation; maintain records of changes in the systems.
  • Apply Agile methodologies to continuously deliver value to the customers.
  • Act as point of contact for legacy/new Compute system components.

Requirements

  • 2+ years of experience in AWS Cloud infrastructure management (preferably backend/infrastructure-focused like AMI, EC2, IAM policies/roles, etc.).
  • Strong knowledge of ASG (Auto Scaling Groups) to design, implement, and support scalable cloud-native environments.
  • Experience with hardened base AMIs and AL23.
  • Proficiency with one or more programming languages such as Java or Python (including software architecture patterns, clean code, debugging, etc.).
  • Proficient in shell scripting to streamline repetitive tasks and enhance efficiency in operations.
  • Skills to work independently with multiple global teams, developing, configuring, deploying, and operating the global Twilio Infrastructure Platform, blending operational excellence with development best practices.
  • Knowledge of container-based application/services.

Nice-to-Haves

  • Knowledge of deployment tools and frameworks like infrastructure as code and continuous deployment processes (e.g. GitHub, Buildkite, Terraform-TFC, ArgoCD, Harness, Cloud network).
  • Operational experience in complex distributed systems, including experience with SLO/SLAs towards high availability and reliability goals, including tools like DataDog or Prometheus.
  • Exposure to File Integrity Monitoring (FIM) tools, specifically Falco, and awareness of compliance frameworks (PCI, SOX).
  • Knowledge in Kubernetes.
  • Experience with Claude AI or similar.

Compensation

The estimated pay ranges for this role are as follows:

  • Based in Colorado, Hawaii, Illinois, Maryland, Massachusetts, Minnesota, Vermont or Washington D.C.: $116,960 - $146,200
  • Based in New York, New Jersey, Washington State, or California (outside of the San Francisco Bay area): $123,760 - $154,700
  • Based in the San Francisco Bay area, California: $137,520 - $171,900

This role may be eligible to participate in Twilio’s equity plan and corporate bonus plan.

Skills

AWS, EC2, Ami, IAM, Asg, Java, Python, Shell Scripting, Terraform, Kubernetes, Docker, Prometheus, Datadog

Mercury

Mercury

San Francisco, CA
Software Engineer - Infrastructure
$116k+/yrRemote2+ YOEDevOps / SRE

Build Mercury’s secure, observable infrastructure platform across AWS, networking, containers, and developer tooling. The role requires strong Linux fundamentals, cloud-native experience, technical writing ability, and software development skills, with opportunities to support AI-agent infrastructure.

Fab2

Fab2

Austin, TX
Infrastructure Software Engineering Intern
$114k+/yrOn-siteDevOps / SRE

Infrastructure and site reliability intern building and operating on-premises backend infrastructure for a semiconductor fabrication environment. The role emphasizes systems programming, Linux, networking, reliability, observability, automation, and performance engineering.

Fab2

Fab2

Austin, TX
Infrastructure Software Engineering Intern
$108k+/yrOn-siteDevOps / SRE

Winter infrastructure and site reliability internship focused on building and operating minimal, on-premises backend infrastructure for a semiconductor fabrication facility. The role requires systems programming, Linux, networking, distributed systems, and hands-on infrastructure or automation experience.

Ontic

Ontic

Austin, TX

Associate DevOps Engineer
$100k+/yrHybridDevOps / SRE

Supports cloud infrastructure, automation, CI/CD, monitoring, and service reliability while learning alongside a global DevOps team. The entry-level role requires a bachelor’s degree, foundational systems knowledge, and exposure to cloud and DevOps tools.

PagerDuty

PagerDuty

Atlanta, GA

Site Reliability Engineer I
$98k+/yrHybridDevOps / SRE

Supports and evolves the networking, compute, Kubernetes, and ingress infrastructure powering PagerDuty’s real-time platform. Requires 0–1+ years of relevant experience, Linux production operations, cloud infrastructure knowledge, programming proficiency, and Infrastructure as Code experience.