Skip to content
TwentyTwenty

DevOps Engineer

Design, build, and operate AWS infrastructure with Terraform, CI/CD pipelines, observability, and security for a national security technology platform.

About the job

What You’ll Do

  • Design, build, and operate AWS-based infrastructure.
  • Implement and maintain Infrastructure-as-Code for single-tenant and multi-tenant environments using Terraform.
  • Build and maintain deployment and environment automation (Ansible or similar).
  • Own and evolve CI/CD pipelines.
  • Design, implement, and refine observability: metrics, logs, traces, dashboards, and alerting.
  • Partner with application teams on architecture decisions, performance tuning, and operational readiness.
  • Contribute to security and governance: IAM policies, network security, secrets management, and security scanning.
  • Document systems, patterns, and runbooks so others can operate and extend the platform reliably.

Must Haves

  • Experience administering infrastructure and operating applications deployed on AWS.
  • Experience using Terraform to manage single-tenant and multi-tenant systems.
  • Strong instincts and practical experience with IP networking (VPCs, routing, subnets, proxies, DNS), network security (security groups, NACLs, firewalls), and PKI management (TLS certificates, CAs, mTLS, certificate lifecycle).

Should Have

  • Experience with Ansible or another deployment automation/configuration management framework.
  • Experience with GitHub Actions or another CI/CD platform (GitLab CI/CD, CircleCI, etc.).
  • Experience working on reliability projects, such as setting up alert management tools and on-call practices, and debugging failures in distributed systems.
  • Experience setting up observability and operations programs, including collecting and representing telemetry in Grafana (dashboards, panels, alerts) and instrumenting applications using OpenTelemetry or other log/metric/trace aggregation frameworks.
  • Experience managing PostgreSQL databases in production (backups, migrations, performance, monitoring).
  • Experience managing a pub/sub or queue technology, such as NATS, RabbitMQ, Kafka, AWS SQS, or Google Pub/Sub.
  • Familiarity with secrets management (AWS SSM/Secrets Manager, Vault, or similar).

Nice to Have

  • Experience or strong interest in cybersecurity (threat modeling, hardening, secure defaults).
  • Proficiency in Python or another scripting language for tooling and automation.
  • Experience operating a security scanning tool, such as Trivy (or similar vulnerability/container scanners).
  • Experience or interest working with large datasets (performance, storage trade-offs, retention policies).
  • Experience managing graph databases, such as Neo4j, AWS Neptune, or similar.
  • Experience designing or contributing to runbooks and internal platform documentation for non-infrastructure teams.

Benefits

  • Health: Medical, dental, and vision plan options. Life / AD&D, disability coverage options.
  • Family: Paid parental leave for eligible full-time employees (12 weeks for birthing parents, 4 for non-birthing parents, 6 weeks for adoptive, foster, or intended parents through surrogacy).
  • Vacation: Paid holidays and flexible PTO.
  • Retirement: 401(k) with pre-tax and Roth options. HSA/FSA options, dependent care FSA.
  • At the office: Commuter benefits. On-site garage parking. Bike storage. Building fitness center. Desk setup stipend.

Skills

AWS, Terraform, Ansible, CI/CD, GitHub Actions, Observability, Grafana, OpenTelemetry, Postgres, Networking, Security, Pki, Python

Granica

Granica

Remote

Software Engineer, Infrastructure
No salary listedRemote5+ YOEDevOps / SRE

Build and operate cloud infrastructure, Kubernetes platforms, CI/CD systems, and observability for exabyte-scale data systems and reliable enterprise AI workloads. The role requires 5+ years of infrastructure, platform, or distributed systems experience and strong programming and cloud skills.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.

Teleport

Teleport

United States

IT Security and Automation Engineer
$149k+/yrRemoteDevOps / SRE

Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.

Crusoe

Crusoe

United States

Electrical Field Engineer - Data Center
$196k+/yrRemote5+ YOEDevOps / SRE

Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.

Beacon AI

Beacon AI

San Carlos, CA

Software Engineer, Cloud Infrastructure
$135k+/yrHybridDevOps / SRE

Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.