Senior Infrastructure Engineer
Own and scale Tennr’s AWS infrastructure, Kubernetes environments, and infrastructure-as-code foundation across development, staging, and production. The role requires 5–8 years of infrastructure, platform, or DevOps experience, strong Kubernetes and AWS expertise, and hands-on production ownership.
About the job
Responsibilities
- Build the EKS cluster module and reusable infrastructure patterns.
- Own cluster management and end-to-end environments for development, staging, and production.
- Drive and unblock infrastructure migrations using safe, staged approaches.
- Consolidate infrastructure observability onto Datadog and establish operational standards.
- Build and maintain AWS infrastructure, including networking, IAM, and core services, using infrastructure as code.
- Keep infrastructure secure by default, reliable, and cost-aware.
Requirements
- 5–8 years of experience in infrastructure, platform, or DevOps engineering with ownership of production systems.
- Hands-on experience with Kubernetes, preferably Amazon EKS.
- Experience with Terraform or comparable infrastructure-as-code tools.
- Strong AWS fundamentals, including networking, IAM, and core cloud services.
- Experience with Datadog or a comparable observability stack.
- Ability to execute autonomously and make pragmatic infrastructure decisions.
- Bias toward simple, durable systems and shipping over over-engineering.
Nice-to-haves
- Experience working in regulated or healthcare environments.
- Familiarity with HIPAA and protected health information (PHI).
Compensation and Benefits
- Annual salary: $200,000–$230,000.
- Unlimited PTO.
- 100% paid employee health benefit options.
- Employer-funded 401(k) match.
- Competitive parental leave.
- Free lunch and snacks.
Skills
Kubernetes, Amazon Eks, Terraform, AWS, Aws Networking, Aws Iam, Datadog, Infrastructure As Code, Observability, Cloud Infrastructure
Similar jobs
DevOps / SRE jobsBuild and improve cloud infrastructure, developer workflows, and internal tooling that make software development, testing, and releases more efficient and reliable. The role requires cloud architecture knowledge, CI/CD experience, Terraform and Bazel proficiency, and software development skills in Go, Python, or C++.
Build and operate scalable control-plane and data-plane infrastructure for distributed AI workloads, including Ray cluster orchestration, scheduling, observability, and accelerator integration. Requires a bachelor's degree or equivalent experience, 3+ years of production coding, cloud-native expertise, Kubernetes, and Go/Python proficiency.
Leads infrastructure and platform strategy for a production healthcare AI platform, owning AWS, reliability, disaster recovery, compliance, CI/CD, and secure AI-agent operations. Requires deep cloud and Terraform expertise, audit-cycle experience, and prior technical leadership.
The Production Engineer will build and operate secure, scalable, and reliable infrastructure and production services, while developing engineering frameworks and supporting on-call operations. The role requires 5+ years of production, site reliability, or DevOps experience and familiarity with AWS, Kubernetes, and Terraform.
Own the reliability, resilience, observability, and automation of AWS and Kubernetes infrastructure supporting production products and AI/ML workloads. The role requires 4+ years of cloud infrastructure experience, strong Kubernetes and Terraform expertise, and senior-level incident response and software engineering skills.