DevOps Engineer
Design, implement, and maintain cloud infrastructure and CI/CD pipelines. Collaborate with developers, SRE, and Security to ensure system reliability, scalability, and security.
About the job
Core Responsibilities
- Design, implement, and maintain CI/CD pipelines to automate the build, test, and deployment processes.
- Effectively prioritize workload with a constant focus on increasing developer efficiency.
- Manage and optimize cloud infrastructure to ensure scalability, performance, and cost-effectiveness.
- Work closely with SRE to develop and maintain monitoring and alerting systems to ensure the health and performance of our systems and applications.
- Work closely with Security to implement and enforce security best practices across our infrastructure and applications.
- Collaborate with software developers to optimize application performance and reliability.
- Participate in the planning and execution of disaster recovery and business continuity strategies.
- Effectively communicate with stakeholders to provide updates on project statuses, issues, and recommendations.
- Participate in on-call rotation to provide system support.
Core Requirements
- At least 5 years of experience in a DevOps role.
- At least 2+ years of experience with Terraform.
- Expert knowledge of cloud concepts and services (EC2, ECS, S3, RDS, A/ELB, EBS, VPC, Route53 or equivalent).
- Expert level knowledge of Linux. Strong scripting skills for automation and tooling.
- Ability to troubleshoot, diagnose, resolve & document RCA of incidents.
- Good knowledge of database configuration, concepts, maintenance, and troubleshooting.
- Thorough understanding of Security principles and experience in designing, implementing, monitoring proper controls.
- Extensive experience with Infrastructure-as-code & configuration management tools (Terraform, SaltStack or equivalent).
- Experience with containerization technologies such as Docker and orchestration tools like Kubernetes, ECS including managing microservices architectures.
- Knowledge of CI/CD tools such as Jenkins, Github Actions or equivalent.
- Experience with monitoring and logging tools such as Datadog, Cloud Watch or equivalent.
- Excellent communication skill, with the ability to effectively collaborate with cross-functional teams.
Skills
Terraform, AWS, EC2, ECS, S3, Rds, Linux, Docker, Kubernetes, Jenkins, GitHub Actions, Datadog, CloudWatch, Saltstack, CI/CD
Similar jobs
DevOps / SRE jobsInfrastructure Engineer responsible for securing, scaling, and operating cloud infrastructure and machine-learning platforms across a distributed startup. The role requires Kubernetes, infrastructure-as-code, cloud operations, CI/CD, programming, observability, and security experience.
Build and operate continuous delivery infrastructure for Kubernetes deployments across global regions, including progressive rollouts, automated health evaluation, and rollback systems. The role requires strong Go or Python skills, large-scale Kubernetes experience, and familiarity with GitOps tooling.
Production Engineer responsible for building and operating scalable infrastructure, driving reliability and architecture initiatives, and enabling product teams through platform tooling. Requires software engineering experience, distributed-systems expertise, cloud experience, and cross-team technical leadership.
Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.
Build and operate Hebbia’s AWS infrastructure and developer platform entirely through code. The role focuses on multi-account architecture, CI/CD, container orchestration, cloud cost controls, security compliance, and scalable platform foundations, requiring 5+ years of production cloud infrastructure experience.