Build and maintain infrastructure, CI/CD pipelines, monitoring, and automation for Openly's insurance platform on Google Cloud. Requires 2+ years infrastructure automation experience, IaC (Terraform), cloud expertise, and strong scripting in Python or Go.
115k – 173k/yr
Remote2+ YOEDevOps / SRE
About the role
Key Responsibilities
Build internal tooling to help other engineers and the rest of the company understand and operate our system
Design and implement security best practices for our team and infrastructure
Reduce toil through automation, including building and maintaining CI/CD infrastructure
Build infrastructure as code using declarative provisioning tools
Develop high signal-to-noise ratio monitoring and alerting policies and technology to help us meet our SLOs
Lead incident response and postmortems
Contribute to important architectural and operational decisions like microservices vs. monoliths, deployment techniques, technologies, policies, etc.
Requirements
2+ years of professional/production experience developing and using infrastructure automation tools and techniques
Proven track record of creating improvements in business-critical systems around stability, performance, and scalability
Demonstrated ability to deliver complete systems from start to finish in a reasonable time frame
Understands the consequences of running software in production and are willing to share your knowledge with the rest of the team
Ability to explain complex technical challenges to non-technical audiences
Strong scripting skills in one or more of the following: Python, Go
Experience working with Infrastructure as Code (IaC) tooling, preferably Terraform
Software Engineer L2 responsible for evolving and maintaining Twilio's Compute infrastructure, including VM orchestration, AWS Auto Scaling Groups, hardened AMIs, secure container images, and automation of operational tasks in a remote-first environment.
117k – 172k/yrRemote2+ YOEDevOps / SRE
Software Engineer - Developer Infrastructure
Applied IntuitionSunnyvale, CA
Builds and improves core libraries, frameworks, and developer tools like Bazel and Buildkite CI/CD to boost engineering productivity. Requires 2+ years experience, Bachelor's in CS, and expertise in Go/C++/Python/TypeScript.
120k – 300k/yrOn-site2+ YOEDevOps / SRE
Site Reliability Engineer
The Voleon GroupNew York, NY +1
Site Reliability Engineer improves, manages, and monitors production-critical infrastructure and data pipelines in a finance AI/ML firm. Collaborates on fault-tolerance, deployments, automation, and on-call incident response using Python, Linux, and cloud tools. Requires 2+ years experience and quantitative degree.
120k – 160k/yrRemote2+ YOEDevOps / SRE
Capacity Ops Associate
BasetenSan Francisco, CA +1
Manages GPU fleet operations, including node maintenance, capacity fulfillment, and technical orchestration between SRE/infra teams and customers. Requires 2+ years experience, Kubernetes familiarity, and strong communication skills.
120k – 160k/yrHybrid2+ YOEDevOps / SRE
Platform Operations Engineer
EliseAINew York, NY
Leads cross-functional technical projects to optimize tech stack, build custom automation and analytics solutions for business operations, and integrate systems using AWS, APIs, and databases. Requires 2+ years experience with Python/SQL proficiency and onsite presence in New York.