DevOps Engineer, DevEx
Builds and evolves internal developer platforms using Kubernetes, Terraform, and GitOps to enhance reliability, scalability, and DevEx. Requires 5+ years in platform engineering, strong AWS and cloud-native expertise, with on-call responsibilities.
About the job
What you’ll work on
- Owning and evolving Mark43's internal developer platform, focusing on reliability, scalability, and developer productivity
- Building self-service platform capabilities (Infrastructure-as-Code modules, GitOps workflows, golden paths, and onboarding templates)
- Improving developer experience (DevEx) by reducing friction in local development, CI/CD pipelines, deployments, and debugging workflows
- Designing and optimizing pipeline-as-code systems to improve build performance, reliability, and feedback cycles
- Driving adoption of GitOps and Infrastructure as Code best practices across the organization
- Leading initiatives to improve operational excellence, including system reliability, resilience, and incident response
- Partnering closely with Product, Engineering, and Security teams to align platform capabilities with business needs
- Owning and delivering large, cross-team platform initiatives that shape how engineers build and operate software at MARK43
- Participating in a shared on-call rotation to support critical infrastructure
What we expect from you
- 5+ years of experience in Platform Engineering, DevOps, and/or Developer Experience
- Strong experience with AWS and cloud-native architectures
- Deep experience with Kubernetes and containerized systems
- Strong experience with Infrastructure as Code (Terraform) and GitOps (e.g., ArgoCD)
- Experience building and optimizing CI/CD systems and pipeline-as-code workflows
- Demonstrated ability to improve developer experience and platform usability at scale.
- Strong systems thinking, including tradeoffs between reliability, cost, and developer velocity
- Understanding of security and compliance best practices in cloud environments
What sets you apart
- You take ownership of complex platform problems, from design through adoption
- You build platforms and abstractions that are adopted across engineering teams
- You proactively identify gaps in developer workflows and drive improvements
- You influence technical direction beyond your immediate team
- You balance developer velocity, reliability, and operational excellence
- You treat internal platforms as products, with a focus on usability and impact
Preferred Qualifications
- Familiarity with defining SLOs, alerting strategies, instrumentation best practices
- Experience with developer platforms and modern build systems
- Experience improving local development environments (e.g., devcontainers, Docker workflows, Tilt/Skaffold)
- Exposure to FinOps practices and cost optimization
- Experience integrating AI/LLM capabilities into developer workflows, such as: AI-assisted debugging or incident triage, Automated incident summaries or developer assistants
- Knowledge of compliance and data governance requirements in regulated environments
Compensation
Total compensation for this role is market competitive, including a base salary range of $140,000–$170,000, plus bonus, equity, and a comprehensive benefits package.
Skills
Kubernetes, Terraform, AWS, Argo CD, GitOps, CI/CD, Datadog, Helm, Docker, SLOs
Similar jobs
DevOps / SRE jobsBuild and operate highly available infrastructure for an enterprise AI platform, spanning cloud systems, Kubernetes, automation, observability, and reliability engineering. Requires 5+ years of production infrastructure experience, strong Python or Go skills, and daily use of AI-assisted workflows.
Build and operate deployment platforms, automation, and developer tooling that make software releases safer, more reliable, and self-service. The role requires a bachelor’s degree or equivalent, three years of software engineering experience, and experience with production systems and cloud or distributed infrastructure.
The DevOps Engineer will build and operate reliable infrastructure, deployment workflows, and observability for data pipelines and AI/ML systems. The role requires at least three years of DevOps, SRE, or infrastructure experience plus strong cloud, Terraform, containerization, and MLOps expertise.
Build and operate AWS cloud, ML, LLM, RAG, and IoT infrastructure, including deployment platforms, data pipelines, vector search, observability, security, and cost controls. The role requires deep AWS experience and production experience with LLM-powered applications.
Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.