Senior Platform Engineer
Build and operate a self-service internal development platform that helps engineering teams deploy and run production services reliably. The role requires strong backend programming, production Kubernetes operations, cloud infrastructure, observability, networking, and distributed-systems experience.
About the job
Responsibilities
- Investigate gaps and limitations in engineering development workflows and understand infrastructure and platform requirements.
- Design self-service platform services and developer tooling focused on reliability, usability, and appropriate abstraction from cloud infrastructure.
- Write and review automation, configuration management, and application code.
- Author and review functional specifications and scoping documents for large platform projects and services.
- Own and operate the internal development platform.
- Collaborate with distributed engineering teams across multiple time zones.
- Contribute to tooling and services, resolve incidents, and respond to user requests.
- Investigate, scope, execute, and document medium-to-large platform projects.
- Drive platform adoption and become a subject matter expert in a platform component.
Requirements
- Experience building and operating large-scale distributed systems in cloud providers; AWS strongly preferred.
- Strong backend programming experience; fluency in Go strongly preferred, with deep experience in another compiled or strongly typed backend language acceptable.
- Experience working with AI coding agents and building high-quality context for high-quality outputs.
- Experience designing and implementing medium-to-large software projects, driving design reviews, and mentoring less-senior engineers.
- Strong experience operating production Kubernetes clusters.
- Practical experience defining and operating against SLI/SLOs for owned services.
- Strong observability experience across metrics, logging, and traces.
- Strong Linux and TCP/IP networking skills.
- Familiarity with OAuth and OIDC authentication protocols.
- Familiarity with web services and/or Kubernetes controller development.
- Strong experience using CI/CD workflows to deploy production services.
- Pragmatic, detail-oriented, self-motivated, and collaborative.
Compensation and Benefits
- MongoDB describes a supportive and enriching culture with employee affinity groups, fertility assistance, and generous parental leave.
- Workplace accommodations are available throughout the application and interview process.
Skills
Go, AWS, Kubernetes, Crossplane, Terraform, Helm, Drone, Prometheus, Grafana, OpenTelemetry, Linux, TCP/IP, OAuth, OIDC, CI/CD
Similar jobs
DevOps / SRE jobsSenior Site Reliability Engineer responsible for operating and improving reliable, scalable cloud services through automation, observability, incident response, and platform engineering. Requires strong Kubernetes, cloud infrastructure, Terraform, Go or Python, distributed systems, and reliability engineering expertise.
Senior Release Engineer responsible for building reliable CI/CD pipelines and release automation for enterprise SaaS platforms such as Salesforce and Zuora. The role requires 7+ years of release engineering or DevOps experience, strong Python skills, and hands-on use of approved AI-assisted tools.
Senior site reliability engineer who will build and operate observability, anomaly detection, reconciliation, and reliability tooling for GitLab’s monetization systems. The role requires Ruby on Rails and observability experience, with knowledge of monitoring platforms, data pipelines, and business-critical billing systems.
The Senior Network Engineer will design, automate, and operate large-scale, high-performance network infrastructure for AI data centers and GPU clusters. The role requires 5+ years of data center networking experience, expertise in spine-leaf fabrics and routing protocols, and familiarity with HPC or GPU-dense environments.
The Senior DevOps Engineer will evolve multi-cloud infrastructure, production Kubernetes platforms, AI workloads, databases, observability, networking, and automation. The role requires 7+ years in infrastructure, DevOps, or SRE, strong Terraform and Kubernetes expertise, and proficiency in Python or Go.