Senior Platform Engineer
Build and operate scalable cloud infrastructure and application platforms, enabling frequent deployments, resilient systems, observability, and self-healing capabilities. The role requires strong troubleshooting, Linux and cloud experience, networking knowledge, and expertise in one or more platform engineering focus areas.
About the job
Responsibilities
- Build cloud infrastructure and the application platform powering healthcare products.
- Assess risk and impact for changes to live environments.
- Participate in an on-call rotation.
- Troubleshoot complex interactions across technology, governance, and people, including under time pressure.
- Design and build autonomous, self-healing processes.
- Support product and data engineers in areas of platform expertise while developing additional expertise.
- Communicate with stakeholders and internal platform customers.
Requirements
- Experience working in a startup environment.
- Experience building scalable infrastructure for web applications or data workloads.
- Deep troubleshooting skills, including identifying root causes across architectural and abstraction layers and implementing solutions.
- Fluency with the command line and comfort working with Linux.
- Strong communication skills.
- Experience with cloud infrastructure on GCP, AWS, or Azure; GCP is preferred.
- Familiarity with GitOps technologies such as Terraform.
- Understanding of VPCs, subnets, routes, peering, DNS, load balancers, L4/L7 routing, NAT gateways, CDNs, and TLS.
- Experience with observability tools such as Datadog and OpenTelemetry for infrastructure, application, and database performance monitoring.
- Ability to help instrument application code, build dashboards, and collaborate with application engineers.
Focus Areas
Experience in one or more of the following areas:
- Site reliability, including improving performance, reliability, and cost across the stack.
- Kubernetes, including scalable, resilient, and secure workload operations and troubleshooting.
- Polyglot development and the ability to read code and understand runtime operational characteristics.
- DevOps, including building and maintaining CI/CD systems.
- Developer product experience, including low-friction tools and workflows for product and data engineers.
- Data infrastructure, including databases, warehouses, connectors, replicators, access controls, and data pipelines.
- Distributed systems, including multi-region systems and scalability, reliability, and resilience tradeoffs.
- Security and cloud infrastructure experience in highly regulated industries such as healthcare.
Skills
GCP, AWS, Azure, Linux, Terraform, GitOps, Kubernetes, Datadog, OpenTelemetry, Networking, DNS, Tls, CI/CD, Distributed Systems, Cloud Infrastructure
Similar jobs
DevOps / SRE jobsDesigns and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.
Designs, deploys, and operates secure, resilient enterprise and cloud networks across data centers, on-premises environments, and AWS and Azure. Requires 6+ years of production network experience plus expertise in routing, switching, firewalls, automation, and hybrid connectivity.
Build and operate core platform infrastructure, developer tooling, CI/CD, observability, and cloud reliability systems for a regulated payments platform. Requires 5+ years of infrastructure or backend experience, strong infrastructure-as-code skills, and production cloud expertise.
Senior software engineer building standardized, self-service cloud infrastructure across AWS, Google Cloud, and networking systems. Requires 5+ years of software engineering experience, production cloud infrastructure expertise, and proficiency in Go or Python.
Designs and supports physical IT infrastructure across offices, labs, manufacturing facilities, and data centers, including racks, cabling, power, cooling, documentation, and capacity planning. Requires 5+ years of physical infrastructure engineering experience and strong cross-functional project execution.