Infrastructure Engineer
Infrastructure Engineer scales ML inference systems serving 150+ biological models using Kubernetes and AWS. Requires containerization expertise, cloud knowledge, and onsite presence in San Francisco.
Responsibilities
- Architect and maintain infrastructure serving 150+ biological ML models
- Scale platform several orders of magnitude to meet growing demand
- Orchestrate containerized workloads using Kubernetes
- Optimize resource allocation and ensure high availability
- Work closely with founders on customer needs, unpredictable workloads, and Bio-ML models
Requirements
- Solid programming and automation skills
- Experience with containerization and orchestration concepts
- Cloud platform knowledge (AWS/GCP/Azure)
- Located in the SF Bay Area or able to relocate
Preferred
- Experience scaling production systems
- Kubernetes experience
- Infrastructure as code tools (Terraform, Pulumi)
- Monitoring and observability tools
- Experience with GPU workloads
Tech Stack: Python, React, AWS (EC2, S3, DynamoDB), Docker, CUDA, Conda, TensorFlow/PyTorch, notebooks, bash/Slurm, APIs & web apps
Senior Network Engineer
Design, deploy, and operate enterprise network infrastructure for corporate facilities and hybrid cloud environments with zero-trust architecture and compliance requirements. Requires 5+ years enterprise networking experience and ability to obtain TS/SCI clearance.
Site Reliability Engineer
Senior or Staff Site Reliability Engineer focused on continuous delivery infrastructure using Argo Workflows, ArgoCD, and Kubernetes. Owns deployment tooling, onboarding flows, and participates in 24/7 on-call. Requires 6+ years building and operating distributed systems.