Senior Software Engineer - Infrastructure and Tools
Build and extend scalable infrastructure for Databricks' data and AI platform, including multi-cloud systems and Kubernetes at massive scale. Requires 5+ years experience in Java/Scala/Go/C++/Python, distributed systems, and cloud technologies.
About the job
The impact you will have:
- Build and extend components of the core Databricks infrastructure
- Architect multi-cloud systems and abstractions to allow the Databricks product to run on top of existing Cloud providers
- Improve software development workflows for engineering and operational efficiency
- Use our own data and AI platform (yes!) to analyze build and test logs and metrics to identify areas for improvement
- Develop automated build, test, and release infrastructures
- Set and uphold the standard for engineering processes to support high-quality engineering, including style and code checking, test harnesses, and release packaging
What we look for:
- BS (or higher) in Computer Science, or a related field
- 5+ years of experience writing production code in one of: Java, Scala, Go, C++ or Python
- Passion for building highly scalable and reliable infrastructure
- Experience architecting, developing and deploying large-scale distributed systems at scale
- Experience with cloud APIs (e.g., a public cloud such as AWS, Azure, GCP or an advanced private cloud such as Google, Facebook)
- Experience with cloud technologies, e.g. AWS, Azure, GCP, Docker, Kubernetes, or Terraform
Benefits
- Comprehensive health coverage including medical, dental, and vision
- 401(k) Plan
- Equity awards
- Flexible time off
- Paid parental leave
- Family Planning
- Gym reimbursement
- Annual personal development fund
- Work headphones reimbursement
- Employee Assistance Program (EAP)
- Business travel accident insurance
Skills
Java, Scala, Go, C++, Python, Kubernetes, AWS, Azure, GCP, Docker
Similar jobs
DevOps / SRE jobsOwn and evolve VSCO’s AWS/EKS platform, including infrastructure as code, GitOps, CI/CD, observability, networking, and production reliability. The role requires 5+ years of hands-on infrastructure or SRE experience and strong Kubernetes, Terraform, and AWS expertise.
Own foundational cloud infrastructure and the internal developer platform supporting Commure’s engineering teams. The role requires 6+ years of infrastructure, platform, or SRE experience and hands-on expertise across Kubernetes, infrastructure as code, GitOps, observability, and cloud environments.
Leads cloud infrastructure, platform strategy, deployment pipelines, and infrastructure automation for a growing consumer platform. Requires 5+ years in infrastructure, DevOps, platform engineering, or SRE, plus deep AWS, coding, containerization, and infrastructure-as-code experience.
Own reliability, deployments, observability, compliance, and AI infrastructure across AWS and Kubernetes for a fintech platform. The role requires strong DevOps/SRE depth, backend software engineering experience, and hands-on ownership of SOC 2 and PCI-DSS controls.
Own and evolve secure, highly available AWS and Azure infrastructure, including Terraform automation, Kubernetes, CI/CD, observability, networking, and incident response. The role requires 7+ years of DevOps or related experience and strong cross-functional partnership across engineering and security.