Infrastructure Software Engineer, Enterprise GenAI
Build and scale enterprise GenAI infrastructure across multi-cloud providers (AWS, Azure, GCP), implementing integrations and architecting systems for regulated industries. Requires 4+ years experience, proficiency in Python/JS/SQL, Kubernetes, and AI technologies like LLMs.
About the job
What You’ll Do
- Architect multi-cloud systems and abstractions to allow the SGP platform to run on top of existing Cloud providers
- Implement custom integrations between Scale AI's platform and customer data environments (cloud platforms, data warehouses, internal APIs)
- Collaborate with platform, product teams and our customers directly to develop and implement innovative infrastructure that scales to meet evolving needs
- Deliver experiments at a high velocity and level of quality to engage our customers
- Work across the entire product lifecycle from conceptualization through production
- Be able, and willing, to multi-task and learn new technologies quickly
What We’re Looking For
- 4+ years of full-time engineering experience, post-graduation
- Experience scaling products at hyper growth startups
- Experience tinkering with or productizing LLMs, vector databases, and the other latest AI technologies
- Proficient in Python or Javascript/Typescript, and SQL
- Experience with Kubernetes
- Experience with major cloud providers (AWS, Azure, GCP)
- Excellent communication skills with the ability to explain technical concepts to both technical and non-technical audiences
Skills
Python, JavaScript, TypeScript, SQL, Kubernetes, AWS, Azure, GCP, LLMs, Vector Databases
Similar jobs
DevOps / SRE jobsBuild and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.
Build and own production-grade AI agent infrastructure across multiple clouds, with responsibility for Kubernetes, Terraform, observability, security, reliability, and automation. Requires 5+ years of cloud infrastructure experience and strong CI/CD, networking, and production operations expertise.
Electrical Field Engineer supports on-site installation, testing, and commissioning of data center power systems like switchgear, transformers, UPS, and generators. Requires 5+ years experience, Bachelor's in Electrical Engineering, and 50%+ travel to sites.
Build and operate the Kubernetes-based cloud and on-premises infrastructure powering large-scale crawling, search, and ML workloads. The role requires 5+ years in DevOps, platform engineering, or cloud infrastructure, with strong Kubernetes, cloud, Docker, Terraform, and distributed-systems experience.
Owns reliability standards, incident management, observability, failure testing, and automation for a high-throughput AI infrastructure platform. The role requires deep Linux, networking, software, cloud-native, and distributed-systems experience, along with the ability to influence teams across the organization.