Staff Backend Engineer – AI
Own end-to-end development, evaluation, and production deployment of AI models serving high-volume real-time products. The role requires 5+ years of production Python experience, hands-on fine-tuning and ML operations, cloud infrastructure expertise, and strong technical ownership.
About the job
Responsibilities
- Own development, fine-tuning, and evaluation of in-house AI models from dataset design through production deployment.
- Run supervised fine-tuning and post-training experiments; establish benchmarks and evaluation harnesses.
- Build and maintain training and evaluation data pipelines, including data quality, labeling, and reproducibility.
- Deploy models on Stream’s serving stack and optimize latency, cost, and reliability at high volume.
- Set technical direction for ambiguous AI problems and decide what to build, test, or abandon.
- Integrate models with Go-based API teams and infrastructure across the engineering organization.
- Contribute to open-source projects and share work through code, writing, or community engagement.
- Raise engineering standards through code review, mentorship, and pragmatic best practices.
Requirements
- 5+ years of production-level Python engineering experience.
- Hands-on experience with supervised fine-tuning and post-training of models.
- Familiarity with fine-tuning and serving tools such as Unsloth, Fireworks, Baseten, or equivalent technologies.
- Experience with GCP or AWS, including infrastructure as code with Terraform.
- Experience operating ML-based products in production, including deployment, monitoring, retraining, and iteration.
- Experience designing and operating data pipelines for training and evaluation.
- Demonstrated ownership of ambiguous problems from definition through delivery.
- Strong communication skills and comfort working in a small, distributed, fast-moving team.
Nice to Have
- Visible open-source contributions or maintained libraries.
- Experience with Go.
- Deep understanding of Python concurrency and asyncio limitations in high-throughput systems.
- Experience with real-time or low-latency inference systems.
- Experience as an early engineer, founder, or in a startup environment with an undefined roadmap.
Compensation & Benefits
- 28 days paid time off plus paid Dutch holidays.
- Company equity.
- Pension scheme.
- Learning and Development budget.
- Commute expenses to Amsterdam covered or company bike option within the city.
- Fitness stipend.
- MacBook Pro.
- Healthy team lunches and snacks.
- Generous relocation package.
- Benefits are adjusted according to the employee’s location of residence.
Skills
Python, Machine Learning, Supervised Fine-Tuning, Terraform, GCP, AWS, Go, Unsloth, Fireworks, Baseten, Data Pipelines, Asyncio, Model Serving, Open Source
Similar jobs
ML Engineering jobsSets the technical direction for production machine learning across a payments platform, building and scaling models for risk, authorization, disputes, and forecasting. Requires 8+ years of ML engineering experience, including production model ownership and strong technical leadership.
Build and operate production AI agent systems that help Sales and Marketing teams with account planning, deal support, competitive intelligence, and content creation. The role requires 6+ years of experience shipping reliable LLM workflows with retrieval, tool use, permissions, evaluation, and observability.
The Staff Machine Learning Engineer will architect and deploy scalable generative AI and machine learning systems, including retrieval, inference, evaluation, and agentic workflows. The role requires 7+ years of software development experience, strong Python skills, applied ML expertise, and deep familiarity with modern GenAI platforms and frameworks.
Builds and owns production multi-agent AI infrastructure, backend integrations, and workflow automation for marketing operations. Requires 8+ years of software engineering experience, strong Python and JavaScript/Node.js skills, production LLM experience, and deep Google Cloud expertise.
Senior AI engineer owning production systems for real-time speech models and AI voice agents. The role requires 5+ years of production software experience, Python, model serving and inference optimization, cloud and distributed systems expertise, and strong reliability and operations skills.