Build and maintain scalable backend infrastructure for Fireworks AI's generative AI platform, including LLM CI/CD pipelines, control planes, and model serving systems. Requires 5+ years software engineering experience focused on ML/infrastructure, strong Python/Go skills, and familiarity with PyTorch, Kubernetes, and LLM concepts.
175k – 220k/yr
On-site5+ YOEML Engineering
About the role
Key Responsibilities
Contribute to the design and development of scalable backend infrastructure that supports distributed training, inference, and data pipelines.
Build and maintain core backend services such as LLM CI/CD pipeline, control plane, and model serving systems.
Support performance optimization, cost efficiency, and reliability improvements across compute, storage, and networking layers.
Building frameworks and safeguards to ensure Fireworks AI has the best model quality in the industry.
Collaborate with performance, training, and product teams to translate research and product needs into infrastructure solutions.
Participate in code reviews, technical discussions, and continuous integration and deployment processes.
Minimum Qualifications
Bachelor’s degree in Computer Science, Engineering, or a related technical field (or equivalent practical experience).
3 years of experience in software engineering, with a focus on infrastructure or machine learning systems.
Strong programming skills in Python, Go, or a similar language.
Proven experience in ML infrastructure and tooling (e.g., PyTorch, MLflow, Vertex AI, SageMaker, Kubernetes, etc.).
Build multimodal agentic systems that use large language models for creative video analysis and editing. The role spans model training, structured generation, evaluation, failure analysis, and deployment of production ML pipelines.
175k – 275k/yrOn-siteML Engineering
Software Engineer
xAIPalo Alto, CA
Build and own a mission-critical research platform for evaluating AI model capabilities and behaviors at xAI. Design instruments, datasets, grading schemes, and infrastructure to measure, diagnose, and improve models while shipping delightful internal tools.
175k – 275k/yrOn-site3+ YOEML Engineering
Software Engineer, ML Products
MirageNew York, NY
Build and ship end-to-end agentic systems and architectures for creative video workflows at an AI-native video platform. Requires 5+ years building production ML/agentic pipelines, deep RAG/context engineering experience, and strong evaluation skills.
175k – 275k/yrOn-site5+ YOEML Engineering
Robotic Software Engineer, Perception
Applied IntuitionSunnyvale, CA +1
Develop and integrate real-time AI/ML perception algorithms and sensor fusion software for autonomous vehicles across land, air, sea, and space domains. Requires MS/PhD or 5+ years experience with multi-modal sensors, ML deployment, and Linux/Docker; US citizenship and security clearance eligibility mandatory.
175k – 250k/yrOn-site5+ YOEML Engineering
Research Engineer
HedraSan Francisco, CA
Leads pre-training and post-training of action-conditioned world models and VLA models for physical AI applications. Requires PyTorch expertise, distributed training, and ML fundamentals; robotics background preferred.