Skip to content
Sweep360Sweep360

Machine Learning Engineer

Build and lead Sweep’s production AI decision system, converting noisy cyber-physical signals into reliable operational decisions across edge and cloud environments. The role requires 5–10 years of end-to-end production systems experience, strong Python and systems design skills, and comfort working in high-stakes deployments.

About the job

Responsibilities

  • Shape how the production AI system behaves in real-world environments.
  • Design ingestion → reasoning → decision systems.
  • Drive inference reliability, predictable logic, and transparent reasoning.
  • Close the loop from deployments to system learning.
  • Build trusted AI systems for high-stakes operational use.
  • Partner with RF, hardware, and field teams.
  • Work across edge devices, cloud environments, and intermittent connectivity.
  • Travel approximately 10–15% for domestic and global deployments.

Requirements

  • 5–10 years owning production systems end-to-end.
  • Strong system design skills across APIs, pipelines, and data storage.
  • Experience building production AI systems trusted in real-world operations.
  • Strong Python skills, plus Go, TypeScript, or similar.
  • Ability to build systems spanning edge devices and cloud infrastructure.
  • Ability to debug production systems quickly and decisively.
  • Clear communication and independent operation.
  • U.S. Person status required; the role may involve export-controlled data.

Nice-to-Haves

  • Experience with streaming systems such as Kafka or Pub/Sub.
  • Production LLM or inference pipelines, including prompting, retrieval, and evaluation.
  • Designing systems for adversarial or security environments.
  • Building systems that run both on-device and in the cloud.
  • Experience working in an early-stage startup environment.

Compensation and Benefits

  • Base salary up to $240,000, depending on qualifications, experience, and impact.
  • Total compensation includes equity, premium insurance, 401(k), flexible PTO, and other individual benefits.
  • On-site work at the company headquarters in New York City.

Skills

Python, Go, TypeScript, APIs, Data Pipelines, Data Storage, Kafka, Pub/Sub, Llm Pipelines, Inference, Retrieval, Edge Computing, Cloud Infrastructure, Cyber-Physical Security, System Design

Baseten

Baseten

San Francisco, CA

AI Engineer
$175k+/yrHybrid3+ YOEML Engineering

Build and ship AI-powered agents, automations, dashboards, and integrations that improve GPU capacity operations across Compute and C3. The role requires 3+ years of AI automation or technical operations experience, production Vercel expertise, agent-tool fluency, and strong API integration skills.

Anthropic

Anthropic

San Francisco, CA
Performance Engineer, Inference Engine
$350k+/yrHybridML Engineering

Build and optimize a high-scale LLM inference engine spanning accelerator programming, host-device coordination, and distributed systems. The role requires strong systems programming, performance analysis, and an understanding of LLM inference across compute, memory, and interconnects.

OnePay

OnePay

United States

Forward Deployed Engineer (AI and Automation)
$150k+/yrRemote5+ YOEML Engineering

Build and operate production AI agents, automation workflows, and integrations that improve complex business processes. The role requires 5+ years of software engineering experience, modern LLM and agent-framework expertise, systems integration skills, and strong cross-functional collaboration.

Thinking Machines Lab

Thinking Machines Lab

San Francisco, CA

Research Software Engineer, Post Training
$350k+/yrHybridML Engineering

Build and operate the engineering systems that support post-training research, including reinforcement learning infrastructure, sandboxed execution, data pipelines, and agent scaffolding. The role requires strong Python and systems engineering skills, project ownership, and a relevant bachelor’s degree or equivalent experience.

OpenAI

OpenAI

San Francisco, CA

Software Engineer, AI for Chip Design
$266k+/yrHybridML Engineering

Build research infrastructure and tooling that enables AI models to design silicon, including reinforcement learning environments, EDA integrations, evaluations, and experiment workflows. The role requires strong software engineering fundamentals and comfort working across research, tooling, and chip-design systems.