Skip to content
TrabaTrabaNew York, NY

Senior Software Engineer

Build and own production AI agent systems (harnesses, evals, orchestration) on frontier LLMs for industrial supply chain workflows at Traba. Requires 5+ years software engineering with 1+ year shipping LLM/agent features, strong Python/TS, and high-agency in ambiguous customer environments.

200k – 240k/yr
Hybrid5+ YOEML Engineering

About the role

Responsibilities

  • Embed with customers and operators to understand how supply chains run today—then design and ship agents that take meaningful work off their plate.
  • Build production agent systems on frontier LLMs: tool use, sub-agents, retrieval, structured outputs, MCP servers, and the orchestration that ties them together.
  • Own evaluation as a first-class discipline—datasets from real traces, rubrics and graders, experiments, and improvements you can prove move the needle.
  • Architect the data, services, and APIs the agent layer depends on—integrating our internal systems with customers' WMS, TMS, and ERP environments.
  • Codify repeatable deployment patterns so each new customer rollout is faster than the last.

Requirements

  • 5+ years of software engineering, with 1+ year shipping LLM- or agent-based features into production.
  • Strong in Python and/or TypeScript/Node.js, and comfortable designing APIs, distributed systems, and data models in PostgreSQL.
  • Hands-on with the modern agent stack: production-scale prompt engineering, evaluation frameworks, orchestration patterns, and frontier model APIs.
  • A track record of building in fast, ambiguous environments—ideally at a vertical AI, AI-agent, forward-deployed, or data-product company.
  • Excellent written and verbal communication—you can run a customer workshop in the morning and write a clean design doc in the afternoon.

Nice-to-Haves

  • Builder with an AI operator's instinct. You've shipped real product on top of LLMs—not just chat wrappers. You've designed agent harnesses, structured tools, written evals, and tuned prompts against production traces. You think in capability, reliability, and unit economics—not just whether the model says the right thing.
  • Domain-immersed. You enjoy time with the people who do the work—learning an industry's vocabulary, edge cases, and operational tempo—and let that shape what you build.
  • High-agency in ambiguity. Dropped into a fuzzy customer problem with a half-formed hypothesis and a deadline, you scope, build, evaluate, and ship without waiting for a spec.
  • Sweat both ends of the stack. You move between prompt iteration, eval design, backend services, and customer-facing UI in the same week, and you care about evals that catch regressions and traces that are easy to debug.

Compensation

  • Competitive salary in the range of $200,000 - $240,000.
  • Start-up equity.
  • 100% Paid health, dental & vision coverage.
  • Dinner Provided via DoorDash, free DashPass & stocked kitchen for NY employees.
  • Commuter benefit.
  • Gympass Benefit.
  • Additional: One Medical Membership, Gympass, HSA via Optum, Talkspace, HealthAdvocate, Teledoc Health.

Skills

PythonTypeScriptNode.jsPostgresLLMsPrompt EngineeringEvaluation Frameworksagent orchestrationDistributed SystemsAPIswmstmsERP

Similar roles

ML Engineering jobs
Rad AI

Machine Learning Research Manager

Rad AISan Francisco, CA

Lead and mentor a team of applied and clinical researchers as a player-coach. Guide ML research in NLP, LLMs, and clinical applications for radiology, translating ideas into production systems while partnering with clinicians and engineers. Requires MS/PhD and 6+ years applied ML research experience.

200k – 230k/yr
On-site7+ YOEML Engineering
Reducto

Machine Learning Infrastructure Tech Lead

ReductoSan Francisco, CA

Lead ML infrastructure at Reducto by owning the training and inference stack. Hands-on role (80% building/optimizing) focused on GPU utilization, distributed systems, Kubernetes, kernels, and high-performance serving for AI document workflows. Requires 5+ years production ML infra experience and strong systems engineering skills.

200k – 300k/yr
On-site7+ YOEML Engineering
Airbnb

Senior Machine Learning Engineer, Relevance and Personalization

AirbnbUnited States

Build and productionize cutting-edge ML models for Airbnb's query intelligence, including autocomplete, query tagging, expansion, intent modeling, and LLM-powered natural language search to understand guest intent.

200k – 235k/yr
Remote5+ YOEML Engineering
Roger Healthcare

Senior Applied AI Engineer

Roger HealthcareSan Francisco, CA

Senior Applied AI Engineer building the core intelligence layer for Roger, an AI platform for home health clinicians. Responsibilities include training/fine-tuning LLMs on proprietary clinical data, building rigorous eval and monitoring systems, and shipping reliable agentic LLM workflows that improve patient care.

200k – 250k/yr
On-site7+ YOEML Engineering
Astronomer

Senior Software Engineer, Build

AstronomerNew York, NY

Build and scale Astronomer's AI-powered global context layer for data, focusing on semantic search, retrieval, code generation, and applied AI for data engineering workflows. Requires 5+ years software engineering experience with Python or Go, plus strong interest in LLMs and data tools.

200k – 230k/yr
Hybrid5+ YOEML Engineering