Skip to content
InstabaseInstabaseSan Francisco, CA

AI Engineer

Staff AI Engineer building the Agent Harness runtime for Instabase's SuperApp: design secure sandboxed execution environments, state machines for agent orchestration, tool-calling frameworks, and guardrails connecting LLMs to production systems. Requires 8+ years distributed systems experience plus agentic AI expertise.

230k – 315k/yr
Hybrid8+ YOEML Engineering

About the role

Responsibilities

  • Architect the next-generation execution runtime (Agent Harness) that drives SuperApp’s planning, reasoning, memory retention, and tool-execution loops. Build reliable state machines capable of pausing, resuming, and versioning agent trajectories.
  • Design and scale highly secure, isolated, and ephemeral environments (using Docker, gVisor, WebAssembly, or microVMs) to execute agent-generated code safely without putting host infrastructure at risk.
  • Build and maintain high-throughput APIs and integration layers (such as the Model Context Protocol - MCP) that connect SuperApp to external services, databases, web search utilities, and complex file processors.
  • Engineer robust system-level guardrails to detect and mitigate prompt injection, defend against jailbreak attempts, and ensure strict compliance with system prompt instructions and data isolation boundaries.
  • Implement advanced routing, context window pruning, and context caching strategies to optimize LLM token usage, minimize round-trip latencies, and drastically reduce the operational cost of complex agent loops.
  • Write comprehensive high-level design documents, align cross-functional engineering teams on the technical roadmap, and mentor senior and mid-level software engineers across the organization.

Requirements

  • Minimum 8+ years of professional software engineering experience, with a proven track record of designing, scaling, and operating mission-critical distributed systems.
  • 2+ years of production experience building or modifying agent execution environments, LLM orchestrators, or tool-calling frameworks (e.g., custom runtimes, LangChain, AutoGen, or similar systems).
  • Expert proficiency in Go and Python. Pragmatic, deep understanding of concurrent programming, microservices, gRPC, and highly scalable API design.
  • Strong experience with containerization and virtualization technologies (Docker, Kubernetes) and a solid grasp of sandboxing untrusted code execution.
  • Deep familiarity with message brokers (e.g., Kafka, RabbitMQ), caching infrastructure (Redis), and relational/non-relational database design.
  • Bachelor’s or Master’s degree in Computer Science, engineering, or equivalent practical systems engineering experience.

Nice-to-Haves

  • Prior experience developing developer tools, SDKs, or extensible plug-in architectures.
  • Experience with frontend integration (React, TypeScript) to support human-in-the-loop debugging interfaces or agent visual builders.
  • Experience working in high-growth, fast-paced startup environments.

Compensation

  • Base salary range: $230000 to $315000 + bonus, equity, and benefits.

Skills

GoPythonDockerKubernetesLLMsLangChainautogengRPCKafkaRabbitMQRediswebassemblygvisor

Similar roles

ML Engineering jobs
Baselayer

Senior AI Engineer, Agentic Data Enrichment

BaselayerSan Francisco, CA

Build and own production LLM-driven agents that enrich business identities using web discovery, evidence extraction, classification, and risk signals. The role requires strong asynchronous Python, browser automation, multi-provider LLM experience, evaluation methodology, and production agent ownership.

230k – 340k/yrHybrid5+ YOEML Engineering
OpenAI

Applied AI Engineer, GTM Growth Engineering

OpenAISan Francisco, CA

Build and improve production AI agent systems for OpenAI's GTM workflows. Own the end-to-end improvement loop using feedback, evaluation, experimentation and backend services to drive measurable gains in customer engagement, pipeline and team productivity. Requires 4+ years building reliable LLM-powered production systems plus strong product judgment.

230k – 385k/yrOn-site7+ YOEML Engineering
Instabase

Technical Lead Manager

InstabaseSan Francisco, CA

Player-coach Technical Lead Manager for AI Systems & Agents at Instabase. Hands-on architect and builder of stateful multi-turn agent loops, secure code execution sandboxes, tool orchestration via MCP, and evaluation harnesses while leading and growing a small elite team of AI engineers.

230k – 270k/yrHybrid8+ YOEML Engineering
Otter

Senior Machine Learning Engineer

OtterMountain View, CA

Lead projects building and deploying large-scale ASR, NLP, and LLM systems for meeting intelligence. Requires 5+ years building production ML systems with PyTorch/JAX and experience with speech/language models.

230k – 265k/yrHybrid5+ YOEML Engineering
Cohere

Senior Research Engineer - Safety Tooling and Data

CohereNew York, NY

Senior Research Engineer building data synthesis, analysis, and management tooling for AI safety model training and evaluation. Requires strong software engineering, statistics, and ML framework expertise.

230k – 380k/yrRemote5+ YOEML Engineering