Member of Technical Staff
Build AI agent harnesses, models, and product capabilities that enable agents to perform complex work across digital environments. The role combines applied AI research and software engineering, requiring Python proficiency, strong product judgment, and experience with agent tooling, reinforcement learning, or browser technologies.
About the job
Responsibilities
- Engineer agent harnesses that connect AI models with environments and tools to perform valuable work for users.
- Drive AI capabilities across multiple layers of an AI agents product.
- Develop and leverage AI models, infrastructure, and browser technologies to advance capability and scale.
- Maintain high standards for AI agent performance and user experience.
- Collaborate with engineers, designers, product managers, data scientists, and other partners.
- Contribute to reliability, code quality, AI evaluation, testing, and maintenance.
Requirements
- Strong foundational familiarity with the full AI product stack.
- Proficiency in Python.
- Significant experience with at least one of the following:
- Context engineering and tool interfaces for frontier AI models
- Post-training and reinforcement learning, particularly for multimodal models
- Browser technologies, including CDP, Playwright, or extension development
- Strong product intuition and commitment to excellent user experiences.
- Ability to work independently and take ownership in a small, fast-moving team.
- Passion for shipping high-quality products.
Nice to Have
- TypeScript, Go, or Rust experience.
Skills
Python, TypeScript, Go, Rust, Reinforcement Learning, Multimodal Models, Playwright, Chrome Devtools Protocol, Browser Extensions, AI Agents, AI Infrastructure, Ai Evaluation
Similar jobs
ML Engineering jobsBuild and operate production machine-learning systems for content safety, from messy customer data through classification, evaluation, and inference. The role requires 5+ years of ML engineering experience, strong Python and MLOps skills, and sound judgment across classical models and LLMs.
Builds the platform, verifiers, environments, and grading infrastructure used to evaluate enterprise AI agents at scale. The role combines strong software engineering with expertise in agent runtimes, evaluation design, benchmarks, and production failure analysis.
Optimizes distributed machine learning training and high-throughput offline inference across large accelerator clusters. The role focuses on profiling, scaling efficiency, cluster goodput, GPU performance, and cost-effective processing of autonomy data.
Build and operate production machine-learning systems for search ranking, relevance, extraction quality, and LLM-driven features. The role requires production ML ownership, ranking or relevance expertise, large-scale data experience, Python, and rigorous experimentation skills.
Build and deploy algorithmic systems for high-impact healthcare problems, choosing among machine learning, optimization, heuristics, and hybrid approaches. The role requires 4+ years of relevant industry experience, strong applied problem-solving and evaluation skills, and fluency in modern ML tooling.