Engineer - Member of Technical Staff
Build, optimize, and evaluate long-running and multi-agent AI systems, along with tools for monitoring and analyzing their real-world behavior. The role requires software engineering experience with coding agents, strong independence, and familiarity with current AI developments.
About the job
Responsibilities
- Build, optimize, and evaluate scaffolding for long-running AI agents.
- Structure multi-agent setups involving tens, hundreds, or thousands of interacting agents.
- Develop tools to monitor agents interacting with the real world at scale.
- Develop tools to analyze large volumes of data and extract insights.
Requirements
- Software engineering experience eliciting agent capabilities or developing multi-agent setups.
- Action-oriented and willing to work across unfamiliar areas with a bias for action.
- Able to work independently with limited code review and oversight.
- Extensive experience with coding agents, including understanding their strengths and weaknesses.
- Keeps up with current AI news and trends.
Nice-to-haves
- Interest in writing technical blog posts or AI explainers.
- Experience designing and building strong user experiences.
- Deep concern for beneficial AI outcomes.
- An existing audience on Twitter, Substack, or a similar platform.
Compensation
- Approximately $150,000–$350,000 per year.
Skills
Python, Multi-Agent Systems, AI Agents, Agent Scaffolding, Agent Evaluation, UX Design, Data Analysis, Technical Writing
Similar jobs
AI Research jobsConduct research on long-horizon, multi-agent AI behavior by designing agent environments, analyzing large-scale data, and running experiments. The role requires strong research judgment, rapid execution, independence, and familiarity with current AI developments.
Research Scientist developing and evaluating health-focused AI models, large language models, and agentic systems for clinical applications. The role requires advanced research experience, strong coding skills, healthcare or clinical-data experience, and top-tier AI/ML publications.
Research Engineer focused on designing benchmarks, evaluation systems, rubrics, and failure-analysis workflows for frontier language models. The role requires strong applied AI research and coding experience, with expertise in model evaluation, data quality, and backend systems.
Own and optimize an enterprise customer-support AI agent by improving prompts, conversational quality, guardrails, intent recognition, and performance. The role requires deep conversational AI experience, strong analytical skills, platform expertise, and strategic collaboration across support teams.
Develops experimental AI techniques and prototypes for agentic marketing applications, with emphasis on image and video generation. The role requires strong backend or probabilistic systems expertise, quantitative thinking, creativity with LLM applications, and product intuition.