AI Research Engineer
Develops advanced AI agents for healthcare revenue recovery, focusing on human-like conversational AI, model improvements, LLM orchestration, and evaluation frameworks to handle insurance interactions and billing tasks.
About the job
Responsibilities
- Improve models, LLM orchestration, prompt optimizations, evaluation framework, and AI infra
- Enable agent/workflow product engineers
- Own and raise the bar on AI quality
- Own ultimate vision of end-to-end AI orchestration of modules
- Work on hardest research problems (e.g., Anti-AI)
- Mix of ML, data science, and applied LLM work (heavy Python)
Requirements
- Like reading latest LLM research papers
- Deep experience implementing and building AI agents
- Eager to try latest releases (e.g., o1, OpenAI voice mode, Claude Computer-Use) for real-world healthcare billing
- Worked with complex distributed systems with many async operations
- Excited about pushing boundaries of conversational AI
- Balance quick experimentation with systematic evaluation
- Very high Python mastery
Perks & Benefits
- In-person culture at Flatiron office in NYC with paid lunch and dinner
- Flexible hours and time off
- Gym stipend
- Commuter benefits
- Health, dental, vision insurance
- 401(k) with matching
- Annual offsite
Skills
Python, LLMs, AI Agents, Llm Orchestration, Prompt Engineering, Machine Learning, Evaluation Frameworks, Distributed Systems, Async Operations, Conversational AI
Similar jobs
AI Research jobsLeads the design, measurement, publication, and adoption of APEX benchmarks evaluating frontier models on economically valuable professional work. The role requires rigorous research judgment, strong coding and statistical skills, and excellent communication across technical, commercial, and research audiences.
Research Scientist defining and executing research on reliable long-horizon agents in enterprise environments. The role focuses on post-training and reinforcement learning, agent memory, evaluation, verification, and structured representations, combining hands-on experimentation with product delivery and publication.
Conduct rigorous people research and applied data science to evaluate talent programs, organizational health, and employee experiences. The role requires advanced expertise in research design, experimentation, measurement, causal inference, statistical modeling, and responsible handling of sensitive employee data.
Build agent-driven chatbots and generative AI workflows for financial-wellness products, owning features from design through impact measurement. The role requires at least three years of software engineering experience, strong system design, maintainable coding practices, and a bachelor’s degree or equivalent experience.
Research Scientist focused on evaluating frontier language and multimodal models, diagnosing failure modes, and building rigorous benchmarks. The role requires advanced training in AI or a related field, post-training expertise, and published machine learning research.