Technical Lead owning architecture, execution, and evolution of an AI-driven telehealth platform built on GCP. Hands-on player-coach role integrating production LLMs, leading a small senior team, and driving scalable distributed systems in a healthcare setting.
Salary not listed
Remote8+ YOEML Engineering
About the role
Responsibilities
Own System Architecture: Design and evolve distributed systems on GCP (Cloud Run, Pub/Sub, BigQuery), making final decisions on service boundaries, data models, and technology trade-offs across Go, Python, and TypeScript.
AI & LLM Strategy: Lead hands-on integration of AI systems, including open-source LLMs, embeddings, and speech models. Define prompt optimization, evaluation strategies, and safety practices using tools like LangFuse.
Operational Excellence: Champion observability, reliability, and scale. Oversee event-driven ingestion and automated QA pipelines for call transcripts and healthcare data (FHIR / Medplum).
Delivery & Execution: Contribute code to critical paths while leading major technical initiatives. Identify gaps, unblock teams, and drive projects to completion.
Cross-Functional Leadership: Translate complex technical constraints into clear options and trade-offs for Product, Design, and Clinical Operations.
Requirements
8+ years of software engineering experience, with 2+ years leading projects or serving as a Tech Lead.
Staff-level scope or equivalent technical impact.
Strong LLM experience in production environments (hands-on required).
Expert-level proficiency in Go, Python, or TypeScript/Node, with the ability to review code across the stack.
Deep experience with distributed systems, containerized deployments (Docker, Kubernetes, Cloud Run), and event-driven architectures.
Nice-to-Haves
ML background, especially applied machine learning or AI infrastructure.
Experience hosting or integrating open-source models (e.g., vLLM, Ollama), with a strong understanding of RAG architectures.
Background in regulated environments (HIPAA, SOC 2) or healthcare data standards (FHIR, HL7).
Strong written and verbal communication skills, including writing RFCs and explaining technical strategy to non-technical stakeholders.
Architects and operates production machine-learning systems that classify web and API traffic, detect bots and scrapers, and support real-time mitigation at internet edge latency. The role requires 9+ years of applied ML experience in adversarial domains and strong expertise in evaluation, data pipelines, and large-scale systems.
212k – 265k/yrRemote9+ YOEML Engineering
Member of Technical Staff, Agentic Environments
CohereNew York, NY
Build scalable software and tools for frontier model training, research experimentation, and production machine learning systems. The role requires strong Python and distributed-training expertise, experience with ML frameworks and infrastructure, and the ability to optimize and debug large language model systems.
Leads the technical direction of large-scale ML infrastructure for embedding, recommendation, and personalization systems. The role requires 8+ years of ML engineering experience, expertise in deep learning and distributed training, and strong leadership across research, infrastructure, and production deployment.
Leads the technical direction and development of large-scale, GenAI-powered recommendation and feed-ranking systems. Requires 10+ years of industry experience in relevance-driven products, deep expertise in machine learning and recommendations, and strong organizational influence and mentoring skills.
266k – 372k/yrRemote10+ YOEML Engineering
Staff Applied Scientist
Garner HealthNew York, NY
Leads end-to-end development of production algorithmic systems for healthcare, spanning machine learning, optimization, and LLM applications. The player-coach role requires 6+ years of industry experience, strong problem-solving and metrics judgment, and technical leadership of a small team.