Skip to content
HebbiaHebbia

Platform Engineer, Document Intelligence

Platform Engineer building high-scale distributed document indexing and search systems for an AI platform serving top financial institutions. Requires 5+ years experience in backend/distributed systems with Python/Java/Go and cloud infrastructure.

About the job

Responsibilities

  • Own critical system components: Take complex requirements and turn them into robust, scaled solutions that solve real customer needs.
  • Unlock O(1) universal indexing: Build and iterate on our high-scale document build system that enables constant time latency for indexing any content in the world, regardless of data volume.
  • Drive performance optimization: Architect and implement performance-tuning solutions to ensure our systems operate efficiently at scale, minimizing latency and maximizing throughput across millions of documents.
  • Mentor and guide: Provide technical leadership, mentorship, and guidance to junior engineers, fostering a culture of learning and growth.

Requirements

  • Bachelor's or Master's degree in Computer Science, Data Science, Statistics, or a related field with strong academic background in data structures, algorithms, and software development.
  • 5+ years software development experience at a venture-backed startup or top technology firm, with a focus on distributed systems and platform engineering.
  • Proficiency in building backend and distributed systems using Python, Java, or Go.
  • Deep understanding of scalable system design, performance optimization, and resilience engineering.
  • Extensive experience with cloud platforms (e.g., AWS).
  • Working experience with one or more of: Kafka, ElasticSearch, PostgreSQL, Redis.
  • Knowledge of workflow orchestration and execution platforms like Airflow, Temporal, or Prefect.
  • Proven experience enabling observability patterns.
  • Ability to analyze complex problems, propose innovative solutions, and communicate technical concepts effectively.
  • Proven experience leading software development projects and collaborating with cross-functional teams.
  • Strong interpersonal and communication skills; autonomous with strong ownership mindset.

Nice-to-Haves

  • Experience building distributed systems leveraging etcd or Apache Zookeeper.
  • Frequent user of AI products (e.g., Cursor, Claude Code) during the development lifecycle.

Benefits

  • Unlimited PTO
  • Medical + Dental + Vision + 401K
  • Catered lunch daily + DoorDash dinner credit
  • 3-4 months parental leave
  • $15k lifetime fertility benefits
  • Competitive equity package

Skills

Python, Java, Go, AWS, Kafka, Elasticsearch, Postgres, Redis, Airflow, Temporal, Prefect, Observability, Distributed Systems, Scalable System Design, Performance Optimization

Hebbia

Hebbia

New York, NY
Software Engineer, Infrastructure
$160k+/yrOn-site5+ YOEDevOps / SRE

Build and operate Hebbia’s AWS infrastructure and developer platform entirely through code. The role focuses on multi-account architecture, CI/CD, container orchestration, cloud cost controls, security compliance, and scalable platform foundations, requiring 5+ years of production cloud infrastructure experience.

Roboflow

Roboflow

New York, NY
Infrastructure Engineer
$165k+/yrRemoteDevOps / SRE

Infrastructure Engineer responsible for securing, scaling, and operating cloud infrastructure and machine-learning platforms across a distributed startup. The role requires Kubernetes, infrastructure-as-code, cloud operations, CI/CD, programming, observability, and security experience.

Baseten

Baseten

San Francisco, CA
Software Engineer - Continuous Delivery
$165k+/yrHybridDevOps / SRE

Build and operate continuous delivery infrastructure for Kubernetes deployments across global regions, including progressive rollouts, automated health evaluation, and rollback systems. The role requires strong Go or Python skills, large-scale Kubernetes experience, and familiarity with GitOps tooling.

Ramp

Ramp

New York, NY
TLM, Production Engineering
$168k+/yrHybrid3+ YOEDevOps / SRE

Production Engineer responsible for building and operating scalable infrastructure, driving reliability and architecture initiatives, and enabling product teams through platform tooling. Requires software engineering experience, distributed-systems expertise, cloud experience, and cross-team technical leadership.

Baseten

Baseten

San Francisco, CA

Capacity Ops Engineer
$170k+/yrHybrid5+ YOEDevOps / SRE

Leads global GPU capacity management across acquisition, orchestration, infrastructure automation, and incident response. The role requires 5+ years of experience, deep Kubernetes expertise, production Go or Python skills, and the ability to balance reliability with unit economics.