Software Engineer, Data Platform
Build and operate the streaming, storage, query, and self-service infrastructure underlying the company’s data platform. The role suits an early-career engineer with 1–3 years of experience, a computer science bachelor’s degree, programming skills, and interest in distributed systems.
About the job
Responsibilities
- Build reusable infrastructure that ingests, moves, and serves data across the organization.
- Work with Kafka and streaming tools to support real-time data movement.
- Build monitoring, alerting, and observability capabilities.
- Develop internal tools and abstractions across ClickHouse, Tinybird, and Snowflake.
- Implement platform features and fixes with increasing independence.
- Help implement access controls, data protection, and compliance practices.
- Contribute to platform architecture and build-versus-buy decisions supporting analytics and AI/ML workloads.
Requirements
- 1–3 years of experience in data engineering, backend/infrastructure engineering, or a related field.
- Bachelor’s degree in Computer Science or a related field.
- Programming experience with Python, Java, Scala, Go, or a similar language.
- Exposure to or strong interest in distributed systems and streaming/event-driven architecture.
- Comfort with SQL and curiosity about storage and query systems.
- Infrastructure mindset focused on building foundational systems.
- Strong communication and collaboration skills.
Compensation and Benefits
- Base pay range: $130,000–$200,000 for San Francisco, CA.
- Competitive compensation package, including equity.
- Inclusive healthcare package.
- Mentorship and professional development opportunities.
- Flexible time off.
- Company-provided equipment and work-from-home budget.
Skills
Kafka, ClickHouse, Tinybird, Snowflake, Python, Java, Scala, Go, SQL, Distributed Systems, Streaming Architecture, Event-Driven Architecture, Monitoring, Observability, Access Controls
Similar jobs
Data Engineering jobsManages the full lifecycle of scientific research data, including governance, metadata, repositories, open-science publishing, and AI/ML compatibility. The role also builds partnerships with federal research organizations and requires a bachelor’s degree plus 5–7+ years of relevant experience.
Build and operate production data pipelines and transformation layers that turn heterogeneous business, identity, and fraud data into reliable inputs for entity resolution, scoring, and customer APIs. The role requires at least one year of data engineering experience with Python, SQL, cloud platforms, and modern pipeline tooling.
Build and maintain reliable data pipelines, warehouses, and lightweight data applications supporting analytics, operational models, and AI/ML workflows. The role requires SQL, Python, ETL, and data warehousing knowledge, with 1+ year of relevant experience preferred.
Builds customer intelligence workflows that turn product usage, engagement, and contract data into actionable insights for Customer Success. The role requires strong Python and SQL skills, a bachelor's degree, and an interest in applied AI, automation, and predictive customer health modeling.
Build and optimize scalable data pipelines, reusable datasets, and federated data quality systems for healthcare analytics. The role requires at least 2 years of data or software engineering experience and strong Python, SQL, AWS, orchestration, database, and warehouse expertise.