Senior Software Engineer - Java Streaming - Connectors
Build and own high-performance JVM-based connectors, drivers, and integrations connecting ClickHouse with streaming frameworks and the broader data ecosystem. The role requires 6+ years of software development experience, deep Java expertise, and production experience with streaming connectors and distributed messaging systems.
About the job
Responsibilities
- Build and maintain ClickHouse data connectors and integrations across the data ecosystem.
- Own the full lifecycle of JVM-based data framework integrations, including database drivers, SDKs, sources, and sinks.
- Develop and maintain production-grade connectors for streaming and data integration frameworks.
- Improve integration performance, reliability, throughput, and developer experience.
- Collaborate with open-source contributors, internal engineering teams, and enterprise users.
- Contribute to high-performance systems supporting real-time analytics and observability workloads.
Requirements
- 6+ years of software development experience building and delivering high-quality, data-intensive solutions.
- Experience with streaming or data integration frameworks, especially Apache Kafka, Kafka Connect, Apache Flink, or Apache Beam.
- Experience developing or extending connectors, sinks, or sources for a streaming processing framework.
- Hands-on experience with distributed messaging systems, including topic design, consumer groups, and performance tuning.
- Strong understanding of SQL, data modeling, query optimization, and OLAP databases.
- Experience building scalable data integration systems beyond simple ETL jobs.
- Strong proficiency in Java and the JVM ecosystem, including memory management, garbage collection tuning, and performance profiling.
- Experience with concurrent programming in Java, including threads, executors, and reactive or asynchronous patterns.
- Understanding of JDBC, TCP/IP, HTTP, and data-throughput optimization.
- Strong written and verbal communication skills.
- Passion for open-source development.
Nice-to-haves
- Contributions to open-source projects and active engagement with open-source communities.
- Familiarity with ClickHouse or similar high-performance data platforms.
- Working knowledge of Python and data engineering tools such as Pandas, PySpark, or Airflow.
Compensation and Benefits
- Flexible, remote-friendly work environment.
- Employer healthcare contributions.
- Company stock options.
- Flexible time off in the United States and generous entitlement in other countries.
- USD $500 home-office setup allowance for remote employees.
- Company-wide offsites and global gatherings.
Skills
Java, Jvm, Apache Kafka, Kafka Connect, Apache Flink, Apache Beam, Apache Pulsar, Amazon Kinesis, SQL, Data Modeling, Olap Databases, Jdbc, Python, Spark, Apache Airflow
Similar jobs
Data Engineering jobsStaff Software Engineer building scalable frameworks for high-performance financial data ingestion/distribution and AI-native products. Requires 7+ years experience with distributed systems, microservices, and data architectures; partners with product teams to drive technical direction.
Build and operate scalable lakehouse infrastructure, streaming and CDC pipelines, query systems, and self-serve BI capabilities. Requires 5+ years of data engineering experience, strong Kubernetes and infrastructure-as-code expertise, and hands-on experience with distributed data platforms.
Leads database architecture, performance, reliability, and developer-tooling initiatives for high-volume trading applications. Requires 8+ years of software engineering experience, expert MySQL skills, backend development expertise, and strong knowledge of distributed systems and database operations.
Senior Data Engineer responsible for building and operating reliable clinical and claims data pipelines, CDC systems, quality controls, and de-identified exports. The role requires 5+ years of production pipeline experience plus strong SQL, Python, Spark, and data-governance skills.
Own the company’s metric governance program by defining canonical metrics, enforcing them in semantic and catalog systems, improving data quality, and validating AI-agent outputs. Requires 5+ years in analytics or analytics engineering, strong SQL, production semantic-layer ownership, and experience with AI evaluation and data governance.