Senior Software Engineer - Java Streaming - Connectors
Build and maintain high-performance Java/JVM connectors, drivers, and SDKs that integrate ClickHouse with streaming and data-processing ecosystems. The role requires 6+ years of software development experience, strong Java concurrency and performance expertise, and production experience with streaming connectors.
About the job
Responsibilities
- Build and maintain ClickHouse data ecosystem integrations, including database drivers, SDKs, connectors, sinks, and sources for JVM-based applications.
- Own the full lifecycle of integrations with streaming and data frameworks such as Kafka, Kafka Connect, Flink, Beam, Spark, dbt, and Fivetran.
- Develop scalable, high-performance data integration systems for real-time analytics and observability workloads.
- Optimize connector performance, reliability, developer experience, and data throughput.
- Collaborate with the open-source community, internal engineering teams, and enterprise users.
Requirements
- 6+ years of software development experience building and delivering high-quality, data-intensive solutions.
- Experience with streaming or data integration frameworks, preferably Apache Kafka, Kafka Connect, Apache Flink, or Apache Beam.
- Experience developing or extending production-grade connectors, sinks, or sources for a streaming processing framework.
- Hands-on experience with Apache Kafka or comparable distributed messaging systems, including topic design, consumer groups, and performance tuning.
- Strong understanding of SQL, database fundamentals, data modeling, query optimization, and OLAP databases.
- Strong proficiency in Java and the JVM ecosystem, including memory management, garbage collection tuning, and performance profiling.
- Experience with concurrent Java programming, including threads, executors, and reactive or asynchronous patterns.
- Understanding of JDBC, TCP/IP, HTTP, and network-throughput optimization.
- Excellent written and verbal communication skills.
- Passion for open-source development.
Nice-to-haves
- Contributions to open-source projects and active engagement with the OSS community.
- Familiarity with ClickHouse or similar high-performance data platforms.
- Working knowledge of Python for data engineering, including Pandas, PySpark, or Airflow.
Compensation and Benefits
- Healthcare contributions.
- Company stock options.
- Flexible time off in the United States and generous leave in other countries.
- USD $500 home-office setup allowance for remote employees.
- Opportunities to participate in company-wide global gatherings.
Skills
Java, Jvm, Apache Kafka, Kafka Connect, Apache Flink, Apache Beam, Spark, SQL, Jdbc, Python, pandas, Pyspark, Apache Airflow, ClickHouse, TCP/IP
Similar jobs
Data Engineering jobsStaff Software Engineer building scalable frameworks for high-performance financial data ingestion/distribution and AI-native products. Requires 7+ years experience with distributed systems, microservices, and data architectures; partners with product teams to drive technical direction.
Build and operate scalable lakehouse infrastructure, streaming and CDC pipelines, query systems, and self-serve BI capabilities. Requires 5+ years of data engineering experience, strong Kubernetes and infrastructure-as-code expertise, and hands-on experience with distributed data platforms.
Leads database architecture, performance, reliability, and developer-tooling initiatives for high-volume trading applications. Requires 8+ years of software engineering experience, expert MySQL skills, backend development expertise, and strong knowledge of distributed systems and database operations.
Senior Data Engineer responsible for building and operating reliable clinical and claims data pipelines, CDC systems, quality controls, and de-identified exports. The role requires 5+ years of production pipeline experience plus strong SQL, Python, Spark, and data-governance skills.
Own the company’s metric governance program by defining canonical metrics, enforcing them in semantic and catalog systems, improving data quality, and validating AI-agent outputs. Requires 5+ years in analytics or analytics engineering, strong SQL, production semantic-layer ownership, and experience with AI evaluation and data governance.