Senior Software Engineer - Java Streaming - Connectors
Senior software engineer building and maintaining high-performance Java/JVM connectors, drivers, SDKs, and integrations for ClickHouse across streaming and data-processing ecosystems. Requires 6+ years of software development experience and deep expertise in Java concurrency, distributed messaging, databases, and performance optimization.
About the job
Responsibilities
- Build and maintain ClickHouse data ecosystem connectors and integrations, including database drivers, SDKs, sinks, and sources for JVM-based applications.
- Own the full lifecycle of data framework integrations, from design and implementation through maintenance and performance optimization.
- Develop high-performance integrations for streaming and data processing frameworks such as Apache Kafka, Kafka Connect, Apache Flink, Apache Beam, Spark, and related platforms.
- Build scalable data integration systems for real-time analytics, observability, and large-scale data workloads.
- Optimize reliability, developer experience, memory usage, garbage collection, concurrency, and data throughput.
- Collaborate with the open-source community, internal engineering teams, and enterprise users.
Requirements
- 6+ years of software development experience building and delivering high-quality, data-intensive solutions.
- Experience with streaming or data integration framework internals, preferably Apache Kafka, Kafka Connect, Apache Flink, or Apache Beam.
- Experience developing or extending production-grade connectors, sinks, or sources for a streaming processing framework.
- Hands-on experience with Apache Kafka or similar distributed messaging systems, including topic design, consumer groups, and performance tuning.
- Strong understanding of SQL, database fundamentals, data modeling, query optimization, and OLAP databases.
- Strong proficiency in Java and the JVM ecosystem, including memory management, garbage collection tuning, and performance profiling.
- Experience with concurrent programming in Java, including threads, executors, and reactive or asynchronous patterns.
- Understanding of JDBC, TCP/IP, HTTP, and network throughput optimization.
- Strong written and verbal communication skills.
- Passion for open-source development.
Nice to Have
- Contributions to open-source projects and active engagement with the open-source community.
- Familiarity with ClickHouse or similar high-performance data platforms.
- Working knowledge of Python for data engineering, including Pandas, PySpark, or Airflow.
Compensation and Benefits
- USD$500 home office setup allowance for remote employees.
- Healthcare contributions.
- Company stock options.
- Flexible time off in the United States and generous entitlement in other countries.
- Flexible, remote-friendly work environment.
- Opportunities to attend company-wide offsites.
Skills
Java, Jvm, Apache Kafka, Kafka Connect, Apache Flink, Apache Beam, SQL, Jdbc, Python, pandas, Pyspark, Apache Airflow, TCP/IP, Http, Garbage Collection
Similar jobs
Data Engineering jobsStaff Software Engineer building scalable frameworks for high-performance financial data ingestion/distribution and AI-native products. Requires 7+ years experience with distributed systems, microservices, and data architectures; partners with product teams to drive technical direction.
Build and operate scalable lakehouse infrastructure, streaming and CDC pipelines, query systems, and self-serve BI capabilities. Requires 5+ years of data engineering experience, strong Kubernetes and infrastructure-as-code expertise, and hands-on experience with distributed data platforms.
Leads database architecture, performance, reliability, and developer-tooling initiatives for high-volume trading applications. Requires 8+ years of software engineering experience, expert MySQL skills, backend development expertise, and strong knowledge of distributed systems and database operations.
Senior Data Engineer responsible for building and operating reliable clinical and claims data pipelines, CDC systems, quality controls, and de-identified exports. The role requires 5+ years of production pipeline experience plus strong SQL, Python, Spark, and data-governance skills.
Own the company’s metric governance program by defining canonical metrics, enforcing them in semantic and catalog systems, improving data quality, and validating AI-agent outputs. Requires 5+ years in analytics or analytics engineering, strong SQL, production semantic-layer ownership, and experience with AI evaluation and data governance.