
Senior Software Engineer - Data Integration & JVM Ecosystem
ClickHouse3 months ago
Responsibilities
- Own and maintain critical components of ClickHouse’s data engineering ecosystem
- Develop and maintain database drivers, SDKs, and connectors for JVM-based applications
- Build integrations with data-processing frameworks including Spark, Flink, Beam, and Kafka Connect
- Improve the performance, reliability, and developer experience of JVM integrations
- Optimize memory usage, concurrency, query performance, and data throughput for large-scale workloads
- Collaborate with the open-source community, internal teams, and enterprise users
Requirements
- 6+ years of software development experience focused on high-quality, data-intensive solutions
- Proven experience with the internals of at least one of Apache Spark, Apache Flink, Kafka Connect, or Apache Beam
- Experience developing or extending connectors, sinks, or sources for a big data processing framework
- Strong understanding of SQL, data modeling, query optimization, and OLAP or analytical databases
- Track record of building scalable data integration systems beyond simple ETL jobs
- Strong proficiency in Java and the JVM ecosystem, including memory management, garbage collection tuning, and performance profiling
- Experience with concurrent Java programming, including threads, executors, and reactive or asynchronous patterns
- Understanding of JDBC, TCP/IP, HTTP, and techniques for optimizing network data throughput
- Strong written and verbal communication skills
- Passion for open-source development
- Prior open-source contributions are preferred
- Familiarity with ClickHouse or similar high-performance data platforms is preferred
- Working knowledge of Python and data engineering tools such as Pandas, PySpark, and Airflow is preferred
Benefits
- Flexible, remote-friendly work environment across more than 20 countries
- Employer healthcare contributions
- Company stock options
- Flexible time off in the US and generous time-off entitlement in other countries
- $500 home office setup for remote employees
- Opportunities to attend company-wide global offsites
Tech Stack
Apache AirflowApache BeamApache FlinkApache KafkaApache SparkClickHousedbtGrafanaJavaMetabasePandasPythonSQL
Categories
BackendData Engineering
About ClickHouse
ClickHouse is a fast, open-source columnar database built for real-time data processing and analytics at scale. ClickHouse Cloud delivers the query speed and concurrency that applications demanding instant insight from large volumes of data require. As AI agents become more embedded in software, generating higher query volumes at tighter latency, ClickHouse provides a high-throughput, low-latency engine purpose-built for that workload.