
Senior Software Engineer - Data Integration & JVM Ecosystem
ClickHouse3 months ago
Responsibilities
- Own the full lifecycle of JVM-based data framework integrations.
- Build and maintain core database drivers, SDKs, and connectors for ClickHouse.
- Develop tools that enable data engineers to use ClickHouse in JVM-based applications.
- Improve performance, reliability, and developer experience for integrations handling massive datasets.
- Collaborate with the open-source community, internal engineering teams, and enterprise users.
- Contribute to official language clients, data connectors, and related ecosystem integrations as needed.
Requirements
- At least 5 years of software development experience focused on high-quality, data-intensive solutions.
- Strong proficiency in Java and the JVM ecosystem, including memory management, garbage collection tuning, and performance profiling.
- Experience with concurrent Java programming, including threads, executors, and reactive or asynchronous patterns.
- Experience developing, extending, or using connectors, sinks, or sources for a big data framework such as Apache Spark, Flink, Beam, or Kafka Connect.
- Strong understanding of SQL, data modeling, query optimization, and OLAP or analytical databases.
- Strong written and verbal communication skills.
- Passion for open-source development and active engagement with OSS communities.
- Open-source contributions, ClickHouse or similar platform experience, and expertise building big data sinks or source connectors are preferred.
- Python experience in data engineering contexts, including Pandas, PySpark, or Airflow, is preferred.
- Knowledge of JDBC, TCP/IP, HTTP, and data-throughput optimization is preferred.
Benefits
- Flexible, remote-friendly work environment across more than 20 countries.
- Employer healthcare contributions.
- Company equity through stock options.
- Flexible time off in the US and generous leave entitlement in other countries.
- $500 home office setup benefit for remote employees.
- Opportunities to attend company-wide global offsites.
Tech Stack
Apache AirflowApache BeamApache FlinkApache KafkaApache SparkC++ClickHousedbtGoGrafanaJavaJavaScriptMetabasePandasPythonRustSQL
Categories
BackendData Engineering
About ClickHouse
ClickHouse is a fast, open-source columnar database built for real-time data processing and analytics at scale. ClickHouse Cloud delivers the query speed and concurrency that applications demanding instant insight from large volumes of data require. As AI agents become more embedded in software, generating higher query volumes at tighter latency, ClickHouse provides a high-throughput, low-latency engine purpose-built for that workload.