almost 2 years ago
Bengaluru, IndiaStaff+
Responsibilities
- Design, implement, and optimize distributed data systems built on Apache open-source technologies.
- Lead technical discussions, define architecture improvements, and guide teams in implementing systems and subsystems.
- Submit patches, review contributions, mentor contributors, and help drive the technical direction of the project.
- Troubleshoot complex software defects, reproduce reported issues, identify root causes, and develop workarounds and permanent solutions.
- Address bug reports and security vulnerabilities while maintaining software quality, reliability, performance, and currency.
- Collaborate with internal teams, senior leadership, external contributors, and the broader open-source community.
Requirements
- 9+ years of software engineering experience, including at least 4–5 years actively contributing to open-source projects, ideally in the Apache ecosystem.
- Deep knowledge of one or more Apache projects such as Hadoop, Kafka, Spark, Hive, or Flink and their underlying architectures.
- Strong coding, debugging, troubleshooting, and large-codebase navigation skills.
- Proficiency in one or more of Java, Python, Scala, or Go; experience with distributed computing and high-performance applications is highly preferred.
- A proven record of significant open-source contributions, including submitting patches, reviewing contributions, and engaging with the community.
- Experience deploying and managing open-source solutions in AWS, Azure, or GCP cloud environments and on-premise environments.
- Ability to communicate complex technical concepts clearly, collaborate across teams, and guide technical implementation.
Tech Stack
Apache FlinkApache HadoopApache HiveApache KafkaApache SparkAWSAzureGoGoogle Cloud PlatformJavaPythonScala
Categories
BackendData Engineering
