2 months ago
Remote, India or Pune, IndiaSenior

Responsibilities

  • Contribute fixes and enhancements to the Apache Spark open-source project.
  • Upgrade and maintain compatibility with newer Spark releases.
  • Optimize Spark internals for performance, scalability, and reliability.
  • Design and develop distributed data processing solutions.
  • Debug complex Spark-related performance and stability issues.
  • Collaborate with product and platform teams on architecture and technical initiatives.

Requirements

  • Strong experience with Apache Spark internals and distributed systems.
  • Hands-on experience contributing to open-source projects.
  • Strong Java and/or Scala programming skills.
  • Expertise in query execution, distributed processing, and performance optimization.
  • Strong debugging and problem-solving abilities.
  • Ability to work independently and drive technical initiatives.
  • Preferred: experience with Kubernetes and cloud-native environments, JVM tuning and system optimization, Apache or other open-source communities, and large-scale analytics or data platforms.

Categories

Data Engineering
InfraCloud Technologies

About InfraCloud Technologies

51-200 employees
Contact me