13 hours ago
Chennai, IndiaEntry Level
Responsibilities
- Develop and maintain batch and real-time data processing pipelines using Apache Spark Core and Spark SQL.
- Write scalable, reusable Scala code for large-scale distributed datasets.
- Optimize Spark jobs for performance and resource utilization.
- Implement data transformations, data cleansing, and aggregations for ETL/ELT processing.
- Collaborate with data engineers, analysts, and stakeholders to understand data requirements.
- Monitor and troubleshoot data workflows and production issues.
- Maintain code quality through unit testing and GitHub-based version control.
Requirements
- Bachelor’s degree in Computer Science, IT, or a related field.
- 1–2 years of experience with Big Data technologies.
- Hands-on experience building scalable data processing solutions with Apache Spark and Scala.
- Relevant Spark or Big Data certifications are optional.
Benefits
- Flexible working environment.
- Volunteer time off.
- LinkedIn Learning.
- Employee Assistance Program (EAP).
Tech Stack
Apache SparkScala
Categories
Data Engineering
