2 hours ago
Bengaluru, IndiaSenior
Responsibilities
- Design, develop, and maintain large-scale data pipelines for ads reporting, attribution, and analytics.
- Use Hive, Spark, SQL, Scala, and Kafka to process and manage petabyte-scale datasets.
- Optimize data workflows for performance, scalability, and cost efficiency.
- Partner with data scientists, product managers, and platform engineers to deliver reliable datasets and APIs.
- Ensure data quality, integrity, and consistency across multiple data sources.
- Troubleshoot real-time streaming pipelines and batch data jobs.
- Evaluate new technologies to improve the Ads Data platform.
Requirements
- Strong programming experience in Scala, Java, or Python, with Scala preferred.
- Hands-on experience with Apache Spark for batch and streaming large-scale data processing.
- Proficiency in Hive, SQL, and data modeling for analytical workloads.
- Experience with Kafka for real-time event streaming.
- Understanding of big data ecosystems including S4, Hive, Presto, and Delta.
- Strong debugging, performance tuning, and problem-solving skills.
- Bachelor’s or master’s degree in computer science, engineering, or a related field.
- 4–6 years of experience in backend development.
- Experience in adtech, attribution, or campaign analytics.
- Familiarity with AWS EMR, GCP BigQuery, Databricks, and scheduling services such as AirFlow.
- Knowledge of data governance, security, and compliance best practices.
Tech Stack
Apache AirflowApache HiveApache KafkaApache SparkAWSDatabricksGoogle Cloud PlatformJavaPrestoPythonScalaSQL
Categories
Data Engineering
About Hotstar
We’ve moved! 🌟 Follow us on @JioHotstar for all the latest stories, updates & more.
