1 month ago
Bengaluru, IndiaSenior
Responsibilities
- Develop ETL/ELT pipelines using AWS Glue, Step Functions, Lambda, and DMS.
- Implement direct-to-Aurora ingestion strategies for snapshots and delta loads.
- Ensure referential integrity, processing order, idempotency, and recovery mechanisms.
- Build monitoring, validation, and reconciliation frameworks for production data pipelines.
- Manage scalability and throughput for high-volume end-of-day ingestion workloads.
- Own production readiness and end-to-end solution delivery, including monitoring, alerting, and runbooks.
Requirements
- Minimum 3 years of relevant experience, typically reflecting 5+ years in data engineering, data integration, or related roles.
- Expert-level SQL development and performance optimization skills.
- Strong proficiency in SQL, Python, and PySpark.
- Hands-on experience with ETL pipelines and orchestration frameworks.
- Understanding of ACID-compliant data ingestion, schema evolution, and CDC patterns.
- Experience building resilient and idempotent ingestion frameworks and supporting high-volume, time-sensitive processing.
- BE, B.Tech, ME, M.Tech, MBA, MCA, or equivalent qualification preferred.
- Preferred experience with SAP ODP, enterprise data replication technologies, distributed PostgreSQL systems, or Kafka.
Benefits
- PwC offers inclusive benefits, flexibility programs, wellbeing support, mentorship, and professional growth opportunities.
- The role is based in Bangalore, with no specific remote or hybrid arrangement stated.
- The posting end date is August 25, 2026.
Tech Stack
Categories
Data Engineering
About PwC
PwC is a global network of professional services firms providing assurance, tax, and advisory work for businesses and public-sector clients. It sells project-based and managed consulting, audit, and deals services, and implements and integrates business applications. Formed in 1998 by the merger of Price Waterhouse and Coopers & Lybrand, it is headquartered in London and is one of the Big Four accounting firms.
