2 hours ago
Bengaluru, IndiaMid Level
Responsibilities
- Own and deliver major data-platform components from design through production rollout.
- Design and evolve distributed systems for ingestion, streaming, lakehouse and warehouse workloads, cataloging, and governance.
- Build computation-graph management, pipeline-optimization, and dataset-lifecycle systems for the Golden Data Sets Platform.
- Author architecture documents and participate in design reviews focused on scalability, resilience, latency, correctness, cost, security, and compliance.
- Drive operational excellence through observability and incident response for owned services.
- Collaborate with product, infrastructure, and analytics teams and contribute to governance, orchestration, and privacy engineering.
Requirements
- 10+ years of software engineering experience, primarily in distributed systems or data platforms.
- Experience designing and delivering large-scale distributed systems involving compute, storage, APIs, or streaming.
- Ability to independently deliver complex projects from requirements through production.
- Proficiency in Java and Python and experience with containerized environments.
- Hands-on expertise with Kafka or Flink, Spark, Delta or Iceberg, Kubernetes, and NoSQL or columnar stores.
- Experience with streaming and batch data platforms and a strong foundation in algorithms and distributed design.
- BS/MS in Computer Science or equivalent experience.
- Strong communication, systems thinking, and ability to anticipate bottlenecks, schema evolution, and reliability issues.
Benefits
- Flexibility, benefits, and career development resources.
- Collaborative and inclusive culture with knowledge-sharing opportunities.
- Focus on reliability and sustainable on-call practices.
- Work contributing to global analytics and machine-learning systems.
Tech Stack
Categories
BackendData Engineering
