4 months ago
Hyderābād, IndiaSenior
Responsibilities
- Design and develop batch and streaming data pipelines using Python and Spark/PySpark.
- Build AI-ready datasets, feature pipelines, training data foundations, feature-store capabilities, and reproducible datasets.
- Develop backend services and APIs using Python with FastAPI or an equivalent framework.
- Implement microservices and event-driven architectures for data ingestion, processing, and serving.
- Support LLM/AI integrations, including RAG, embeddings, inference APIs, vectorization, and retrieval workflows.
- Collaborate on Azure, AWS, or GCP infrastructure, including data lakes, lakehouses, SQL/NoSQL systems, containerized architectures, and serverless architectures.
- Implement data quality, metadata, lineage, governance, CI/CD, observability, reliability, and performance practices.
- Mentor engineers on data engineering patterns and best practices.
Requirements
- 8+ years of experience in data engineering and backend development.
- Strong Python, API, and microservices experience.
- Deep experience with Spark/PySpark or an equivalent technology.
- Expertise in ETL/ELT pipelines and lakehouse architectures.
- Experience with feature stores and AI-ready datasets.
- Experience with Azure, AWS, or GCP.
- Familiarity with Docker, Kubernetes, CI/CD, LLM/GenAI integration patterns, and metadata and governance tooling.
- Strong system design and problem-solving skills.
Tech Stack
Categories
BackendData Engineering
About Cognizant
Cognizant is a public IT services and consulting firm that designs, builds, and runs enterprise technology, including digital engineering, cloud modernization, data/AI, and managed services. It sells consulting, systems integration, and outsourcing on multi-year engagements to large enterprises in healthcare, banking, retail, communications, and manufacturing. Founded in 1994 and headquartered in Teaneck, New Jersey, Cognizant is NASDAQ-listed (CTSH) and a Fortune 500 company.
