1 day ago
Pune, IndiaStaff+
Responsibilities
- Design, develop, and maintain scalable enterprise-grade AI agents and ELT/ETL processes for large data volumes.
- Build and deploy generative AI agents using Google ADK and Google Flash 2.5+ LLMs with human-in-the-loop architecture.
- Develop data federation layers supporting lambda and Data Mesh architectures using tools such as Starburst.
- Develop, deploy, and automate microservice integrations for data-intensive applications.
- Implement cloud-native infrastructure and OpenShift or Kubernetes architectures, including CI/CD pipelines.
- Integrate agentic AI tools and platforms through advanced prompt engineering.
- Ensure data quality, integrity, security, regulatory compliance, and appropriate risk management across the data lifecycle.
- Improve data engineering processes, standards, and best practices while collaborating with business and technology stakeholders.
Requirements
- Bachelor’s degree in Computer Science, Engineering, or a related field.
- At least 8 years of overall experience in large-scale application development.
- At least 5 years of experience in a Python and PySpark engineering lead role focused on enterprise-grade, high-volume ELT/ETL processes.
- Hands-on experience developing agentic AI solutions with YAML, JSON, FastAPI or Spring Boot, Google ADK, LLM integrations, Devin.AI or GitHub Copilot, MCP, and advanced prompt engineering.
- Experience developing and automating microservice integrations for data-intensive applications.
- Proficiency in Python, Scala, or another programming language used for data analytics and engineering.
- Strong SQL skills and experience with relational databases.
- Deep understanding of data modeling, data warehousing, Data Mesh architecture, and data federation.
- Preferred experience with Cloudera, Databricks, AWS, Azure, or GCP cloud-based Big Data platforms.
- Preferred experience with Angular or React.js for data-driven interfaces.
- Preferred practical experience applying AI/ML techniques to business problems.
- Preferred familiarity with Docker and Kubernetes.
- Preferred experience in retail banking products such as Cards, Mortgage, Deposits, or Wealth Management.
- Relevant industry certifications, such as AWS Certified Big Data - Specialty or Azure Data Engineer Associate, are preferred.
- A master’s degree is a plus.
Benefits
- Full-time position.
Tech Stack
AngularApache KafkaAWSAzureDatabricksDockerFastAPIGoogle Cloud PlatformKubernetesOpenShiftPythonReactScalaSpring BootSQL
Categories
AI ApplicationsData Engineering
About Citi
Citi is a public financial-services company offering consumer and institutional banking, credit cards, wealth management, treasury and trade solutions, and capital-markets services. It serves individuals, corporations, financial institutions, and governments in more than 160 countries and jurisdictions, earning interest and fee income from lending, payments, trading, and advisory. Founded in 1812 and headquartered in New York, it trades on the NYSE under the ticker C.
