JPMorgan Chase

Lead Software Engineer - Java/Python, AWS, Spark

JPMorgan Chase
Apply
2 days ago
Pune, IndiaStaff+
H1B Sponsor

Responsibilities

  • Lead the design, development, maintenance, and governance of scalable cloud-based data pipelines and infrastructure.
  • Architect enterprise data models and optimize large-scale datasets for storage, retrieval, analytics, integrity, and quality.
  • Translate business requirements into scalable data engineering solutions and align technical practices with organizational objectives.
  • Manage end-to-end data infrastructure from design and construction through installation and ongoing maintenance.
  • Drive data quality, accessibility, regulatory compliance, process improvement, and reusable engineering patterns.
  • Author, review, and approve technical requirements and architectural designs while leading design workshops, coding sessions, and hackathons.
  • Lead responsible adoption of enterprise-authorized AI-assisted engineering and automation tools, including validation of generated outputs for correctness, performance, and security.
  • Coach engineers on secure, compliant, and effective use of AI-assisted software development practices.
  • Support budgeting, resource allocation, and vendor relationship management for data engineering projects.

Requirements

  • Formal training or certification in software engineering concepts and at least 5 years of applied experience.
  • Expertise with Apache Spark and at least one cloud data lakehouse platform, such as AWS data lake services or Databricks.
  • Expertise with Airflow, AWS Step Functions, or a similar orchestration tool, plus relational and NoSQL databases.
  • Strong knowledge of data structures, JSON, Avro, Protobuf, Parquet, Iceberg, or similar serialization and storage formats.
  • Professional programming experience with Python and SQL plus Java, Scala, or another additional language for data engineering.
  • Hands-on experience developing and optimizing Apache Spark pipelines for batch processing, real-time analytics, machine learning, and data transformation.
  • Experience with Docker, Kubernetes, serverless computing, distributed cluster computing, and data modeling techniques such as Dimensional, Data Vault, Kimball, or Inmon.
  • Experience with TDD or BDD and CI/CD tools.
  • Experience with streaming platforms such as Kafka or MQ.
  • Experience leading approved AI-assisted development tools and establishing standards for validating AI outputs and handling sensitive data securely.
  • Preferred experience with Terraform, AWS CloudFormation, Spinnaker, Snowflake, project budgeting and resource allocation, and vendor management.

Tech Stack

Apache AirflowApache HadoopApache KafkaApache SparkAWSDatabricksDockerJavaKubernetesPythonScalaSnowflakeSpinnakerSQLTerraform

Categories

Data Engineering
JPMorgan Chase

About JPMorgan Chase

10,000+ employees

With a history tracing its roots to 1799 in New York City, JPMorganChase is one of the world's oldest, largest, and best-known financial institutions—carrying forth the innovative spirit of our heritage firms in global operations across 100 markets. We serve millions of customers and many of the world’s most prominent corporate, institutional, and government clients daily, managing assets and investments, offering business advice and strategies, and providing innovative banking solutions and services. Social Media Terms and Conditions: https://bit.ly/JPMCSocialTerms JPMorgan Chase & Co. is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Contact me