
Lead Software Engineer (AI Platform) - Vice President
JPMorgan Chase11 hours ago
Base Salary
$157k - $215k/yr
Responsibilities
- Own the architecture, technical roadmap, and day-to-day engineering of the AI research platform.
- Design and operate secure, scalable cloud and CPU/GPU infrastructure, including clusters, scheduling, storage, networking, capacity, performance, and cost management.
- Build reproducible self-service research environments and platform automation using containers, GPU software stacks, dependency and artifact management, infrastructure-as-code, CI/CD, and orchestration.
- Translate requirements from LLM, agentic AI, model evaluation, retrieval, inference, and accelerator benchmarking projects into dependable platform capabilities.
- Package research prototypes for downstream integration through service interfaces and operational-readiness reviews.
- Evaluate and onboard AI infrastructure technologies, cloud services, accelerators, and developer tooling through benchmarks, architecture reviews, and controlled pilots.
- Maintain platform reliability through observability, incident response, patching, lifecycle management, access and data controls, risk and control evidence, documentation, and runbooks.
Requirements
- Master’s degree in computer science, software engineering, computer engineering, information systems, or a related technical field.
- At least 4 years of relevant industry experience building or operating cloud, platform, or infrastructure systems.
- Hands-on experience operating production environments, including EC2, networking, storage, monitoring, and cost management.
- Strong knowledge of Linux systems, containers, networking, storage, distributed systems, and orchestration or workload scheduling.
- Experience with infrastructure-as-code and delivery automation using Terraform or CloudFormation, configuration management, and CI/CD pipelines.
- Proficiency in Python and shell scripting, with sound software engineering practices for maintainable automation, services, and platform integrations.
- Working knowledge of AI/ML development workflows and the ability to translate research requirements into platform designs, communicate trade-offs, manage dependencies, and close risk and control items.
- Experience with cloud and accelerated-compute platforms, research and developer environments, AI workload enablement, or reliability and operations.
- Experience supporting AI/ML research or development teams and GPU-intensive, LLM, agentic, or distributed AI workloads is preferred.
- Experience with Kubernetes or managed container platforms, GPU software stacks, Slurm or Ray, observability platforms, and cloud cost optimization is preferred.
- Experience delivering infrastructure in a regulated enterprise and working with cybersecurity, architecture, technology risk, and audit stakeholders is preferred.
- Familiarity with finance or financial use cases is preferred.
Benefits
- Competitive total rewards package with base salary determined by role, experience, skills, and location; eligible roles may also receive incentive compensation.
- Comprehensive health care coverage, on-site health and wellness centers, retirement savings plan, backup childcare, tuition reimbursement, mental health support, and financial coaching.
- Equal opportunity employer with reasonable accommodations for religious practices, mental health needs, and physical disabilities.
About JPMorgan Chase
JPMorgan Chase provides consumer and commercial banking, payments, credit card, wealth management, and corporate and investment banking services to individuals, businesses, institutions, and governments. The public company (NYSE: JPM) earns revenue from interest, fees, trading, and asset management across operations in more than 100 markets. Headquartered in New York City with roots dating to 1799, it serves retail customers and prominent corporate and government clients through brands including Chase and J.P. Morgan.