
Senior Lead Site Reliability Engineer
JPMorgan Chase5 hours ago
Base Salary
$176k - $260k/yr
Responsibilities
- Lead the Production Management team supporting Audit and Credit Review, including setting direction, priorities, and performance expectations.
- Own stability, availability, resiliency, and end-to-end operational performance for business-critical Sales platforms.
- Serve as a senior escalation point during critical incidents and lead triage, coordination, recovery, root-cause analysis, and remediation.
- Drive operational consistency through standardization, governance participation, service reviews, and support-model integration.
- Improve reliability through observability, monitoring, automation, operational analytics, systems thinking, and SRE adoption.
- Lead incident, problem, and change management, including readiness testing, monitoring, and resiliency or disaster-recovery validation.
- Use and scale enterprise-authorized AI capabilities for incident triage, troubleshooting, knowledge capture, SDLC/toolchain workflows, testing, validation, and operational readiness while enforcing appropriate controls.
Requirements
- Formal training or certification in site reliability engineering concepts and 10+ years of applied experience.
- Demonstrated experience using enterprise-authorized AI capabilities to improve SRE workflows, with strong validation practices and awareness of data sensitivity.
- Ability to assess AI-assisted operational recommendations, define team-use guardrails, and align outcomes with resiliency and security expectations.
- Leadership experience across Production Support, Production Management, SRE, and Technology Operations teams.
- Strong systems thinking and problem-solving skills for complex, cross-domain production issues.
- Demonstrated major-incident leadership and experience coordinating service restoration and root-cause remediation.
- Experience partnering with Application Development, Product, Sales, and business stakeholders to improve reliability and operational maturity.
- Strong observability and service-management expertise, including Dynatrace, Splunk, Geneos, Grafana, and ITIL incident, problem, change, and availability practices.
- Preferred experience integrating support teams, standardizing operating models, and working with AWS, Azure, GCP, Python, Shell, PowerShell, Ansible, Terraform, containers, microservices, Kubernetes, and OpenShift.
Benefits
- Competitive total rewards package with base salary determined by role, experience, skill set, and location, plus possible incentive compensation for eligible roles.
- Benefits may include health care coverage, on-site health and wellness centers, retirement savings, backup childcare, tuition reimbursement, mental health support, and financial coaching.
- JPMorgan Chase is an equal opportunity employer and provides reasonable accommodations.
Tech Stack
Categories
Site Reliability
About JPMorgan Chase
JPMorgan Chase provides consumer and commercial banking, payments, credit card, wealth management, and corporate and investment banking services to individuals, businesses, institutions, and governments. The public company (NYSE: JPM) earns revenue from interest, fees, trading, and asset management across operations in more than 100 markets. Headquartered in New York City with roots dating to 1799, it serves retail customers and prominent corporate and government clients through brands including Chase and J.P. Morgan.