Charles Schwab

Senior Site Reliability Engineer

Charles Schwab
Apply
1 day ago
Austin, TX, USA or Southlake, TX, USASenior

Base Salary

$129k - $175k/yr

Responsibilities

  • Respond to system alerts and production incident escalations and lead or support triage, resolution, root cause analysis, and post-incident reviews.
  • Improve monitoring coverage, alerting, telemetry, dashboards, and visibility into system performance and health.
  • Design and build automation, scripts, and tooling using Python and shell scripting to reduce operational toil and improve resilience.
  • Contribute to CI/CD and deployment pipeline improvements, including automated service recovery, system maintenance, and certificate management.
  • Partner with development, architecture, infrastructure, and product teams to embed reliability and production-readiness practices into the software development lifecycle.
  • Drive improvements in reliability, scalability, resilience, observability, incident reduction, and operational maturity.
  • Strengthen on-call practices and explore AI and automation opportunities for incident detection, triage, and response.
  • Mentor junior engineers and influence teams to adopt proactive reliability and observability practices.

Requirements

  • 10+ years of experience in software development and site reliability engineering, including cloud-native architectures and distributed systems.
  • 8+ years of experience in DevOps and/or site reliability engineering focused on production operations, automation, and system reliability at scale.
  • 8+ years of experience with CI/CD pipelines, observability, and monitoring or telemetry platforms.
  • 5+ years of experience leading reliability engineering practices such as service level objectives, monitoring strategies, incident reviews, and automation-driven improvements.
  • Experience designing, developing, and maintaining production-grade systems, automation frameworks, and reliability tooling across the software development lifecycle.
  • Experience supporting high-availability distributed systems at scale, with deep monitoring, observability, incident management, automation, scripting, and operational tooling experience.
  • Bachelor of Science degree in Computer Science, a related field, or equivalent work experience.
  • Preferred experience with Python or Java, Splunk, Kubernetes, Terraform or similar infrastructure-as-code tools, and AWS, GCP, or Azure.
  • Strong technical leadership, cross-functional influence, communication, analytical, problem-solving, and incident coordination skills.

Benefits

  • The role is intended to be performed onsite in the specified location(s).
  • The role is eligible for bonus or incentive opportunities in addition to the salary range.

Categories

Site Reliability
Charles Schwab

About Charles Schwab

10,000+ employees

Charles Schwab provides brokerage, banking, and wealth management services for individual investors and independent investment advisors. Its business spans trading platforms, advisory and custody services (Schwab Advisor Services), ETFs and mutual funds, and a U.S. bank offering deposits and lending. Founded in 1971 and headquartered in Westlake, Texas, Schwab is publicly traded on the NYSE (SCHW) and expanded its retail and advisor footprint through the acquisition of TD Ameritrade.

Contact me