Fidelity

Principal Site Reliability Engineer

Fidelity
Apply
10 hours ago
Durham, NC, USAStaff+

Responsibilities

  • Define and lead enterprise-level reliability strategies.
  • Architect resilient systems and infrastructure for scalability, availability, and fault tolerance.
  • Design, implement, and maintain performance, load, stress, and chaos testing frameworks.
  • Create performance test reports and recommendations for quality improvement.
  • Implement observability practices using metrics, logs, tracing, dashboards, and alerting.
  • Establish and track SLOs, SLIs, error budgets, monitoring strategies, and incident response workflows.
  • Analyze system and application performance, diagnose bottlenecks, and perform capacity planning.
  • Automate operational workflows and build, deployment, and orchestration processes.
  • Develop innovative solutions to improve system availability, scalability, and performance.
  • Advise senior leadership on reliability engineering practices and mentor junior engineers.
  • Perform complex technical and functional analysis for multiple divisional initiatives.

Requirements

  • Bachelor’s degree in Computer Science, Engineering, Information Technology Management, Information Systems Security, Business Administration, or a closely related field and five years of experience as a Principal Site Reliability Engineer or closely related occupation.
  • Alternatively, a relevant master’s degree and three years of experience as a Principal Site Reliability Engineer or closely related occupation.
  • Experience implementing highly available trading systems in a financial services environment.
  • Expertise in performance benchmarking and engineering for online financial web applications, APIs, and mobile transactions using Rushhour, Locust, K6, and JMeter.
  • Experience configuring CI/CD and test automation using Jenkins, Sonar, Ant, Maven, Artifactory, and Terraform in AWS.
  • Experience designing scalable, resilient enterprise software platforms using AWS services including EC2, ECS, Lambda, EMR, and CloudFormation.
  • Experience developing microservices on EKS and using Bitbucket, GitHub, Artifactory, Sonar, Veracode, Helm, Java, Python, Spring Boot, Docker, and AWS.
  • Experience monitoring Apache, NGINX, Java, Node.js, Linux, and Windows environments using Splunk, Datadog, Kibana, Grafana, and AWS CloudWatch.
  • Experience with APM tools including Dynatrace, New Relic, Splunk, and Datadog, as well as capacity planning and performance tuning.
  • Experience implementing cloud-native and hybrid observability, Python automation, Infrastructure as Code, dashboards, alerting, incident response, and trace-level correlation.

Benefits

  • Fidelity is transitioning toward a full-time onsite working model through a phased rollout; onsite requirements vary by region and role and may evolve.
  • The position does not provide immigration sponsorship.

Tech Stack

Apache AntApache JMeterAWSDatadogDockerElasticsearchGrafanaHelmJavaJenkinsKibanaKubernetesLinuxLogstashMavenNode.jsPythonSonarQubeSplunkSpring BootTerraformWindows

Categories

Site ReliabilityTesting
Fidelity

About Fidelity

10,000+ employees

Fidelity Investments provides brokerage, retirement plan recordkeeping, wealth management, and asset management services to individuals, employers, advisors, and institutions, plus online trading platforms and mutual funds and ETFs. It earns fees from managing and administering assets, advisory services, and brokerage transactions. Founded in 1946 and headquartered in Boston, it is privately held and administers trillions of dollars for U.S. and global customers.

Contact me