6 days ago
Bengaluru, IndiaSenior
Responsibilities
- Design and implement monitoring, alerting, and observability for key services.
- Provide reliability, scalability, and performance expertise for new features and system changes.
- Write scripts and tools to automate operational tasks, reduce toil, and improve engineering velocity.
- Participate in on-call rotations and assist with incident response and blameless post-mortems.
- Contribute to disaster recovery and resiliency testing.
- Support migration of Markets applications to Google Cloud Platform.
Requirements
- Require 5-8+ years of professional experience in Site Reliability, DevOps, Software, or Systems Engineering.
- Demonstrate the ability to automate operational tasks using Python or Bash.
- Have an interest in and understanding of Site Reliability Engineering principles.
- Preferred experience includes Google Cloud Platform, GCE, GKE, Prometheus, Grafana, OpenTelemetry, Splunk, Kubernetes, and Docker.
- Preferred knowledge includes large-scale distributed systems, HTTP, TCP/UDP, IP networking, financial markets, or message-oriented middleware.
- Strong problem-solving, analytical, communication, and teamwork skills are required.
Benefits
- Competitive compensation and benefits package.
- Career growth in Site Reliability Engineering within a collaborative and innovative technology organization.
Tech Stack
Categories
Site Reliability
About CME Group
CME Group operates global derivatives exchanges and a central counterparty clearinghouse, offering futures and options across interest rates, equities, FX, energy, agriculture, and metals via CME Globex and CME Clearing. It serves banks, asset managers, corporations, and professional traders with trading, market data, and risk management tools, earning fees from transactions, clearing, and data services. A public company headquartered in Chicago, it owns the CME, CBOT, NYMEX, and COMEX marketplaces.
