
Staff Site Reliability Engineer
Modernizing Medicine, Inc.7 days ago
Hyderābād, IndiaStaff+
Responsibilities
- Lead the long-term vision and architecture for the AWS cloud ecosystem, including fault-tolerant, scalable, and cost-effective infrastructure.
- Establish organization-wide observability standards using DataDog.
- Improve Jenkins-based CI/CD automation and platform efficiency for engineering teams.
- Lead Kubernetes optimization, security, and deployment patterns across the organization.
- Partner with security teams on infrastructure protection and HIPAA and SOC 2 compliance.
- Act as the ultimate escalation point for complex system failures and lead incident post-mortems.
- Mentor senior engineers and lead cross-functional technical initiatives.
Requirements
- 10–13+ years of experience in site reliability engineering or cloud architecture.
- Deep expertise in AWS, including EC2, Lambda, RDS, and S3.
- Professional-level experience with Kubernetes, DataDog, Terraform, and Ansible.
- Expert scripting skills in Python and Bash.
- Ability to communicate technical vision to executives and conduct deep architectural reviews.
- Preferred: AWS Professional or Specialty certifications in Architect, Security, or DevOps.
- Preferred: experience with Kafka, Kinesis, or Redshift.
- Preferred: open-source contributions or speaking at industry conferences.
Benefits
- Complimentary office lunches and dinners on select days, plus healthy snacks delivered to desks.
- Comprehensive health, accidental, and life insurance plans, including family coverage, at no cost to employees.
- Annual wellness allowance.
- Earned, casual, sick, bereavement, extended medical, and company-paid holiday leave.
- Paid parental leave including maternity, paternity, adoption, surrogacy, and abortion leave.
- Celebration leave.
Tech Stack
Categories
DevOpsSite Reliability