
Site Reliability Engineer
Luma Financial Technologies2 months ago
Cincinnati, OH, USASenior
Responsibilities
- Design and build infrastructure for product engineering services.
- Operate and scale Kubernetes clusters on AWS EKS.
- Design resilience strategies covering multi-region architecture, backups, disaster recovery, and failover.
- Automate infrastructure with Terraform and Infrastructure-as-Code.
- Improve CI/CD pipelines and deployment practices.
- Monitor system performance and reliability with observability tools.
- Participate in on-call rotations and lead incident response.
- Perform root-cause analysis and implement long-term reliability fixes.
Requirements
- 5+ years of applicable experience in Site Reliability or Software Development Engineering.
- Strong programming experience with Java, JavaScript, Python, Bash, and Go.
- Strong experience with AWS services including RDS, CloudFront, IAM, and VPCs.
- Strong experience with Terraform and Kubernetes.
- Experience designing and operating systems that remain dependable during failures and recover seamlessly.
- Hands-on experience improving and operating CI/CD pipelines, such as CircleCI or GitHub Actions.
- Incident response, root-cause analysis, communication, collaboration, and continuous-improvement skills.
- Bachelor’s degree in Computer Science, Software Engineering, or a related concentration is highly preferred.
Benefits
- Hybrid work arrangement.
- On-call rotation participation is part of the role.
- Luma does not provide current or future employment-based immigration sponsorship for this position.
Tech Stack
Categories
DevOpsSite Reliability