5 days ago
Remote, IndiaSenior
Responsibilities
- Architect, upgrade, design, and build scalable infrastructure using Kubernetes, AWS, and RDS.
- Drive infrastructure roadmap initiatives focused on reliability, recoverability, scalability, and performance.
- Perform capacity planning, benchmarking, stress testing, bottleneck analysis, and growth preparation.
- Define and maintain SLAs, alerts, anomaly detection, and observability practices.
- Support AI enablement for infrastructure reliability, developer productivity, and internal tooling.
- Build consistency and scalability across distributed microservices architectures.
- Establish engineering practices for observability, security, and CI/CD.
- Mentor engineers, guide technical initiatives, and collaborate on technical roadmaps aligned with business goals.
Requirements
- 6+ years of professional software engineering or infrastructure engineering experience, including significant SRE and backend experience.
- Recent experience deploying significant changes to a production application or infrastructure configuration within the past 30 days.
- Strong proficiency in Golang and experience building and maintaining commercial APIs.
- Expertise with MySQL or PostgreSQL databases, including schema and query optimization at scale.
- Proficiency with Prometheus, Grafana, Datadog, or New Relic.
- Understanding of distributed systems patterns, event-driven architecture, stream processing, and queues.
- Demonstrated ability to influence decisions and lead complex technical initiatives.
- Required experience with Claude or equivalent large language model tools.
- Bachelor's degree in Computer Science, Engineering, or a related technical field.
- Preferred experience with AWS cloud infrastructure, especially Aurora RDS for MySQL and PostgreSQL.
- Preferred experience with data engineering, data pipelines, data warehousing, CI/CD pipelines, and containerized microservices in Kubernetes.
- Familiarity with Claude Code, Gemini CLI, Codex, or Cursor is preferred.
- Proven leadership in guiding technical direction, improving reliability, and scaling high-traffic services.
Benefits
- Full-time role based in India with a fixed schedule of 6:00 AM–2:00 PM EST.
- Remote work is indicated by the #Li-remote posting tag.
- Opportunities for career advancement, technical leadership, mentoring, and work on AI enablement initiatives.
Tech Stack
AWSDatadogElasticsearchGitGoGrafanaKubernetesMySQLPostgreSQLPrometheusPythonReactReact NativeTypeScript
Categories
DevOpsSite Reliability
