4 months ago
Bengaluru, IndiaStaff+
Responsibilities
- Define and evolve the end-to-end architecture for core platform services and distributed systems.
- Own architectural decisions involving service orchestration, data flows, extensibility, scalability, reliability, security, privacy, and operational excellence.
- Lead architectural reviews, RFCs, long-term technical planning, and cross-system workflow design.
- Set direction for cloud-native compute, storage, networking, deployment, containerization, orchestration, and infrastructure-as-code practices.
- Design foundational APIs, frameworks, abstractions, platform contracts, service boundaries, and reusable internal standards.
- Own production readiness, availability, latency, scalability, fault isolation, SLAs, SLOs, observability, alerting, and incident response for mission-critical components.
- Mentor Staff and Senior engineers and influence platform direction across engineering, AI research, product, security, and design teams.
Requirements
- 12+ years of experience designing, building, and operating large-scale, production-grade software systems.
- Proven ownership of complex distributed systems or platform architectures.
- Deep expertise in cloud-native systems, service-oriented design, and distributed architectures.
- Strong hands-on experience with Docker, Kubernetes, and modern cloud platforms.
- Expert understanding of software design principles, clean code practices, and system-level trade-offs.
- Strong proficiency in Python and deep experience with at least one of Java, Go, or JavaScript.
- Ownership across the full software development lifecycle, including design, implementation, testing, deployment, and operations.
- Experience with Infrastructure as Code such as Terraform or CloudFormation.
- Understanding of observability, metrics, logging, and distributed tracing.
- Experience with event-driven systems, streaming platforms such as Kafka, and SQL and NoSQL databases.
- Familiarity with enterprise security, access control, and compliance considerations.
- Experience designing cross-system workflows and implementing BPMN-based or workflow-driven systems.
- Experience with AI/ML or data-intensive platforms, MLOps, AI service deployment, LLMs, agents, retrieval, or evaluation pipelines is a strong plus.
- Ability to make architectural decisions under ambiguity, lead through technical depth and influence, and own long-term architectural consequences.
Benefits
- Opportunity to build a foundational enterprise AI platform with real-world production impact.
- Work with deep technology, AI systems, and auditable, human-aligned enterprise applications.
- Headquartered in Los Altos, California.
