2 hours ago
Pune, IndiaStaff+
Responsibilities
- Lead architecture and design for major platform initiatives spanning multiple services and engineering teams.
- Design and build highly available, fault-tolerant distributed systems for the Prism management plane.
- Analyze performance and conduct micro-benchmarking to improve scalability and resource efficiency.
- Drive technical decisions through design reviews, architecture discussions, and technical RFCs.
- Evolve Prism APIs, management workflows, lifecycle management, and platform infrastructure.
- Improve deployment, upgrade, observability, diagnostics, and production-operations workflows, including zero-downtime upgrades.
- Lead cross-team technical execution, identify risks, and balance short-term delivery with long-term maintainability.
- Collaborate with product management, customer support, and platform teams to translate requirements into scalable solutions.
- Raise engineering quality through code reviews, design reviews, technical coaching, testing automation, performance measurement, and continuous improvement.
- Mentor engineers and help develop future technical leaders.
Requirements
- 10+ years of professional software engineering experience building distributed systems, infrastructure software, or large-scale backend platforms.
- Demonstrated experience leading the technical design and delivery of complex, cross-team engineering initiatives.
- Strong understanding of distributed systems architecture, consensus and coordination, high availability, fault tolerance, concurrency, synchronization, replication, consistency models, and service-oriented architectures.
- Strong proficiency in Go or another systems programming language, with the ability to contribute across multiple technology stacks.
- Production systems experience on Linux, including networking fundamentals, containers, Kubernetes, storage systems, virtualization concepts, performance analysis, and debugging.
- Experience designing and evolving REST and/or gRPC APIs with attention to compatibility, versioning, security, and developer usability.
- Working knowledge of authentication, authorization, TLS, X.509 certificates, identity management, or secure service-to-service communication.
- Experience driving architecture reviews, design documentation, operational readiness, production incident analysis, CI/CD improvements, and reliability engineering.
- Preferred experience with management or control plane software, cloud infrastructure, virtualization platforms, Kubernetes operators, cloud-native architectures, observability platforms, highly available upgrade and lifecycle systems, or open-source infrastructure and distributed systems projects.
Benefits
- Hybrid work arrangement with remote and in-person collaboration; in applicable locations, employees are expected to work onsite at least 3 days per week.
- Workplace type and team-specific guidance may vary by location and team requirements.
Tech Stack
About Nutanix
Nutanix is a global leader in cloud software, offering organizations a single platform for running apps and data across clouds. With Nutanix, companies can reduce complexity and simplify operations, freeing them to focus on their business outcomes. Building on its legacy as the pioneer of hyperconverged infrastructure, Nutanix is trusted by companies worldwide to power hybrid multicloud environments consistently, simply, and cost-effectively. Learn more at www.nutanix.com or follow us on social media @nutanix.
