13 days ago
Bengaluru, IndiaStaff+
Responsibilities
- Own and evolve the architecture for services and tools that accelerate OCI compute and network capacity growth.
- Lead the design and delivery of resilient, scalable, secure, and operable services for the physical asset lifecycle.
- Translate infrastructure and business goals into technical strategy, architecture roadmaps, and phased execution plans.
- Prototype critical paths, review code and designs, solve difficult technical problems, and guide high-risk implementation.
- Drive alignment across engineering, product, operations, security, and partner organizations through durable interfaces and integration contracts.
- Define operational-excellence standards for observability, service health, incident learning, capacity planning, and reliability.
- Identify technical debt, systemic risks, and scaling bottlenecks and prioritize improvements using customer impact and data.
- Raise engineering standards through architectural reviews, mentorship, technical guidance, and constructive feedback.
Requirements
- BTech in Computer Science, Engineering, or a related field, or equivalent practical experience.
- 10+ years of professional software development experience designing, shipping, and operating scalable cloud-native distributed systems.
- Technical leadership experience for complex multi-service initiatives, including architecture, roadmaps, design reviews, and delivery across teams.
- Strong proficiency in one or more high-level programming languages, preferably Go or Java, plus scripting or automation experience such as Python.
- Deep knowledge of distributed systems, microservice design patterns, service-to-service communication, data consistency, fault tolerance, and highly available services.
- Experience designing production observability using service-level indicators and objectives, dashboards, metrics, alarms, capacity planning, and operational runbooks.
- Production-operations experience including incident response, root-cause analysis, on-call practices, and systemic reliability improvements.
- Strong computer science foundation in data structures, algorithms, concurrency, and programming paradigms.
- Ability to author and communicate technical proposals, architecture documents, design specifications, and presentations.
- Experience mentoring engineers and improving practices through code reviews, design reviews, and technical guidance.
- Preferred: MTech or PhD in Computer Science, Engineering, or a related field, or equivalent experience.
- Preferred: experience with cloud-provider control-plane or data-plane solutions, cloud-scale data-center physical asset lifecycles, DCIM systems, capacity-planning tools, Kubernetes, container orchestration, Infrastructure as Code, and Terraform.
- Preferred: experience diagnosing large-scale performance, scaling, and reliability issues and establishing technical standards or reusable platform capabilities adopted across teams.
Benefits
- Competitive benefits including flexible medical, life insurance, and retirement options.
- Employee volunteer programs and community-giving opportunities.
- Accessibility assistance and accommodation are available throughout the employment process.
About Oracle
Oracle builds database technology, enterprise applications, and Oracle Cloud Infrastructure for companies and governments. Its business spans cloud subscriptions, software licenses, and support services across ERP, HCM, CX, and industry suites, plus NetSuite. Founded in 1977 and headquartered in Austin, Texas, Oracle is a public company traded on the NYSE and serves global customers migrating and running critical workloads in its cloud.
