1 day ago
Base Salary
$115k - $235k/yr
Responsibilities
- Lead the design, implementation, and evolution of core distributed systems and data-plane services at hyperscale.
- Define scalability, elasticity, durability, availability, and reliability requirements for owned components.
- Optimize high-throughput retrieval, storage, and processing paths using distributed state, replication, and synchronization patterns.
- Design fault-tolerant systems with redundancy, automatic failover, recovery-oriented design, and in-service updates.
- Apply reliability techniques including load shedding, throttling, rate limiting, retries, and timeouts.
- Establish SLOs, KPIs, telemetry, dashboards, and proactive alerting for critical systems.
- Lead performance, load, fault-injection, and brownout testing.
- Diagnose and recover from production incidents, guide root-cause analysis, and mentor engineers.
- Build Infrastructure as Code and operational automation for safe patching, updates, rollbacks, and change management.
- Apply security controls and remediation practices, including encryption, access controls, and compliance readiness.
Requirements
- Bachelor’s or master’s degree in Computer Science, Computer Engineering, or a related field, or equivalent practical experience.
- 7+ years of professional software-engineering experience with demonstrated impact on large-scale distributed systems or cloud infrastructure.
- Strong experience designing and operating highly available, scalable, and fault-tolerant distributed systems.
- Proficiency in one or more object-oriented or systems programming languages such as Java, C++, C#, or Go.
- Deep understanding of distributed-systems design, data structures, algorithms, operating systems, networking, and secure software development.
- Experience with system-level test automation, performance and load testing, reliability engineering, and production incident response.
- Experience leading or influencing technical architecture and mentoring engineers.
- Strong problem-solving, communication, and cross-functional collaboration skills.
- Preferred experience with Oracle Cloud, AWS, Azure, Google Cloud, data-plane platforms, distributed storage, microservices, replication, state management, high-throughput data processing, SLOs, observability, 24x7 production operations, Infrastructure as Code, service automation, security controls, and cloud compliance.
Benefits
- Office-based role requiring onsite presence in Nashville, Tennessee; relocation assistance may be available under Oracle policies.
- Medical, dental, vision, disability, life insurance, flexible spending accounts, commuter and parking benefits, and a 401(k) plan with company match.
- Paid vacation or accrued vacation, 11 paid holidays, paid sick leave, paid parental leave, and adoption assistance.
- Employee Stock Purchase Plan, financial planning, group legal benefits, and voluntary insurance benefits.
About Oracle
Oracle builds database technology, enterprise applications, and Oracle Cloud Infrastructure for companies and governments. Its business spans cloud subscriptions, software licenses, and support services across ERP, HCM, CX, and industry suites, plus NetSuite. Founded in 1977 and headquartered in Austin, Texas, Oracle is a public company traded on the NYSE and serves global customers migrating and running critical workloads in its cloud.
