1 day ago
Base Salary
$129k - $185k/yr
Responsibilities
- Deploy, operate, maintain, monitor, and troubleshoot resilient AWS and Kubernetes infrastructure and microservices.
- Respond to complex production incidents in a 24x7 environment and help restore services quickly.
- Automate operational work and infrastructure provisioning to improve deployment safety, efficiency, and consistency.
- Support identity and access management, logging, hardening, vulnerability remediation, and FedRAMP continuous monitoring.
- Collaborate with engineering, security, and compliance teams to identify operational risks and improve reliability.
- Create runbooks, troubleshooting guides, operational procedures, documentation, and incident reports.
- Participate in on-call rotations, root-cause analysis, post-incident reviews, and Incident Commander responsibilities.
Requirements
- Bachelor’s degree in Computer Science, engineering, or a related field plus 5+ years of related experience, or equivalent practical experience.
- 5+ years of software development or automation experience using Java, Go, Python, or a comparable programming language.
- Experience with testing, troubleshooting, and maintaining code.
- Experience supporting AWS cloud infrastructure, production services, DevOps, platform engineering, or Site Reliability Engineering.
- Experience with Linux administration and troubleshooting application, system, or networking issues in production or production-like environments.
- Experience with Kubernetes, Docker, microservices, cloud-native application deployment and troubleshooting, Git, CI/CD pipelines, and infrastructure-as-code or automation.
- Preferred experience with FedRAMP, government cloud, or other regulated environments.
- Preferred experience with Prometheus, Grafana, CloudWatch, CloudTrail, Elastic Stack, Splunk, GitLab, or Jenkins.
- Preferred experience with highly available multi-region distributed systems, capacity planning, disaster recovery, on-call operations, incident management, and audit documentation.
Benefits
- Hybrid work in RTP, North Carolina, or Boxborough, Massachusetts, with 2–3 days in the office.
- Medical, dental, and vision insurance; 401(k) with Cisco matching contribution; paid parental leave; disability coverage; and basic life insurance.
- Potential eligibility for Cisco restricted stock unit grants and annual bonuses for non-sales roles.
- Paid holidays, floating holidays, birthday leave, year-end shutdown, personal wellness days, vacation or flexible vacation time, sick time, family emergency leave, and optional volunteer days.
- The role requires work that the U.S. government has specified can only be performed by a U.S. citizen on U.S. soil.
Categories
DevOpsSite Reliability
About Cisco
Cisco designs and sells networking, security, and collaboration platforms for enterprises, service providers, and governments, spanning routers and switches, Wi‑Fi, firewalls, zero‑trust, observability, and cloud-managed IT (Meraki) plus Webex. Its business model mixes hardware, software subscriptions, and support/consulting services. Founded in 1984 and headquartered in San Jose, California, Cisco is a public company traded on Nasdaq and serves customers across data centers, campuses, and service provider networks.
