Sr. Staff Cloud Engineer
Bloom Energy1 hour ago
San Jose, CA, USAStaff+
Base Salary
$156k - $224k/yr
Responsibilities
- Design, implement, maintain, and operate scalable AWS and hybrid cloud infrastructure for enterprise applications, production workloads, AI/data platforms, and CI/CD ecosystems.
- Architect highly available, resilient, secure, and cost-optimized cloud solutions and develop reusable infrastructure modules, platform standards, and engineering patterns.
- Lead Infrastructure-as-Code adoption using Terraform, CloudFormation, and related automation frameworks.
- Build observability capabilities with metrics, logs, traces, dashboards, and automated alerting, and drive SRE practices including SLOs, error budgets, incident management, and reliability improvement.
- Lead production readiness reviews, root cause analysis, performance optimization, resiliency planning, operational risk assessments, and critical-system incident escalation.
- Design and optimize CI/CD pipelines using DevOps, automation, and GitOps practices.
- Integrate security, compliance, governance, monitoring, patching, change management, disaster recovery, and production support controls into infrastructure workflows.
- Implement automated remediation, rollback, guardrails, and self-healing infrastructure patterns.
- Evaluate cloud, automation, AI/ML infrastructure, and platform engineering capabilities to support modernization and scalability.
- Mentor cloud, DevOps, and infrastructure engineers and lead large-scale cross-functional technical initiatives.
- Participate in on-call rotations for critical production systems.
Requirements
- Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related field.
- At least 10 years of experience in cloud, infrastructure, platform engineering, or DevOps, including 3 or more years in a senior, staff, principal, or technical leadership role.
- Deep hands-on expertise in AWS cloud services, cloud architecture, networking, security, resiliency, and production operations.
- Experience designing and operating mission-critical production environments supporting enterprise applications and services.
- Strong cloud networking experience involving VPC, Transit Gateway, Cloud WAN, Direct Connect, firewalls, routing, and BGP.
- Strong experience with Terraform, CloudFormation, CI/CD platforms, Kubernetes, containers, monitoring, and observability technologies.
- Strong scripting and automation skills using Python, Bash, PowerShell, or similar languages.
- Understanding of IAM, cloud security, compliance controls, encryption, resiliency engineering, operational governance, and SRE practices.
- Ability to lead large-scale cross-functional technical initiatives from strategy through execution.
- Master’s degree is preferred.
- Experience with Azure, Google Cloud, Oracle Cloud, or multi-cloud environments is strongly preferred.
- Experience with high-availability SaaS, AI/ML, manufacturing, energy, industrial technology, enterprise environments, FinOps, cloud cost optimization, AI/ML infrastructure, GPU workloads, data platforms, cloud governance, compliance automation, or internal developer platforms is preferred.
- Cloud or platform certifications such as AWS Solutions Architect Professional, AWS DevOps Professional, Kubernetes certifications, or Terraform certification are preferred.
Benefits
- Medical, dental, and vision plans with a large employer contribution.
- 401(k) retirement plan with company match.
- Mental health support services, legal services, virtual physical therapy access, and fertility and family-forming benefits.
- Full-time role based in San Jose, California, with participation in on-call rotations.
- Equal opportunity employer offering reasonable accommodations consistent with applicable law.
Tech Stack
AWSAzureBashDatadogGitHub ActionsGoogle CloudGrafanaJenkinsKubernetesOracle CloudPowerShellPrometheusPythonTerraform