Lambda

Senior Platform Engineer - Core Infrastructure

Lambda
Apply
2 months ago
San Francisco, CA, USA or San Jose, CA, USASenior
H1B Sponsor

Base Salary

$230k - $340k/yr

Responsibilities

  • Architect, deploy, and operate Kubernetes clusters across AWS and Lambda’s bare-metal datacenters
  • Build and maintain automation for Kubernetes cluster provisioning, upgrades, scaling, and lifecycle management
  • Own the reliability, performance, and security of production Kubernetes workloads
  • Implement observability, logging, and alerting for clusters and critical workloads
  • Partner with product teams to design scalable cloud-native services and CI/CD pipelines
  • Set platform standards for resource management, networking, and RBAC
  • Lead incident response, root-cause analysis, and post-mortems for platform issues
  • Mentor engineers and raise the standard for platform engineering across the organization

Requirements

  • At least 5 years of experience in platform, infrastructure, or SRE roles, including running Kubernetes in production at scale
  • Deep knowledge of Kubernetes internals and day-2 operations such as upgrades, scaling, and troubleshooting
  • Strong experience with Helm, Kustomize, or similar tools and GitOps-based delivery
  • Proficiency with infrastructure-as-code using Terraform, Pulumi, or equivalent
  • Solid knowledge of networking, service meshes, and container runtimes
  • Hands-on experience with Prometheus, Grafana, and OpenTelemetry observability stacks
  • Strong coding skills in Go or Python for automation and tooling
  • Practical security experience with network policies, secrets management, and image scanning
  • Preferred experience with multi-cluster, multi-cloud, or hybrid environments
  • Preferred knowledge of GPU scheduling, HPC workloads, or ML/AI infrastructure
  • Preferred experience with Temporal, Cadence, or Argo Workflows
  • Preferred exposure to cost optimization and capacity planning for large clusters
  • Contributions to CNCF or Kubernetes open-source projects are a plus
  • CKA or CKS certification is a plus

Benefits

  • Hybrid schedule requiring presence in the San Francisco, San Jose, or Seattle office 4 days per week, with Tuesday designated as the work-from-home day
  • Generous cash and equity compensation
  • Health, dental, and vision coverage for employees and dependents
  • Wellness and commuter stipends for select roles
  • 401(k) plan with a 2% company match for U.S. employees
  • Flexible paid time off plan

Tech Stack

AWSGoGrafanaHelmKubernetesPrometheusPythonTerraform

Categories

DevOpsSite Reliability
Lambda

About Lambda

501-1,000 employees

The Superintelligence Cloud

Contact me