Robots and Pencils

Staff Platform Engineer

Robots and Pencils
Apply
1 day ago

Base Salary

$126k - $174k/yr

Responsibilities

  • Define DevOps strategy and infrastructure architecture across multi-environment, multi-region cloud systems.
  • Architect and own scalable Kubernetes platforms and containerized infrastructure.
  • Establish infrastructure-as-code standards and lead DevSecOps practices including secrets management, compliance, auditing, IAM, and zero-trust networking.
  • Drive platform reliability, performance SLAs, observability, cost optimization, cloud migrations, and modernization initiatives.
  • Design and operate AI/ML platform infrastructure including model serving, GPU workload orchestration, LLM gateways, vector stores, and AI/ML deployment pipelines.
  • Partner with engineering, product, and leadership; lead design reviews, architecture discussions, and release-readiness assessments.
  • Establish platform standards, mentor junior and mid-level engineers, act as a technical escalation point, and evaluate emerging technologies.

Requirements

  • 7+ years of professional DevOps or platform engineering experience, including leadership of complex platform initiatives.
  • Expert scripting and programming skills in technologies such as Python, Go, Java, or Bash.
  • Deep expertise in at least one major cloud platform, Kubernetes, container orchestration, infrastructure as code, and CI/CD architecture at scale.
  • Strong DevSecOps, secrets management, compliance, auditing, networking, IAM, security architecture, and zero-trust experience.
  • Experience with service mesh, distributed systems, microservices architecture, and AI/ML platform infrastructure.
  • Demonstrated technical leadership, mentoring, stakeholder communication, and ability to navigate ambiguous technical and business challenges.
  • Demonstrable daily use and expert knowledge of AI-forward tools such as Claude and Cursor.
  • Cloud certifications such as AWS DevOps Engineer Professional or CKA, and FinOps experience, are preferred.
  • Experience designing and provisioning HPC cluster infrastructure across AWS, CoreWeave, GCP, and OCI is helpful.
  • Experience with HPC job schedulers and workload managers such as Slurm is helpful.

Benefits

  • Remote role based in the US.
  • Salary range of $126,000 USD to $174,000 USD.
  • Comprehensive benefits package including paid time off, medical, dental, vision insurance, and 401(k) for eligible employees.
  • Employment may be conditional on successful completion of a background check.
Robots and Pencils

About Robots and Pencils

51-200 employees
Contact me