
DevOps Engineer
Mind Robotics4 hours ago
Palo Alto, CA, USAMid Level
Responsibilities
- Design, deploy, and maintain scalable cloud infrastructure using AWS, GCP, or Azure.
- Build and maintain Infrastructure as Code using Terraform or similar tools.
- Develop CI/CD pipelines for robotics software, embedded applications, and machine learning workflows.
- Automate software deployment across development robots, test environments, and production fleets.
- Improve build systems, artifact management, release processes, and version-control workflows.
- Manage Kubernetes clusters and containerized services supporting robotics applications.
- Monitor infrastructure health, application performance, and production systems using observability tools.
- Improve system reliability, uptime, security, and disaster-recovery processes.
- Build internal developer tooling to increase engineering productivity.
- Partner with robotics, firmware, and ML teams to create reproducible development environments.
- Support simulation infrastructure and large-scale testing pipelines.
- Optimize cloud resource utilization and infrastructure costs.
- Implement infrastructure security practices for secrets management, networking, and access controls.
- Participate in incident response, root-cause analysis, and continuous operational improvements.
Requirements
- At least four years of experience in DevOps, Platform Engineering, Site Reliability Engineering, or Infrastructure Engineering.
- Strong experience administering Linux systems and troubleshooting production systems under operational constraints.
- Experience with AWS, GCP, or Azure, as well as Kubernetes and Docker in production environments.
- Strong Infrastructure as Code experience, with Terraform preferred.
- Experience building CI/CD pipelines with GitHub Actions, GitLab CI, Jenkins, Buildkite, or similar platforms.
- Experience with configuration management and automation tools.
- Proficiency with Python, Go, or Bash for infrastructure automation.
- Experience implementing monitoring and logging solutions such as Prometheus, Grafana, Datadog, ELK, or OpenTelemetry.
- Strong understanding of networking, security, authentication, and distributed systems.
- Nice-to-have experience supporting robotics, autonomous systems, embedded software teams, physical-device fleets, GPU infrastructure, ML training clusters, simulation environments, edge computing, or IoT deployments.
- Nice-to-have experience with NVIDIA GPUs, CUDA environments, distributed training systems, artifact repositories, large binary assets, software supply-chain security, and SBOMs.
Tech Stack
AWSAzureBashBuildkiteDatadogDockerGitHub ActionsGitLab CI/CDGoGoogle Cloud PlatformGrafanaJenkinsKubernetesLinuxPrometheusPythonTerraform
Categories
About Mind Robotics
Mind Robotics is building generalizable robotics for real-world industrial deployment. We believe the fastest path to broadly capable robots is to work on clearly defined, high impact problems. That's why we're starting where the need is most acute, and the environment is most exacting: the factory floor, in partnership with Rivian. We are building a lean, collaborative team where every individual has a massive impact. Researchers and engineers work hand in hand with hardware every day. We’re looking for builders who are passionate about robotics and live for the challenge of making systems work in the real world. Explore opportunities at https://www.mindrobotics.com/