5 hours ago
Remote, United StatesStaff+
Base Salary
$180k - $224k/yr
Responsibilities
- Own discovery, technical scoping, system design, implementation, and production rollout for strategic customer and ISV engagements.
- Build physical AI workflows covering real-world and synthetic data, model training, evaluation, deployment, and failure capture.
- Design scenario-based evaluations, regression tests, failure analysis, model comparisons, and real-to-sim-to-real improvement loops.
- Integrate customer datasets, infrastructure, codebases, simulation frameworks, robotics toolchains, and NVIDIA ecosystem tools.
- Develop reusable reference architectures, joint ISV solutions, and product components from customer prototypes.
- Use AI coding tools to rapidly prototype, test, debug, refactor, and ship production-quality software.
- Provide structured customer feedback to Product and Engineering and contribute technical blogs, solution templates, and reference architectures.
- Represent Nebius in customer technical sessions, partner co-builds, and Physical AI industry events.
Requirements
- 6+ years of hands-on engineering experience in applied ML, physical AI, computer vision, robotics, simulation, autonomy, or AI/ML platforms.
- At least two years in a customer-facing or deployment-oriented technical role such as Forward Deployed Engineer, founding engineer, technical co-founder, or embedded technical lead.
- Demonstrated experience building production data pipelines, training pipelines, evaluation harnesses, model versioning, deployment, and monitoring systems.
- Strong Python engineering skills and hands-on experience with PyTorch or similar ML frameworks.
- Practical expertise in evaluation, metrics, failure analysis, data quality, model improvement loops, sim-to-real gaps, domain shift, and production reliability.
- Experience with video, images, telemetry, annotations, simulation outputs, deployment logs, and other multimodal datasets.
- Fluency with Claude Code, Codex, and Cursor as AI-native development tools.
- Strong working knowledge of GPU compute, distributed training infrastructure, high-throughput storage, and orchestration frameworks such as Kubernetes, Ray, or Slurm.
- Ability to work directly in customer environments, navigate unfamiliar codebases and infrastructure, and communicate effectively with engineers and CTOs.
- Preferred experience includes robotics, drones, industrial automation, warehouse robotics, autonomous systems, synthetic data, scenario generation, simulation-based evaluation, world models, vision-language-action models, policy learning, reinforcement learning, representation learning, evaluation platforms, and relevant open-source projects.
Benefits
- 100% company-paid medical, dental, and vision insurance for employees and families.
- 401(k) plan with up to a 4% company match and immediate vesting.
- Paid parental leave of 20 weeks for primary caregivers and 12 weeks for secondary caregivers.
- Up to $85 per month in mobile and internet reimbursement for remote work.
- Company-paid short-term disability, long-term disability, and life insurance.
- Remote work is available from the United States, with the SF Bay Area or Austin, Texas preferred.
- Career growth and learning opportunities, flexibility, ownership, and an international work environment.
Tech Stack
Categories
Forward Deployed
About Nebius
Nebius builds a full-stack AI cloud offering GPU compute, storage, and tools for training and deploying ML models for startups, enterprises, and research labs. It sells consumption-based cloud infrastructure (IaaS/PaaS) and managed services tailored to generative AI workloads, including large-scale model training and inference. The company is headquartered in Amsterdam and operates as an independent provider.
