RIVR

AI Intern – Vision-Language-Action (VLA) & Data

RIVR
Apply
4 days ago
Zürich, SwitzerlandIntern

Responsibilities

  • Process, curate, and analyze multimodal sensor data for training Vision-Language-Action models.
  • Develop software tools to visualize data and debug model performance.
  • Work with senior engineers on data strategies that improve robotic system robustness.
  • Integrate software components to evaluate algorithms in simulation and on hardware.
  • Learn about and contribute to VLA, self-supervised learning, and generative AI developments.

Requirements

  • At least a BSc in computer science, robotics, machine learning, or a related field.
  • Proficiency in Python and experience with deep learning frameworks, preferably PyTorch.
  • Hands-on deep learning for computer vision experience through coursework, internships, or projects.
  • Familiarity with NumPy, Pandas, and OpenCV.
  • Strong problem-solving skills and willingness to work with complex real-world data.
  • Bonus qualifications include an MSc or ongoing PhD, large-scale image or video dataset experience, transformer or vision-language model knowledge, VLA experience, 3D geometry, camera projections, sensor fusion, robotics, and motion or action prediction.

Benefits

  • In-person work at RIVR office locations is required.
  • Hands-on experience with state-of-the-art VLA models and autonomous robotic systems.
  • Exposure to self-supervised learning, generative AI, simulation, and physical robot hardware.
  • Collaborative learning environment within RIVR, part of Amazon.
  • Applications are limited to Schengen Area citizens, except ETHZ and EPFL students completing compulsory internships.

Tech Stack

NumPyOpenCVPandasPythonPyTorch

Categories

RIVR

About RIVR

51-200 employees
Contact me