Reka

Member of Technical Staff (GPU Performance Engineer)

Reka
Apply
9 months ago
Remote, United States +5 moreSenior

Responsibilities

  • Design and implement improvements to model training infrastructure.
  • Contribute to technical decisions that optimize model performance.
  • Work on post-training processes, including reinforcement learning and fine-tuning.
  • Improve the efficiency and scalability of model serving infrastructure.
  • Profile GPU-accelerated workloads, identify bottlenecks, and implement performance optimizations.

Requirements

  • Strong engineering skills with fluency in Python and PyTorch or other frameworks.
  • Proven experience implementing and training large deep learning models.
  • Experience writing and debugging low-level GPU code using CUDA and C++.
  • Experience scaling GPU jobs with large-scale compute clusters such as Slurm or Kubernetes.
  • Ability to analyze and optimize GPU-accelerated workloads through profiling, bottleneck identification, and performance tuning.

Benefits

  • Remote-first work arrangement with a globally distributed team
  • Five weeks of paid leave
  • Comprehensive healthcare benefits including vision and dental
  • Additional well-being perks
  • Visa assistance, including H1B and OPT transfers, for US employees
Reka

About Reka

51-200 employees

Reka is a privately held AI research and product company building multimodal foundation models and tools for organizations and businesses. Headquartered in Sunnyvale with a remote-first team, it offers products such as Reka Vision and Reka Edge, a 7B vision-language model. Founded by researchers from Google DeepMind and Meta’s FAIR, it also operates Claru, a data unit supplying licensed training datasets for physical and embodied AI.

Contact me