5 months ago
Remote, United States or Remote, EMEAMid Level / Senior
H1B sponsor
Responsibilities
- Conduct research on Large Multimodal Models for Conversational Avatars.
- Develop methods to model verbal and non-verbal conversation aspects.
- Experiment with fine-tuning and conditioning techniques for AudioVisual Multimodal Models.
- Collaborate with the Applied ML team to transition research from prototype to production.
- Stay updated with advancements in the field and help define future directions.
Requirements
- PhD or near completion in a relevant field, or equivalent research experience.
- Hands-on experience with Large Multimodal Models and generative models.
- Experience in fine-tuning/adapting VLMs for control and downstream tasks.
- Solid background in deep learning and foundation models.
- Strong PyTorch skills and experience building deep learning pipelines.
Benefits
- Hybrid work model with preferred locations in San Francisco or London.
- Remote work options available for exceptional candidates.
Tech Stack
Categories
Data Science
About Tavus
Tavus builds generative AI technology and APIs that let developers and enterprises create real-time, face-to-face conversational agents and digital-twin video experiences. Its platform exposes speech, vision, and video synthesis models for embedding into apps for customer engagement, healthcare assistants, training, and other interactive use cases, sold as usage-based developer services. Founded in 2020 and headquartered in San Francisco, the company is privately held.
