Responsibilities
- Conduct research on Large Multimodal Models for Conversational Avatars.
- Develop methods to model verbal and non-verbal conversation aspects.
- Experiment with fine-tuning and conditioning techniques for AudioVisual Multimodal Models.
- Collaborate with the Applied ML team to transition research from prototype to production.
- Stay updated with advancements in the field and help define future directions.
Requirements
- PhD or near completion in a relevant field, or equivalent research experience.
- Hands-on experience with Large Multimodal Models and generative models.
- Experience in fine-tuning/adapting VLMs for control and downstream tasks.
- Solid background in deep learning and foundation models.
- Strong PyTorch skills and experience building deep learning pipelines.
Benefits
- Hybrid work model with preferred locations in San Francisco or London.
- Remote work options available for exceptional candidates.
Tech Stack
Categories
About Tavus
Tavus is a research lab pioneering human computing. We’re building AI humans: a new interface that closes the gap between us and machines, free from the friction of today’s systems. Our real-time human simulation models let machines see, hear, respond, and even look real, enabling meaningful face-to-face conversations with people. AI Humans connect and act with precision and empathy, making them capable, trusted agents. It’s the best of both worlds: the emotional intelligence of humans, with the reach and reliability of machines. They’re available 24/7, in every language, on our terms. Imagine a therapist that anyone can afford. A personal trainer that adapts to your schedule. A fleet of medical assistants that can give every patient the attention they need. Tavus: teaching machines how to be human.
