
Cantina
Cantina Labs is a social AI company, developing a suite of advanced real-time models that push the boundaries of expression, personality, and realism. We bring characters to life, transforming how people tell stories, connect, and create. We build and power ecosystems. Cantina, our flagship social AI platform, is just the beginning.
Open Positions at Cantina
10 open positions
Build the real-time speech and media infrastructure behind Cantina’s AI conversations, using high-performance C++ and WebRTC across mobile and web platforms. This Senior-Staff-level role tackles low-latency audio, distributed systems, speech technologies, and immersive voice and video experiences.
Build and ship state-of-the-art voice conversion and speech models at Cantina, spanning research, data, evaluation, and production inference. This role combines frontier audio ML with hands-on systems engineering, responsible AI, and large-scale deployment.
Join Cantina Labs as a Machine Learning Engineer specializing in speech and audio generation, focusing on joint audio-video modeling to create advanced AI systems.
Build and scale production inference infrastructure for Cantina’s generative audio models, supporting reliable, low-latency streaming and batch workloads. You’ll bridge research and production through Kubernetes-based deployment, MLOps automation, GPU optimization, and observability.
Build and scale backend services powering Cantina’s social AI platform, with a focus on acquisition, engagement, retention, and scalable product growth. This staff-level role combines hands-on backend engineering with significant influence over architecture and technical infrastructure.
Build polished, high-performance Android experiences for Cantina’s social AI platform, spanning custom Compose interfaces, media pipelines, real-time effects, and AI-powered features. This senior-to-staff role offers substantial ownership over architecture, product direction, and flagship-app experiences.
Build the systems that turn large-scale video and multimodal data into high-quality, training-ready datasets for Cantina’s real-time social AI models. You’ll own distributed pipelines, curation workflows, cloud infrastructure, and data quality improvements from raw content through model training.
Build and ship cutting-edge TTS, voice-cloning, and related speech models while owning the full research-to-production lifecycle. You’ll combine ML research, large-scale data and GPU systems, rigorous evaluation, and responsible AI practices.
Join Cantina Labs as a founding Machine Learning Engineer focused on evaluating speech generation and recognition models. You’ll build scalable evaluation pipelines, metrics, user studies, and dashboards while shaping the company’s future evaluation team.
Build visually rich, high-performance iOS experiences for Cantina’s social AI platform, spanning personalized feeds, custom interfaces, AI-powered media, and real-time effects. You’ll own major features, contribute to architecture, and help shape an app used by a large audience.