
Familiar
AI video translator providing real-time voice and lip-sync dubbing.

About Familiar
Familiar: World Models for Accurate Humans
Familiar is an AI-powered video translator and dubbing model that translates content into 25 languages while preserving the original voice, lip-sync, and background audio. It operates in real-time for livestreams and offers both API and self-serve dashboard access to help creators and platforms localize their content.
Key Features
- Unified Voice and Lip-Sync: Generates voice and lip-sync together in one model, ensuring the speaker's face accurately performs the new language.
- Real-Time Livestream Dubbing: Supports WebRTC and RTMP to dub livestreams in real time, pushing each language's feed to its own channel.
- Audio Preservation: Retains original background noise, music, laughter, and sound effects alongside the translated voice.
- Manual Review and Control: Allows teams to review and edit translations line-by-line before rendering, and supports Do Not Translate lists for catchphrases and names.
- API and Agent Integration: Offers an API for bulk requests and metadata translation, plus a Model Context Protocol (MCP) server for native integration with AI agents like Claude and OpenAI.
- 25 Supported Languages: Translates any-to-any between languages including English, Spanish, Mandarin, Hindi, Arabic, Japanese, and more.
Use Cases
- Live broadcasts and live commerce platforms like eBay, Whatnot, and Fanatics.
- Churches broadcasting sermons, services, and events.
- Travel agencies, medical clinics, law firms, and insurance companies for consults, tours, and client explainers.
- Podcasts, interviews, lectures, and straight-to-camera presentations.
Pricing
- $5 per finished minute per language, which includes translation, voice, scene audio, and lip-sync.
- Studio contracts are available with volume pricing and dedicated support.
Getting Started
Website: https://www.thefamiliarlab.com
Backed by Y Combinator, Familiar provides creators and enterprises with a highly accurate, cost-effective solution for globalizing video and live content without losing the original speaker's identity or scene atmosphere.