Radar / AI media tools / Tavus Griffin
Tavus Griffin — real-time video Turing-test avatar
Tavus's Griffin is a full-duplex "Human Interaction Model" that generates a conversational person — face, voice, body and background — frame by frame from a single reference image at 0.43s average latency.
Why it matters
First system to claim a real-time video Turing test (26 of 54 participants judged it human vs 1 of 41 for Tavus's previous stack), and NVIDIA independently scored it 3.83/5 on VideoFDB generation against a 3.92 human reference — perception, decision and generation run concurrently, not chained pipelines.
What you could build with it
A telehealth clinic could staff an always-on intake avatar that greets patients on video, reads facial cues and collects symptom histories overnight, with every flagged answer routed to a human nurse for review before any clinical use.
Does it hold up?
Too early to judge for builders: only the Griffin-Lite research preview exists for a small group of trusted testers, with no API, pricing or public release date; the 48% figure comes from Tavus's own 54-person study, though NVIDIA's independent VideoFDB scoring corroborates the generation quality.
Built with Tavus Griffin
- Frontier Now: Tavus launches Griffin, the full-duplex Human Interaction Modelfrontiernow.dev · Independent write-up: Griffin-Lite scored 3.83 on NVIDIA VideoFDB generation and 3.73 on perception (the highest on both leaderboards); only a trusted-tester research preview exists, with wider release gated on disclosure and safety work.
- Yomimono: Tavus Griffin AI fools 48% of video callers into thinking it is humanyomimono.id · Editorially reviewed summary of the GIGAZINE reporting: 720p video generated from one still image at sub-second latency, with the system split into a Continuous Conversational Modeling engine and an Audio-Visual Generation engine.
- StyleYourItems: Tavus unveils Griffin AI video model that fools 48% of usersstyleyouritems.com · Breakdown of the full-duplex architecture: Griffin reassesses conversation state at sub-second intervals, allowing it to nod, backchannel mid-sentence and stop the moment a user interrupts — versus relay-style avatar pipelines.
Learn more
First spotted on article: source.
More AI media tools
OpenMontage: open-source agentic video production systemTurns an AI coding assistant into a full video studio: 12 pipelines and 100+ tools for scripting, footage, narration…media · JEV 0.73LingBot-Map: streaming 3D reconstruction foundation model (ECCV 2026 best-paper candidate)Feed-forward 3D foundation model that reconstructs scenes from streaming video at ~20 FPS with drift correction over…media · JEV 0.7HyperFramesHeyGen's open-source framework that turns HTML/CSS/GSAP into deterministic MP4 video via 21 agent skills and a router…media · JEV 0.69ArtCraft pivots to seven open-source Adobe clones built with Claude Opus 5.5Brandon Thomas pivoted his AI art IDE into a seven-app open-source creative suite…media · JEV 0.67Meshy crosses $100M ARR and launches official iOS/Android appsMeshy announced Sep 30 that ARR surpassed $100M (up 100x in under two years, first AI-3D company at that milestone)…media · JEV 0.61Kandinsky 6.0 Video: open MIT-licensed audio-video generation modelsSber's team released Kandinsky 6.0 Video: Lite (3B) and Pro (29B) diffusion models generating 5-second clips with…media · JEV 0.59
Get the week's best AI launches, plus 3 ideas worth building
One email every Saturday. Ranked by traction, not hype. Free.