Tavus 发布 Griffin,称其为首个通过视频图灵测试的人类交互模型(HIM),48% 与之实时对话的人认为它是真人,此前系统通过率低于 3%。Griffin 在 NVIDIA 全双工 AI 视频基准上得 3.83 分(真人为 3.92),次优系统为 2.80 分;它将听、想、说和面部动画整合在单一系统中,可在对话者仍在说话时读取表情、语气并实时反应。
A company built a model good enough to pass as a person.
48% of people who talked to Griffin live (a human interaction model from Tavus) thought it was a real human, I was lucky to get an early access and can definitely agree to that.
On NVIDIA's scoreboard for face-to-face AI, a real human scores 3.92 out of 5.
Griffin scores 3.83. The next best system scores 2.80.
Earlier systems stitched separate models together to hear, think, speak and animate, and every handoff added delay. Griffin does all of it in one system, reading your face and tone as well as your words and reacting while you're still talking.
Introducing Griffin, the first model to pass the video Turing test. 48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video. It’s the first Human Interaction Model (HIM).在 X 查看被引用的帖子
来源:Rohan Paul · x.com