# DF26 基准：现有检测手段已难以分辨 AI 生成视频与真实视频

- 来源：HuggingFace Daily Papers（社区热门论文）
- 发布时间：2026-09-07 08:00
- AIHOT 分数：54
- AIHOT 链接：https://aihot.news/items/cmtv4ahpr0oy6rorpiqhpfkxi
- 原文链接：https://arxiv.org/abs/2609.07369

## AI 摘要

DF26 基准针对单人物公开演讲场景评测 AI 生成视频检测，包含 271 个真实视频和由七个现代视频模型生成的 2,420 个合成视频。结果显示人类和最先进的 deepfake 检测器准确率都接近随机水平，暴露出当前评测协议的局限，并提出了对现代生成模型分布偏移保持鲁棒性的基准需求。

## 正文

We introduce DF26, a novel benchmark for detecting AI-generated videos containing fully synthetic clips produced by recent text-to-video and image-to-video models. The videos capture single-person public-speaking scenarios, spanning direct-to-camera recordings, official statements, and studio interviews - 271 real and 2,420 synthetic videos generated by seven modern video models. The study on DF26 shows that human performance in detecting AI-generated videos, as well as state-of-the-art deepfake detectors, is close to random chance. Our results highlight the limitations of current evaluation protocols and motivate the need for benchmarks that explicitly measure robustness to modern generative model distribution shifts.
