TensorX
返回文献探索

Paper · arXiv 2408.05366

DeepSpeak Dataset v1.0

Sarah Barrington, Matyas Bohacek, Hany Farid

14 upvotesAugust 9, 2024arXiv 预印本
AI 摘要
deepfakeface-swaplip-sync

Abstract

We describe a large-scale dataset--{\em DeepSpeak}--of real and deepfake footage of people talking and gesturing in front of their webcams. The real videos in this first version of the dataset consist of 9 hours of footage from 220 diverse individuals. Constituting more than 25 hours of footage, the fake videos consist of a range of different state-of-the-art face-swap and lip-sync deepfakes with natural and AI-generated voices. We expect to release future versions of this dataset with different and updated deepfake technologies. This dataset is made freely available for research and non-commercial uses; requests for commercial use will be considered.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号