3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement
Siqi Zheng, Luyao Cheng, Yafeng Chen +2 authors
A large-scale speech corpus, 3D-Speaker, facilitates research on disentangling speech representations by varying speakers, recording devices, distances, and dialects, enabling the evaluation of universal speech models and out-of-domain learning.