Strategically involved in conversation Al dataset development for many years, MagicData hasdesigned and produced the Speech E2E Translation Dataset.This dataset originates from rea human natura conversations. where lanquace expressions arenatural, diverse, and exhibit individual characteristics. The emotional expressions are naturalallowing machines to learn human natural expressions effectively. The dataset supports not onlytraditional speech to text-to-text (S2T) translation but also speech-to-speech (S2S) translation.
Language
CN-EN
Style
Conversational & Scripted
Sampling Rate
16kHz
Bit Rate
16bits
ISO/IEC 27001 & ISO/IEC 27701:2019 compliant
Audio, text, image, and video multi-modal data
Conversational, scripted, and spontaneous data covering extensive domains
Expertise secured quality result