en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

593 Hours - English(China) Scripted Monologue Smartphone speech dataset

Chinese speak English
English voice data
and mobile phones collect voice data.china
chinaman
oriental
taiwanese
byzantine
chink
chinese
woman
formosan
asian
chinawoman
japanese
korean
sinaean
celestial
chine
chino
chugoku
non-chinese
ovary
pekin
acupuncture
airframe
airframes
all-china
beijing
patois
vernacular
language
lingo
speech
idiom
argot
tongue
accent
jargon
slang
parlance
cant
patter
brogue
provincialism
localism
regional
language
locution
terminology
vocabulary
local
speech
regionalism
pidgin
regionalisms
colloquialism
phraseology
idioms
mother
tongue
talk
pronunciation
local
language
creole
langue
lingua
franca
phrasing
tongues
idiolect
lexicon
vernacularism
colloquial
word
brogues
dialectal
jive
talk
languages
lingua
localisms
wording
accents
business
language

English(China) Scripted Monologue Smartphone speech dataset, collected from monologue based on given scripts, covering 100,000 common expressions. Transcribed with text content and other attributes. Our dataset was collected from extensive and diversify speakers(3,691 Chinese, covering domestic dialect zones like Jiangsu, Shandong, Beijing, He'nan, and meets the specific accents of Chinese speaking English), geographicly speaking, enhancing model performance in real and complex tasks.Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Paid Datasets
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
SpecificationsSpecifications
Format
16kHz, 16bit, uncompressed wav, mono channel;
Recording condition
Low background noise (indoor), without echo;
Content category
100,000 common expressions;
Recording device
Android smartphone;
Speaker
3,691 Chinese, 34% male and 66% female;
Country
China(CHN);
Language
English;
Features of annotation
Transcription text;
Accuracy Rate
Sentence Accuracy Rate (SAR) 95%
Sample Sample
  • Audio

    Her cheeks had fallen in,making her look old.

  • Audio

    No milk. I'm slimming.

  • Audio

    We're focused on small things: Do I have my pierce?

  • Audio

    No one could know why he did like that.

  • Audio

    The bark scaled off the tree.

Recommended DatasetsRecommended Dataset
Tell Us Your Special Needs

By submitting, I agree to the Privacy Protection

3a4ea118-7f41-4942-846b-e02382f9bc70

0ac30917-68af-471e-8bbc-7810b051600a