
Sold by: DefinedCrowd, Corp.
Deployed on AWS
This dataset contains English Scripted Monologue data, recorded from speakers in enUS.
Overview
This dataset contains English Scripted Monologue data, recorded from speakers in United States.
Packaging description A zip file containing metadata files in tsv format and a folder with all the audio files
- Domain Generic
- Total recordings 73,389
- File size 10.68GB
- Hours 99
- Word error rate (%) 1.1%
- Total prompts 73389
- Unique prompts 73389
- Average amount of recordings per speaker 266.87
Details
New
Introducing multi-product solutions
You can now purchase comprehensive solutions tailored to use cases and industries.
Features and programs
Financing for AWS Marketplace purchases
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing
This product is available free of charge. Free subscriptions have no end date and may be canceled any time.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
Vendor refund policy
Refunds are not available for this product.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA) .
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
Delivery details
AWS Data Exchange (ADX)
AWS Data Exchange is a service that helps AWS easily share and manage data entitlements from other organizations at scale.
Additional details
You will receive access to the following data sets.
Data set name | Type | Historical revisions | Future revisions | Sensitive information | Data dictionaries | Data samples |
|---|---|---|---|---|---|---|
English Speech Data - Scripted Monologue | All historical revisions | All future revisions | Not included | Not included |
Resources
Vendor resources
Similar products
Natural and expressive voices in multiple languages. For voice agents and brand ambassadors.
MiniMax Speech 02 is an advanced AI speech model capable of voice cloning and voice synthesis with high fidelity. It can replicate a speaker unique timbre and generate natural, expressive speech across different languages and styles.
Ranked #1 on the Artificial Analyze Text to Speech leaderboard, Speech 02 is designed for audio production, virtual assistants, call center and content creation, delivering realistic, customizable voice experiences at scale.
Deepgram is the leading voice AI platform for enterprise use cases, offering speech-to-text (STT), text-to-speech (TTS), and full speech-to-speech (STS) capabilities. 200,000+ developers build with Deepgrams voice-native foundational models due to our unmatched accuracy, low latency, and pricing. Having processed over 50,000 years of audio and transcribed over 1 trillion words, there is no organization in the world that understands voice better than Deepgram.
Fano Speech API (AWS Marketplace) provides Streaming and Async Speech-to-Text with low latency, multilingual code-switching (Cantonese/Mandarin/English), keyword biasing, automatic punctuation, speaker diarization, and webhook callbacks, built for production voice workflows.