Professionally recorded emotional speech corpora for TTS, voice cloning, expressive synthesis, and research.
Each voice includes 59–60 vocal effects such as moans, gasps, whimpers, screams, laughs, panting, breathing, inhales, exhales, crying and more.
Commercially Licensed Studio Data — Signed voice-actor agreements — Proof available on request
Every voice actor recorded the exact same manuscript: 1,360 phrases captured in six emotional states (Neutral, Angry, Happy, Scared, Shouting, Whisper) for a total of 8,160 clips per voice.
This shared phrase set ensures perfect alignment across all speakers — the same sentences, the same emotional coverage, and consistent structure across the entire corpus.
The corpus includes 1,344 structured phrases with questions, dramatic lines, pangrams, digits, written numbers, alphabet coverage, plus 16 long monologues.
Current datasets are available in English and Greek, and we are currently recording Norwegian. Additional languages such as French, Danish, Swedish, and other Scandinavian languages can be produced on demand. Contact us at
bedvibe@bedvibe.studio for custom language datasets.
All recordings are professionally produced in studio environments. BedVibe uses owned, contracted, or internally recorded voice data for commercial production. Signed agreements and consent documentation are maintained privately for formal partner review where applicable. Contact us for licensing and commercial use details.
Click a portrait to purchase a dataset. Use the emotion buttons below each speaker card to preview real samples.