Chrisa — English Emotional Speech Dataset
8,158 clips · 3.50 hours · 6 emotions · 48 kHz, 32-bit float. €149 per dataset.
8,158 clips · 3.50 hours · 6 emotions · 48 kHz, 32-bit float. €149 per dataset.
8,158 clips · 4.56 hours · 6 emotions · 48 kHz, 32-bit float. €149 per dataset.
8,157 clips · 4.59 hours · 6 emotions · 48 kHz, 32-bit float. €149 per dataset.
8,158 clips · 4.50 hours · 6 emotions · 48 kHz, 32-bit float. €149 per dataset.
8,160 clips · 4.08 hours · 6 emotions · 48 kHz, 32-bit float. €149 per dataset.
8,159 clips · 4.02 hours · 6 emotions · 48 kHz, 32-bit float. €149 per dataset.
8,160 clips · 5.44 hours · 6 emotions · 48 kHz, 32-bit float. €149 per dataset.
57,110 clips · 30.69 hours · 6 emotions · 48 kHz, 32-bit float. €599 bundle.
Custom multilingual dataset production for TTS, voice cloning, expressive synthesis, and research. Additional speakers, emotions, languages, or scripted content recorded to specification. Custom quote — contact bedvibe@bedvibe.studio.
Every dataset package also includes 59–60 bonus vocal effects per speaker: gasps, moans, sneezes, yawns, screams, crying, laughing, panting, breathing, whimpers, and pain and effort vocalizations.
Full English Pack plus the right to distribute trained models, embed them in SDK, CLI or MCP software you ship, and run hosted inference for your own customers. €1,799 one-time price.
Everything in the commercial license plus the right to sublicense trained models to your own customers, white-label and OEM distribution, performer-rights indemnity, and speaker-disjoint expansion. €4,999 one-time price.
Original studio recordings created before October 2026: the 1,360-phrase script recorded in Greek in six emotional states. Commercial AI training, speech synthesis and dataset licensing rights are available subject to the applicable BedVibe licence terms. Pan is Panagiotis Gkilis, founder of BedVibe, the recorded speaker and owner of his datasets. Giouli recorded under a signed BedVibe performer agreement. Documentation of performer authorisation is available for private due-diligence review on request. Read the blank performer agreement (V4 template).
8,160 clips · 5.53 hours · 6 emotions · 48 kHz, 32-bit float. Contact for licensing.
8,154 clips · 4.75 hours · 6 emotions · 48 kHz, 32-bit float. Contact for licensing.
Pan recorded eleven languages, each in all six emotional states, so the speaker stays the same while language and emotion change. English and Greek are complete 1,360-phrase sets of 8,160 clips each. French, German, Japanese, Mandarin, Norwegian, Portuguese, Russian, Spanish and Swedish share a shorter phrase set, recorded in every emotion. Choose a language, then an emotion.
11 languages · 16,861 clips · 11.22 hours · 6 emotions in every language · 48 kHz, 32-bit float. Contact for licensing.
Further English voices recorded on the same 1,360-phrase script. Emotional coverage differs from voice to voice: each card lists the states recorded and plays one real sample from each. Licensing on request.
6,134 clips · 2.17 hours · 6 emotions · 48 kHz, 32-bit float. Contact for licensing.
4,032 clips · 1.19 hours · Neutral, Happy, Whisper · 48 kHz, 32-bit float. Contact for licensing.
1,937 clips · 0.72 hours · Neutral, Angry, Happy · 48 kHz, 32-bit float. Contact for licensing.
1,440 clips · 0.72 hours · Neutral, Happy · 48 kHz, 32-bit float. Contact for licensing.
1,409 clips · 0.67 hours · Neutral, Happy · 48 kHz, 32-bit float. Contact for licensing.
1,345 clips · 0.56 hours · Neutral · 48 kHz, 32-bit float. Contact for licensing.
1,176 clips · 0.46 hours · Neutral · 48 kHz, 32-bit float. Contact for licensing.
489 clips · 0.22 hours · 6 emotions · 48 kHz, 32-bit float. Contact for licensing.
Four ways to buy: €149 per speaker, €599 for all seven, €1,799 commercial, €4,999 enterprise with sublicensing. Every price is a one-time cost for a perpetual license — no subscription, no per-use fee, no royalty. If you only need to train a model and sell what you generate with it, the €149 or €599 purchase is enough. If you need to distribute the model, embed it in software you ship, or run hosted inference for your own customers, buy the commercial license. If your own customers need rights of their own, buy the enterprise license.
Dataset licenses cover the AI training, fine-tuning, synthesis and commercial-use rights stated for each license tier.
Performer personal information is not part of the dataset license. Where a performer is presented publicly by a first name, stage name or other agreed public name, that name is what the dataset supplies. The license does not grant access to, or rights over, undisclosed legal names, surnames, contact information, addresses, private biographical information, photographs or endorsements.
The source recordings themselves may not be redistributed except where expressly permitted by a separate written agreement.
Ask and we send, the same day: the full license text for any tier, and a sample of the
per-clip manifest — audio_file, text, speaker, emotion, language,
gender — with duration distributions per emotion, so you can check same-text
neutral/emotional alignment, performer uniqueness and clip counts before spending anything.
Each of the seven published English sets includes 59–60 non-verbal vocalizations — breathing, panting, gasps, whimpers, screams, crying, laughter, yawns, sneezes, and pain and effort vocalizations.
Beyond the published English sets we hold roughly thirty voices in total. Recorded hours and coverage of emotions and vocalizations vary from voice to voice, and we supply the exact hours, emotional states and material available for any voice on request. Some of the newer recordings were made on a Neumann U87, and new sessions can be recorded to your specification.