vp-AB2-k3

vprof_vc voice profiles, rule AB2, k=3, T=0.25 C=0.25. One profile is one cloned voice by construction, so there is no speaker-identity risk; there is also no time axis, so the order is chosen rather than observed.

Rule. AB2 — two-sided: emotion A falls by >=T while emotion B rises by >=T, each consecutive step <=C
Source. vprof_vc (mined by gridsel/vpgrid, PROVISIONAL)  |  Family. voice-profile grid (vprof_vc): one cloned voice, chains CONSTRUCTED not discovered
How to read a Script. Each chunk is one line: a short tag of what the models heard in that clip, then the words spoken.

(underlined, plain · delivery, style) — the tag before the words. Emotions first, then how it is delivered. Underlined descriptors are the ones that change across this chain — anything identical on every clip is pulled out and stated once above, because a value that never moves says nothing about a trajectory.
(ahem) — brackets inside the words are a different thing: a real non-speech sound, printed where it happens. Most clips have none; about a quarter do.

The full generated caption for any clip is still there, under “full caption & clip details”. Its perceived-gender and background-noise clauses were re-rendered from the numeric buckets, because the versions stored in the corpus index had those two ladders running backwards.

The emotion clause has been re-derived, and it used to be wrong. The 40 emotion heads are not on a common scale — Interest has a median of 2.08 and is never zero, while Infatuation is zero on 87.7 % of clips — and the caption named an emotion whenever its raw score cleared an absolute 1.0. Interest therefore appeared in 94.8 % of captions and Sadness in almost none: the clause was reporting the scale of the head, not the emotion of the clip. An emotion is now named only when it lands in the top 10 % for that emotion, against a pooled tie-aware mid-rank ECDF over 132,833,726 utterances spanning every dataset and language. Interest now appears in 6.2 %, all 40 emotions occur, and a clip that is ordinary on all 40 says “no dominant emotion” rather than being forced to pick one (21.8 % of clips). This is the same scale the trajectory miner selects on, so the caption and the mining now refer to the same quantity: the mined target emotion is named in the final clip's caption on 73 % of chains, up from 46 %.
Affection ↓  /  Astonishment Surprisevp-AB2-k3 · #1

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Astonishment Surprise up — by at least 0.25 each.

The chain starts with Astonishment Surprise around average — 0.52, higher than 52 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.48.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.23 — an even, steady climb — each clip carries about the same share.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 22 s · en · vprof_vc

hear it un-normalised (raw levels, max seam 1.8 dB)
k 3d_a -0.404d_b 0.475step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker anime_016track anime_016total 21.9slevel spread 2.3 dBmax seam 1.8 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: an adult masculine voice · fairly smooth, good recording, no background noise
(affection, longing, sadness · measured, normally alert, slightly relaxed, narration) I cannot bear to see you walk away, my heart aches at the thought. Four shots, perhaps, is the only way I know to keep you near me forever.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as affection, longing, sadness; style: narration, formal; good recording, no background noise; genuineness 0.6/6; vocal-burst blend 2.4/10; 6.4s, EN.
anime_016__E__Affection__B__en.c027.k1 · in -22.5 dBFS · gain +2.5 dB · vprof_vc-00000
(disgust, contempt, sourness · normal-paced, normally alert, slightly relaxed, monologue) Im Ernst, mit so einer Tinder-artigen Wisch-App Kontakte knüpfen? Das ist einfach... ekelhaft. Warum sollte sich irgendjemand freiwillig so einen riesigen Mist antun?
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; very clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as disgust, contempt, sourness; style: monologue, narration; good recording, no background noise; genuineness 0.9/6; vocal-burst blend 2.2/10; 6.1s, DE.
anime_016__E__Disgust__D__de.c025.k3 · in -24.2 dBFS · gain +4.2 dB · vprof_vc-00000
(astonishment surprise, awe, distress · slow, very low-energy, relaxed, conversational) No way you're telling me that happened like that, for real? That sounds totally wild, dude.
full caption & clip details
An adult masculine voice; delivery is very low-energy, slow, relaxed, moderately variable; timbre is slightly warm, dark, fairly smooth, very thin; somewhat unclear, little disfluency, wide pitch range, normal breath; affect is negative, submissive, vulnerable; reads as astonishment surprise, awe, distress; style: conversational, casual; good recording, no background noise; genuineness 1.4/6; vocal-burst blend 2.6/10; 9.1s, EN.
anime_016__B__hiss__en.c016.k2 · in -24.7 dBFS · gain +4.7 dB · vprof_vc-00000
Affection ↓  /  Contemplationvp-AB2-k3 · #2

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Contemplation up — by at least 0.25 each.

The chain starts with Contemplation clearly present — 0.70, higher than 70 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.30.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.05 — most of the change happening immediately, then levelling off.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 43 s · en · vprof_vc

hear it un-normalised (raw levels, max seam 0.0 dB)
k 3d_a -0.404d_b 0.296step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker anime_016track anime_016total 42.6slevel spread 0.0 dBmax seam 0.0 dB
Script — 3 chunks, 2 with a non-speech sound
Unchanged across all 3 clips: a young adult masculine voice · neutral-bright, balanced body, frequent disfluency
(affection, longing, helplessness · slow, lethargic, relaxed, casual) He feels like love and being loved, sharing and giving it, somehow fueled the cancer. (coughing) It's heartbreaking to think that something so beautiful could be twisted like this.
full caption & clip details
A young adult masculine voice; delivery is lethargic, slow, relaxed, moderately variable; timbre is slightly warm, neutral-bright, slightly rough, balanced body; somewhat unclear, frequent disfluency, narrow pitch range, normal breath; affect is mildly negative, submissive, neutral openness; reads as affection, longing, helplessness; style: casual; below-average recording, quiet background; genuineness 1.4/6; vocal-burst blend 3.6/10; 14.8s, EN.
anime_016__X__ga_sad_cry__en.c039.k3 · in -24.6 dBFS · gain +4.6 dB · vprof_vc-00000
(awe, jealousy and envy, infatuation · slow, very low-energy, neutral tension) (contented sigh) The divine spark inside us demands to shine forth, doesn't it? Just look at it, the sheer brilliance we were meant to reveal.
full caption & clip details
An adult masculine voice; delivery is very low-energy, slow, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, balanced body; somewhat unclear, frequent disfluency, wide pitch range, normal breath; affect is negative, neutral stance, vulnerable; reads as awe, jealousy and envy, infatuation; below-average recording, quiet background; mildly explicit content; genuineness 0.9/6; vocal-burst blend 0.4/10; 12.1s, EN.
anime_016__B__hiss__en.c001.k2 · in -24.6 dBFS · gain +4.6 dB · vprof_vc-00000
(contemplation, doubt, concentration · measured, subdued, slightly relaxed, monologue) How does an individual actually become? What about the beginnings of developmental biology, or the cell theory's deep impact? (low mumble) Hmm, developmental genetics and evolution… let's look at those sections nine through two-twenty-seven.
full caption & clip details
An adult masculine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, frequent disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as contemplation, doubt, concentration; style: monologue; good recording, no background noise; genuineness 1.1/6; vocal-burst blend 2.0/10; 15.4s, EN.
anime_016__B__lip_smack__en.c021.k0 · in -24.6 dBFS · gain +4.6 dB · vprof_vc-00000
Affection ↓  /  Contemplationvp-AB2-k3 · #3

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Contemplation up — by at least 0.25 each.

The chain starts with Contemplation clearly present — 0.64, higher than 64 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.36.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.75 (higher than 75 % of clips in this corpus), a change of -0.25. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.11 — a fairly even climb, though some clips carry more of the change than others.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 27 s · en · vprof_vc

hear it un-normalised (raw levels, max seam 2.2 dB)
k 3d_a -0.252d_b 0.363step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker anime_088track anime_088total 26.6slevel spread 3.6 dBmax seam 2.2 dB
Script — 3 chunks, 1 with a non-speech sound
Unchanged across all 3 clips: an adult masculine voice · neutral-bright, fairly smooth, no background noise, slightly relaxed, moderate pitch range
(affection, jealousy and envy, disappointment · slow, subdued, moderately variable, conversational) (contented sigh) She just doesn't know what happened to her mom. My mom, who always had the perfect thing to say, whose laugh could lift you up, she was the heart of this whole family.
full caption & clip details
An adult masculine voice; delivery is subdued, slow, slightly relaxed, moderately variable; timbre is slightly warm, neutral-bright, fairly smooth, full; average clarity, almost no disfluency, moderate pitch range, normal breath; affect is mildly negative, slightly submissive, fairly guarded; reads as affection, jealousy and envy, disappointment; style: conversational; good recording, no background noise; mildly explicit content; genuineness 0.8/6; vocal-burst blend 6.9/10; 8.4s, EN.
anime_088__V__FOCS__very_high__en.c020.k3 · in -25.6 dBFS · gain +5.6 dB · vprof_vc-00008
(awe, contentment, relief · slow, very low-energy, moderately variable, monologue) Here, the mountains meet the sea, running right out to the water. (smack one s lips) The Great Ocean Road snakes along them in a long, winding path.
full caption & clip details
An adult masculine voice; delivery is very low-energy, slow, slightly relaxed, moderately variable; timbre is slightly warm, neutral-bright, fairly smooth, very thin; clear, little disfluency, moderate pitch range, minimal breath; affect is mildly negative, neutral stance, fairly guarded; reads as awe, contentment, relief; style: monologue, ASMR; good recording, no background noise; mildly explicit content; genuineness 0.7/6; vocal-burst blend 2.9/10; 8.9s, EN.
anime_088__V__EXPL__moderately_high__en.c039.k1 · in -23.4 dBFS · gain +3.4 dB · vprof_vc-00008
(contemplation, awe, longing · normal-paced, normally alert, fairly steady, narration) Manchmal reicht es schon, jemanden durch ein Fenster zu beobachten, um einen Einblick in seine wahren Gedanken zu bekommen. Es ist wie ein Fenster in seine innere Welt.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as contemplation, awe, longing; style: narration, formal; very good recording, no background noise; genuineness 0.4/6; vocal-burst blend 5.7/10; 9.1s, DE.
anime_088__V__DARC__extremely_low__de.c047.k2 · in -22.0 dBFS · gain +2.0 dB · vprof_vc-00008
Affection ↓  /  Contemptvp-AB2-k3 · #4

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Contempt up — by at least 0.25 each.

The chain starts with Contempt clearly present — 0.61, higher than 61 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.39.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.14 — a fairly even climb, though some clips carry more of the change than others.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 33 s · de · vprof_vc

hear it un-normalised (raw levels, max seam 0.4 dB)
k 3d_a -0.404d_b 0.387step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang despeaker anime_088track anime_088total 33.0slevel spread 0.5 dBmax seam 0.4 dB
Script — 3 chunks, 1 with a non-speech sound
Unchanged across all 3 clips: an adult masculine voice · neutral-bright, fairly smooth, good recording, measured
(affection, contentment, thankfulness gratitude · very low-energy, slightly relaxed, fairly steady, monologue) Durch kleine Schritte mit viel Liebe und Geduld blüht der Maltipoo zu einem toleranten Familienfreund auf. (drinking noises) Das lässt ihn zu einem so wunderbaren, vertrauensvollen Begleiter werden.
full caption & clip details
An adult masculine voice; delivery is very low-energy, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as affection, contentment, thankfulness gratitude; style: monologue, casual; good recording, quiet background; genuineness 1.6/6; vocal-burst blend 7.6/10; 11.3s, DE.
anime_088__V__R_MIXD__very_high__de.c012.k3 · in -25.0 dBFS · gain +5.0 dB · vprof_vc-00008
(disgust, sexual lust, teasing · very low-energy, fully relaxed, moderately variable, casual) (contented sigh) A pastry made with a very light yeast dough, topped with cream and various grated or crumbled cheeses. (person whistling to get attention) It's a sweet and savory combination you have to try.
full caption & clip details
An adult masculine voice; delivery is very low-energy, measured, fully relaxed, moderately variable; timbre is slightly warm, neutral-bright, fairly smooth, full; clear, little disfluency, fairly narrow pitch, minimal breath; affect is mildly negative, slightly submissive, neutral openness; reads as disgust, sexual lust, teasing; style: casual, whispered; good recording, no background noise; mildly explicit content; genuineness 0.8/6; vocal-burst blend 0.6/10; 9.7s, EN.
anime_088__V__EMPH__extremely_low__en.c006.k3 · in -24.6 dBFS · gain +4.6 dB · vprof_vc-00008
(contempt, anger, malevolence malice · normally alert, slightly relaxed, fairly steady, storytelling) Du wirst weder deine Bohnen, noch dein Abendessen, noch deinen Darm ruinieren, indem du einen Weg statt des anderen wählst. (nervous giggle) Ehrlich gesagt, macht es zwischen diesen beiden Ansätzen keinen wirklichen Unterschied.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is slightly warm, neutral-bright, fairly smooth, very full; clear, no disfluency, moderate pitch range, light breath; affect is negative, neutral stance, fairly guarded; reads as contempt, anger, malevolence malice; style: storytelling, narration; good recording, no background noise; genuineness 0.7/6; vocal-burst blend 5.6/10; 11.8s, DE.
anime_088__V__REGS__extremely_low__de.c010.k2 · in -24.5 dBFS · gain +4.5 dB · vprof_vc-00008
Affection ↓  /  Disgustvp-AB2-k3 · #5

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Disgust up — by at least 0.25 each.

The chain starts with Disgust clearly present — 0.65, higher than 65 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.35.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.10 — a fairly even climb, though some clips carry more of the change than others.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 24 s · en · vprof_vc

hear it un-normalised (raw levels, max seam 0.8 dB)
k 3d_a -0.404d_b 0.349step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0039track emolia_c0039total 24.5slevel spread 0.8 dBmax seam 0.8 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, no background noise, moderate pitch range
(affection, contentment, jealousy and envy · normal-paced, normally alert, slightly relaxed, narration) Even with Augustus' illness, even knowing how brief this joy is, I feel such profound peace just being with him. (guffaw) Our love, I know that will never fade.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, full; average clarity, little disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as affection, contentment, jealousy and envy; style: narration, conversational; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 2.7/10; 6.7s, EN.
emolia_c0039__E__Relief__C__en.c037.k0 · in -22.9 dBFS · gain +2.9 dB · vprof_vc-00016
(longing, pain, sadness · measured, normally alert, slightly relaxed) I will live on the edge of this precarious cliff, watching the town below. Let the performer truly live, pouring emotion into every line.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, little disfluency, moderate pitch range, minimal breath; affect is mildly negative, slightly dominant, fairly guarded; reads as longing, pain, sadness; good recording, no background noise; genuineness 0.7/6; vocal-burst blend 0.9/10; 8.7s, EN.
emolia_c0039__V__AGEV__moderately_low__en.c030.k2 · in -23.6 dBFS · gain +3.6 dB · vprof_vc-00016
(disgust, shame, embarrassment · normal-paced, very low-energy, neutral tension, playful) Honestly, I feel so ashamed because I know viruses can linger on keyboards and doorknobs for ages. (gulps) It's just... deeply gross.
full caption & clip details
A young adult masculine voice; delivery is very low-energy, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, slightly thin; average clarity, some disfluency, moderate pitch range, normal breath; affect is mildly negative, slightly dominant, slightly guarded; reads as disgust, shame, embarrassment; style: playful, conversational; average recording, no background noise; genuineness 0.4/6; vocal-burst blend 2.0/10; 8.8s, EN.
emolia_c0039__E__Shame__C__en.c010.k2 · in -23.4 dBFS · gain +3.4 dB · vprof_vc-00016
Affection ↓  /  Fatigue Exhaustionvp-AB2-k3 · #6

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Fatigue Exhaustion up — by at least 0.25 each.

The chain starts with Fatigue Exhaustion clearly present — 0.62, higher than 62 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.38.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.13 — a fairly even climb, though some clips carry more of the change than others.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 22 s · en · vprof_vc

k 3d_a -0.404d_b 0.380step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0039track emolia_c0039total 21.7slevel spread 2.1 dBmax seam 1.3 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, good recording, no background noise, normally alert, slightly relaxed, wide pitch range
(affection, infatuation, pain · measured, moderately variable, almost no disfluency, narration) Perhaps this isn't love when I say you are dearest to me; love is knowing you are the blade I use to cut into my own soul.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, almost no disfluency, wide pitch range, minimal breath; affect is mildly negative, slightly dominant, fairly guarded; reads as affection, infatuation, pain; style: narration; good recording, no background noise; explicit content; genuineness 0.7/6; vocal-burst blend 1.2/10; 6.9s, EN.
emolia_c0039__V__ARSH__very_high__en.c025.k1 · in -22.8 dBFS · gain +2.8 dB · vprof_vc-00016
(longing, shame, triumph · normal-paced, moderately variable, almost no disfluency) I wish I could just reach it right now, though knowing it takes some know-how. With these simple steps, I hope I can finally access what I crave.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, almost no disfluency, wide pitch range, light breath; affect is mildly negative, slightly dominant, fairly guarded; reads as longing, shame, triumph; good recording, no background noise; mildly explicit content; genuineness 0.4/6; vocal-burst blend 0.7/10; 7.7s, EN.
emolia_c0039__E__Longing__C__en.c020.k1 · in -23.6 dBFS · gain +3.5 dB · vprof_vc-00016
(fatigue exhaustion, disgust, helplessness · measured, fairly steady, no disfluency, ranting) Mann, ich bin gerade total fertig, echt ausgelaugt. Dieser ganze Tag war ein ziemlicher Krampf, wie ewig.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; very clear, no disfluency, wide pitch range, light breath; affect is neutral, slightly dominant, fairly guarded; reads as fatigue exhaustion, disgust, helplessness; style: ranting, authoritative; good recording, no background noise; genuineness 1.3/6; vocal-burst blend 3.3/10; 6.8s, DE.
emolia_c0039__E__Sourness__B__de.c024.k1 · in -24.9 dBFS · gain +4.8 dB · vprof_vc-00016
Affection ↓  /  Angervp-AB2-k3 · #7

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Anger up — by at least 0.25 each.

The chain starts with Anger clearly present — 0.60, higher than 60 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.40.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.75 (higher than 75 % of clips in this corpus), a change of -0.25. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.15 — a fairly even climb, though some clips carry more of the change than others.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 34 s · en · vprof_vc

k 3d_a -0.252d_b 0.397step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0238track emolia_c0238total 33.8slevel spread 1.5 dBmax seam 1.5 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: a young adult feminine voice · neutral-bright
(affection, contentment, pleasure ecstasy · normal-paced, normally alert, slightly relaxed, conversational) It's just... they are so loving, so gentle, and so utterly sweet. It’s almost too much to ask for perfection, and sometimes it's just heartbreaking.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as affection, contentment, pleasure ecstasy; style: conversational, narration; good recording, no background noise; mildly explicit content; genuineness 1.5/6; vocal-burst blend 2.6/10; 8.7s, EN.
emolia_c0238__E__Disappointment__D__en.c029.k0 · in -25.1 dBFS · gain +5.2 dB · vprof_vc-00024
(fear, sourness, sadness · normal-paced, normally alert, slightly relaxed, formal) It's just so poignant, that in thirteen percent of these cases, even with all the supportive care, this progressive course of Wegener's granulomatosis continues. It weighs heavily on my heart to see that persistence.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as fear, sourness, sadness; style: formal, narration; good recording, no background noise; mildly explicit content; genuineness 0.1/6; vocal-burst blend 0.0/10; 10.4s, EN.
emolia_c0238__E__Affection__C__en.c029.k0 · in -25.1 dBFS · gain +5.1 dB · vprof_vc-00024
(anger, contempt, disgust · slow, very low-energy, relaxed) Seriously, some student thinks this punishment, and the whole sexist dress code mess, is totally unfair and is now speaking up about it. (growl) You gotta laugh at that.
full caption & clip details
A young adult feminine voice; delivery is very low-energy, slow, relaxed, moderately variable; timbre is slightly warm, neutral-bright, slightly rough, slightly thin; slurred, frequent disfluency, wide pitch range, minimal breath; affect is mildly negative, submissive, vulnerable; reads as anger, contempt, disgust; below-average recording, quiet background; mildly explicit content; genuineness 0.6/6; vocal-burst blend 0.0/10; 14.4s, EN.
emolia_c0238__X__amused_laughter__en.c003.k2 · in -26.6 dBFS · gain +6.6 dB · vprof_vc-00024
Affection ↓  /  Astonishment Surprisevp-AB2-k3 · #8

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Astonishment Surprise up — by at least 0.25 each.

The chain starts with Astonishment Surprise around average — 0.52, higher than 52 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.48.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.23 — an even, steady climb — each clip carries about the same share.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 30 s · en · vprof_vc

k 3d_a -0.404d_b 0.475step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0238track emolia_c0238total 29.5slevel spread 1.2 dBmax seam 1.2 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(affection, contentment, infatuation · steady, little disfluency, clear, ASMR) It's just... they are so loving, so gentle, and so utterly sweet. It’s almost too much to ask for perfection, and sometimes it's just heartbreaking.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as affection, contentment, infatuation; style: ASMR, narration; good recording, no background noise; genuineness 1.1/6; vocal-burst blend 2.7/10; 8.7s, EN.
emolia_c0238__E__Disappointment__D__en.c029.k2 · in -25.2 dBFS · gain +5.2 dB · vprof_vc-00024
(doubt, impatience and irritability, sourness · fairly steady, little disfluency, average clarity, narration) Moment, also Kunden mit wirklich kritischen aws Internet of Things Core Sachen können dieses Ding nutzen, um ihre eigenen Daten in einem zweiten aws-Bereich zu behalten und damit zu arbeiten, falls der Hauptbereich von ihren Geräten nicht erreichbar ist? (growl) Ich bin mir nicht ganz sicher, wie das funktioniert.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as doubt, impatience and irritability, sourness; style: narration, storytelling; good recording, no background noise; mildly explicit content; genuineness 1.3/6; vocal-burst blend 4.0/10; 12.9s, DE.
emolia_c0238__E__Confusion__A__de.c002.k2 · in -26.2 dBFS · gain +6.2 dB · vprof_vc-00024
(astonishment surprise, fear, distress · fairly steady, no disfluency, clear, narration) My god, this truck just rammed into a Ford that was going faster, and then that whole thing slammed into a bmw. (heavy breathing) I cannot believe how chaotic this whole mess turned out.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as astonishment surprise, fear, distress; style: narration, formal; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.4/10; 7.7s, EN.
emolia_c0238__E__Distress__B__en.c038.k0 · in -25.1 dBFS · gain +5.0 dB · vprof_vc-00024
Affection ↓  /  Fearvp-AB2-k3 · #9

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Fear up — by at least 0.25 each.

The chain starts with Fear clearly present — 0.70, higher than 70 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.30.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.74 (higher than 74 % of clips in this corpus), a change of -0.26. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.05 — most of the change happening immediately, then levelling off.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 19 s · en · vprof_vc

k 3d_a -0.255d_b 0.297step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0323track emolia_c0323total 19.0slevel spread 1.1 dBmax seam 1.1 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: an adult feminine voice · neutral-toned, fairly smooth, balanced body, good recording, no background noise, slightly relaxed, light breath
(affection, infatuation, malevolence malice · brisk, energised, moderately variable, conversational) Still, I wouldn't hesitate to fight him for you. (spitting) I would do anything to have your love.
full caption & clip details
An adult feminine voice; delivery is energised, brisk, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, no disfluency, wide pitch range, light breath; affect is positive, slightly submissive, neutral openness; reads as affection, infatuation, malevolence malice; style: conversational, storytelling; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 1.4/10; 5.2s, EN.
emolia_c0323__V__ATCK__very_high__en.c038.k3 · in -25.0 dBFS · gain +5.0 dB · vprof_vc-00032
(longing, pleasure ecstasy, infatuation · normal-paced, normally alert, fairly steady, casual) Those wireless phones, like little devices hooked to the routers, crave that wlan connection so badly. (chuckle) I want to sink into that connection, a sweet, deep surge.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as longing, pleasure ecstasy, infatuation; style: casual, monologue; good recording, no background noise; genuineness 1.6/6; vocal-burst blend 0.0/10; 7.7s, EN.
emolia_c0323__E__Sexual_Lust__C__en.c010.k2 · in -25.0 dBFS · gain +5.0 dB · vprof_vc-00032
(fear, distress, helplessness · brisk, normally alert, fairly steady, conversational) I'm so scared that if they don't try, if they don't put in the work, they'll never reach even the smallest thing they've dreamed of. That possibility terrifies me.
full caption & clip details
A young adult feminine voice; delivery is normally alert, brisk, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as fear, distress, helplessness; style: conversational, authoritative; good recording, no background noise; genuineness 1.3/6; vocal-burst blend 0.8/10; 5.9s, EN.
emolia_c0323__E__Fear__B__en.c044.k1 · in -23.9 dBFS · gain +3.9 dB · vprof_vc-00032
Amusement ↓  /  Angervp-AB2-k3 · #10

This chain comes from the two-sided rule: it only counts if both emotions move — Amusement down and Anger up — by at least 0.25 each.

The chain starts with Anger clearly present — 0.70, higher than 70 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.30.

At the same time Amusement goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.71 (higher than 71 % of clips in this corpus), a change of -0.29. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.05 — most of the change happening immediately, then levelling off.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 25 s · en · vprof_vc

k 3d_a -0.292d_b 0.297step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0323track emolia_c0323total 24.7slevel spread 3.1 dBmax seam 2.5 dB
Script — 3 chunks, 1 with a non-speech sound
Unchanged across all 3 clips: a young adult feminine voice · fairly smooth, balanced body, good recording, no background noise, moderately variable, wide pitch range
(amusement, pleasure ecstasy, disgust · slow, very low-energy, neutral tension, casual) Seriously, this whole thing is proper dodgy innit. It's like the absolute worst, proper rubbish, innit. (surprised gasp)
full caption & clip details
A young adult feminine voice; delivery is very low-energy, slow, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, wide pitch range, normal breath; affect is positive, slightly submissive, vulnerable; reads as amusement, pleasure ecstasy, disgust; style: casual, playful; good recording, no background noise; mildly explicit content; genuineness 1.8/6; vocal-burst blend 1.3/10; 11.5s, EN.
emolia_c0323__V__AROU__moderately_high__en.c012.k1 · in -27.0 dBFS · gain +7.0 dB · vprof_vc-00032
(impatience and irritability, contempt, disgust · normal-paced, normally alert, neutral tension, casual) Honestly, stop banking on it for your bachelor's or master's thesis; you need substance, not just that. (yawn) It's completely insufficient as a main pillar.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, neutral openness; reads as impatience and irritability, contempt, disgust; style: casual, conversational; good recording, no background noise; mildly explicit content; genuineness 2.6/6; vocal-burst blend 0.6/10; 7.2s, EN.
emolia_c0323__E__Impatience_and_Irritability__C__en.c037.k1 · in -26.4 dBFS · gain +6.4 dB · vprof_vc-00032
(anger, disgust, contempt · brisk, energised, slightly tense, cartoonish) Hold your tongue, fiend, before the Almighty. You shall not stand here to bring charges against His chosen servant.
full caption & clip details
An adult feminine voice; delivery is energised, brisk, slightly tense, moderately variable; timbre is slightly cool, slightly bright, fairly smooth, balanced body; clear, no disfluency, wide pitch range, light breath; affect is negative, slightly dominant, fairly guarded; reads as anger, disgust, contempt; style: cartoonish, ranting; good recording, no background noise; genuineness 0.5/6; vocal-burst blend 1.7/10; 5.7s, EN.
emolia_c0323__V__AGEV__very_high__en.c029.k3 · in -23.9 dBFS · gain +3.9 dB · vprof_vc-00032
Affection ↓  /  Astonishment Surprisevp-AB2-k3 · #11

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Astonishment Surprise up — by at least 0.25 each.

The chain starts with Astonishment Surprise around average — 0.52, higher than 52 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.48.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.23 — an even, steady climb — each clip carries about the same share.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 22 s · en · vprof_vc

k 3d_a -0.404d_b 0.475step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0382track emolia_c0382total 22.1slevel spread 1.3 dBmax seam 1.3 dB
Script — 3 chunks, 1 with a non-speech sound
Unchanged across all 3 clips: an adult feminine voice · neutral-toned, fairly smooth, full, good recording, no background noise, moderately variable, no disfluency, clear
(affection, contentment, pleasure ecstasy · brisk, energised, slightly relaxed, storytelling) I'd love for you to join us, sweetheart. The IMDb rating plugin is just for our registered community members, and we cherish having you here with us.
full caption & clip details
An adult feminine voice; delivery is energised, brisk, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, full; clear, no disfluency, wide pitch range, minimal breath; affect is positive, slightly dominant, neutral openness; reads as affection, contentment, pleasure ecstasy; style: storytelling, narration; good recording, no background noise; genuineness 0.3/6; vocal-burst blend 0.1/10; 8.1s, EN.
emolia_c0382__E__Affection__A__en.c026.k2 · in -21.0 dBFS · gain +1.0 dB · vprof_vc-00040
(relief, fear, helplessness · measured, normally alert, slightly relaxed, monologue) (breathy giggle) When you break down mediation, it’s so satisfying because there's always that opening, the substance, and then a clear wrap-up. (breathy giggle) That predictable structure just brings such a calm sense of completion.
full caption & clip details
An adult feminine voice; delivery is normally alert, measured, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, no disfluency, wide pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as relief, fear, helplessness; style: monologue, narration; good recording, no background noise; mildly explicit content; genuineness 0.7/6; vocal-burst blend 0.4/10; 7.7s, EN.
emolia_c0382__E__Contentment__A__en.c033.k2 · in -21.6 dBFS · gain +1.6 dB · vprof_vc-00040
(astonishment surprise, awe, impatience and irritability · normal-paced, normally alert, neutral tension, storytelling) No way, you're seriously saying that happened? Like, no fucking way that actually went down, dude. That's wild, for real.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, no disfluency, moderate pitch range, light breath; affect is mildly negative, slightly dominant, neutral openness; reads as astonishment surprise, awe, impatience and irritability; style: storytelling, narration; good recording, no background noise; genuineness 0.9/6; vocal-burst blend 4.3/10; 6.1s, EN.
emolia_c0382__C__undead__en.c025.k1 · in -20.4 dBFS · gain +0.4 dB · vprof_vc-00040
Affection ↓  /  Fatigue Exhaustionvp-AB2-k3 · #12

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Fatigue Exhaustion up — by at least 0.25 each.

The chain starts with Fatigue Exhaustion around average — 0.57, higher than 57 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.43.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.24, then +0.19 — an even, steady climb — each clip carries about the same share.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 25 s · en · vprof_vc

k 3d_a -0.404d_b 0.429step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0382track emolia_c0382total 25.4slevel spread 1.2 dBmax seam 0.7 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: an adult feminine voice · neutral-toned, neutral-bright, fairly smooth, good recording, no background noise, normal-paced, normally alert, slightly relaxed
(affection, contentment, pleasure ecstasy · no disfluency, formal, monologue) They are so sweet, truly docile and cheerful, but I wish they were a bit more... lively. Still, their purrs and cuddles make them the beloved, peaceful members of our family.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, no disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as affection, contentment, pleasure ecstasy; style: formal, monologue; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.8/10; 8.0s, EN.
emolia_c0382__E__Disappointment__D__en.c029.k3 · in -20.3 dBFS · gain +0.3 dB · vprof_vc-00040
(doubt, fear, confusion · little disfluency, conversational, formal) Ich bin mir nicht ganz sicher, wie ich anfangen soll, aber ich denke, wir sollten uns zuerst den Winkelhalbierenden-Satz ansehen. Danach können wir es vielleicht tatsächlich beweisen.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as doubt, fear, confusion; style: conversational, formal; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 1.7/10; 8.9s, DE.
emolia_c0382__E__Doubt__B__de.c043.k3 · in -21.1 dBFS · gain +1.1 dB · vprof_vc-00040
(fatigue exhaustion, pain, helplessness · little disfluency, conversational, ASMR) Sometimes, after using this for too long, I start getting headaches and can't sleep. (fast breathing) I've also felt palpitations, vision issues, and that awful nasal swelling.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as fatigue exhaustion, pain, helplessness; style: conversational, ASMR; good recording, no background noise; genuineness 0.9/6; vocal-burst blend 1.2/10; 8.3s, EN.
emolia_c0382__E__Distress__A__en.c008.k1 · in -21.6 dBFS · gain +1.6 dB · vprof_vc-00040
Affection ↓  /  Angervp-AB2-k3 · #13

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Anger up — by at least 0.25 each.

The chain starts with Anger clearly present — 0.67, higher than 67 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.33.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.08 — most of the change happening immediately, then levelling off.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 23 s · en · vprof_vc

k 3d_a -0.404d_b 0.333step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0645track emolia_c0645total 22.9slevel spread 1.0 dBmax seam 0.6 dB
Script — 3 chunks, 1 with a non-speech sound
Unchanged across all 3 clips: an adult somewhat feminine voice · moderately variable, wide pitch range
(affection, infatuation, contentment · slow, very low-energy, neutral tension) (contented sigh) Marriage, a blessed arrangement, a dream come true, truly. (heavy breathing) And love, true love, it will follow you forever. So treasure your love, every single moment.
full caption & clip details
An adult somewhat feminine voice; delivery is very low-energy, slow, neutral tension, moderately variable; timbre is slightly warm, neutral-bright, slightly rough, full; somewhat unclear, little disfluency, wide pitch range, minimal breath; affect is negative, submissive, vulnerable; reads as affection, infatuation, contentment; below-average recording, quiet background; genuineness 1.4/6; vocal-burst blend 0.3/10; 10.8s, EN.
emolia_c0645__C__seasoned-merchant__en.c013.k0 · in -24.6 dBFS · gain +4.6 dB · vprof_vc-00048
(relief, contentment, pleasure ecstasy · brisk, energised, slightly relaxed, playful) I am so ready for my Feierabend, I just want to kick off my shoes and relax. It is time to finally unwind after this whole Tag.
full caption & clip details
A young adult masculine voice; delivery is energised, brisk, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; clear, no disfluency, wide pitch range, light breath; affect is positive, slightly dominant, neutral openness; reads as relief, contentment, pleasure ecstasy; style: playful, ranting; good recording, no background noise; genuineness 0.7/6; vocal-burst blend 0.7/10; 6.5s, EN.
emolia_c0645__E__Contemplation__B__en.c030.k3 · in -24.1 dBFS · gain +4.1 dB · vprof_vc-00048
(anger, impatience and irritability, contempt · brisk, energised, slightly tense, playful) How dare you think they'll just magically accept them? (yawn) After this ordeal, they'll see what you've done.
full caption & clip details
An adult feminine voice; delivery is energised, brisk, slightly tense, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, full; average clarity, no disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, fairly guarded; reads as anger, impatience and irritability, contempt; style: playful, dramatic; good recording, no background noise; genuineness 0.9/6; vocal-burst blend 0.6/10; 5.3s, EN.
emolia_c0645__E__Anger__A__en.c038.k1 · in -23.6 dBFS · gain +3.6 dB · vprof_vc-00048
Affection ↓  /  Contemplationvp-AB2-k3 · #14

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Contemplation up — by at least 0.25 each.

The chain starts with Contemplation clearly present — 0.69, higher than 69 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.31.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.06 — most of the change happening immediately, then levelling off.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 33 s · en · vprof_vc

k 3d_a -0.404d_b 0.309step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0645track emolia_c0645total 32.6slevel spread 0.6 dBmax seam 0.3 dB
Script — 3 chunks, 2 with a non-speech sound
Unchanged across all 3 clips: an adult somewhat feminine voice · neutral-bright, neutral tension, moderately variable, wide pitch range
(affection, infatuation, contentment · slow, lethargic, little disfluency) (contented sigh) Marriage, a blessed arrangement, a dream come true, truly. (heavy breathing) And love, true love, it will follow you forever. So treasure your love, every single moment.
full caption & clip details
An adult somewhat feminine voice; delivery is lethargic, slow, neutral tension, moderately variable; timbre is slightly warm, neutral-bright, slightly rough, full; somewhat unclear, little disfluency, wide pitch range, minimal breath; affect is negative, submissive, vulnerable; reads as affection, infatuation, contentment; below-average recording, quiet background; genuineness 1.4/6; vocal-burst blend 0.4/10; 10.8s, EN.
emolia_c0645__C__seasoned-merchant__en.c013.k2 · in -24.5 dBFS · gain +4.5 dB · vprof_vc-00048
(jealousy and envy, awe, helplessness · slow, very low-energy, some disfluency, casual) (contented sigh) Oh, those clasped hands in a dream... it just screams of turmoil for the Imam, or the very leader of this land. Such knotted fingers speak of deep, gnawing complications in their world.
full caption & clip details
A young adult feminine voice; delivery is very low-energy, slow, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, slightly thin; clear, some disfluency, wide pitch range, normal breath; affect is negative, neutral stance, vulnerable; reads as jealousy and envy, awe, helplessness; style: casual; average recording, quiet background; genuineness 0.1/6; vocal-burst blend 1.0/10; 14.1s, EN.
emolia_c0645__X__pain_groan__en.c022.k2 · in -24.3 dBFS · gain +4.3 dB · vprof_vc-00048
(contemplation, doubt, distress · brisk, energised, almost no disfluency, monologue) After all this time, can I truly feel anything new from Him? (yawn) Or is it just the same old, hollow ache inside?
full caption & clip details
A young adult masculine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, wide pitch range, light breath; affect is negative, slightly dominant, vulnerable; reads as contemplation, doubt, distress; style: monologue; good recording, no background noise; genuineness 0.5/6; vocal-burst blend 0.3/10; 7.5s, EN.
emolia_c0645__E__Bitterness__C__en.c006.k0 · in -23.9 dBFS · gain +3.9 dB · vprof_vc-00048
Affection ↓  /  Embarrassmentvp-AB2-k3 · #15

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Embarrassment up — by at least 0.25 each.

The chain starts with Embarrassment clearly present — 0.72, higher than 72 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.28.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.75 (higher than 75 % of clips in this corpus), a change of -0.25. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.03 — most of the change happening immediately, then levelling off.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 24 s · en · vprof_vc

k 3d_a -0.252d_b 0.279step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0758track emolia_c0758total 23.5slevel spread 2.6 dBmax seam 2.6 dB
Script — 3 chunks, 1 with a non-speech sound
Unchanged across all 3 clips: a young adult feminine voice · neutral-toned, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert, average clarity
(affection, pleasure ecstasy, contentment · slightly relaxed, fairly steady, little disfluency, casual) Seeing her smile when she gets something she truly wants just fills my heart. (breathy giggle) It feels like the simplest joy to bring her happiness.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as affection, pleasure ecstasy, contentment; style: casual, conversational; good recording, no background noise; mildly explicit content; genuineness 2.6/6; vocal-burst blend 3.7/10; 6.4s, EN.
emolia_c0758__E__Affection__A__en.c010.k1 · in -24.0 dBFS · gain +4.0 dB · vprof_vc-00056
(elation, pleasure ecstasy, hope enthusiasm optimism · slightly relaxed, moderately variable, some disfluency, conversational) (ahem) I'm absolutely thrilled! (breathy (breathy giggle) giggle) We sifted through so many bathing shoes, past the overpriced junk and terrible advice, to find the absolute best value. Our testers only picked the winners for our look selection!
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as elation, pleasure ecstasy, hope enthusiasm optimism; style: conversational, casual; good recording, no background noise; genuineness 3.4/6; vocal-burst blend 2.5/10; 9.6s, EN.
emolia_c0758__E__Elation__D__en.c013.k2 · in -21.4 dBFS · gain +1.4 dB · vprof_vc-00056
(embarrassment, shame, infatuation · neutral tension, fairly steady, little disfluency, conversational) I am so embarrassed; my face feels like it's burning right now. All those loud sighs and groans made me feel completely exposed.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, neutral tension, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as embarrassment, shame, infatuation; style: conversational, casual; good recording, no background noise; genuineness 2.6/6; vocal-burst blend 1.8/10; 7.3s, EN.
emolia_c0758__E__Embarrassment__D__en.c031.k2 · in -23.7 dBFS · gain +3.7 dB · vprof_vc-00056
Affection ↓  /  Emotional Numbnessvp-AB2-k3 · #16

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Emotional Numbness up — by at least 0.25 each.

The chain starts with Emotional Numbness clearly present — 0.63, higher than 63 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.36.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.11 — a fairly even climb, though some clips carry more of the change than others.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 34 s · en · vprof_vc

k 3d_a -0.404d_b 0.364step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0758track emolia_c0758total 34.0slevel spread 5.4 dBmax seam 5.4 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: a young adult feminine voice · neutral-toned, fairly smooth, balanced body, good recording, no background noise, light breath
(affection, infatuation, longing · brisk, normally alert, slightly relaxed, conversational) I know and love you so much, and if you flutter away, my sparkle will fade completely. Then all the brightest bits of my world would just disappear!
full caption & clip details
A young adult feminine voice; delivery is normally alert, brisk, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, little disfluency, wide pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as affection, infatuation, longing; style: conversational, storytelling; good recording, no background noise; genuineness 1.1/6; vocal-burst blend 2.8/10; 6.4s, EN.
emolia_c0758__C__sprightly-pixie__en.c026.k0 · in -21.0 dBFS · gain +1.0 dB · vprof_vc-00056
(fatigue exhaustion, embarrassment, teasing · measured, very low-energy, relaxed, whispered) Ach, du kannst doch nicht einfach auf dem Rücken liegen bleiben, nicht einmal am ersten Tag nach der Geburt. (normal breathing) Beweg dich einfach, du kannst in jeder Position sein, die du brauchst.
full caption & clip details
An adult somewhat feminine voice; delivery is very low-energy, measured, relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, fairly narrow pitch, light breath; affect is mildly negative, submissive, vulnerable; reads as fatigue exhaustion, embarrassment, teasing; style: whispered, monologue; good recording, no background noise; mildly explicit content; genuineness 1.5/6; vocal-burst blend 4.8/10; 12.1s, DE.
emolia_c0758__X__pain_groan__de.c032.k2 · in -26.4 dBFS · gain +6.4 dB · vprof_vc-00056
(emotional numbness, disgust, malevolence malice · normal-paced, normally alert, slightly relaxed, formal) Idiopathische Thrombozytopenische Purpura zeigt sich mit einer komplexen Pathophysiologie, die eine abnormale Thrombozytenzerstörung beinhaltet. Wir müssen den komplizierten Mechanismus der idiopathischen Thrombozytopenischen Purpura berücksichtigen.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as emotional numbness, disgust, malevolence malice; style: formal, newsreading; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.3/10; 15.3s, DE.
emolia_c0758__E__Contemplation__B__de.c036.k0 · in -21.9 dBFS · gain +1.9 dB · vprof_vc-00056
Affection ↓  /  Embarrassmentvp-AB2-k3 · #17

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Embarrassment up — by at least 0.25 each.

The chain starts with Embarrassment clearly present — 0.64, higher than 64 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.36.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.11 — a fairly even climb, though some clips carry more of the change than others.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 25 s · en · vprof_vc

k 3d_a -0.404d_b 0.361step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0955track emolia_c0955total 25.1slevel spread 3.0 dBmax seam 2.0 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: an adult masculine voice · neutral-bright, fairly smooth, no background noise
(affection, awe, infatuation · brisk, energised, slightly tense, casual) Because you shine with a light I can't ignore, every part of you, past and present, is utterly captivating to me. You deserve all the love in the world, always and forever.
full caption & clip details
An adult masculine voice; delivery is energised, brisk, slightly tense, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, almost no disfluency, wide pitch range, light breath; affect is mildly negative, slightly dominant, fairly guarded; reads as affection, awe, infatuation; style: casual, authoritative; good recording, no background noise; mildly explicit content; genuineness 0.7/6; vocal-burst blend 2.2/10; 8.1s, EN.
emolia_c0955__E__Infatuation__B__en.c017.k3 · in -22.5 dBFS · gain +2.5 dB · vprof_vc-00064
(distress, shame, sadness · normal-paced, normally alert, slightly relaxed, narration) Sein Gegner scheitern brachte ihm ein seltsames Gefühl, eine wahre Schadenfreude. Es war reines Deutsch, nur ein kleines bisschen Miteinander in diesem Sieg.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as distress, shame, sadness; style: narration, monologue; good recording, no background noise; genuineness 1.2/6; vocal-burst blend 2.7/10; 8.0s, DE.
emolia_c0955__E__Jealousy_and_Envy__C__de.c034.k3 · in -24.5 dBFS · gain +4.5 dB · vprof_vc-00064
(embarrassment, shame, infatuation · measured, very low-energy, neutral tension, conversational) I hate that this started so fast, this awful swelling. (fearful gasp) The symptoms of epiglottitis came on almost instantly, I feel so embarrassed.
full caption & clip details
A young adult masculine voice; delivery is very low-energy, measured, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, slightly thin; somewhat unclear, some disfluency, wide pitch range, normal breath; affect is mildly negative, submissive, vulnerable; reads as embarrassment, shame, infatuation; style: conversational, playful; average recording, no background noise; genuineness 1.3/6; vocal-burst blend 3.6/10; 8.7s, EN.
emolia_c0955__E__Shame__B__en.c007.k3 · in -25.5 dBFS · gain +5.5 dB · vprof_vc-00064
Affection ↓  /  Fatigue Exhaustionvp-AB2-k3 · #18

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Fatigue Exhaustion up — by at least 0.25 each.

The chain starts with Fatigue Exhaustion clearly present — 0.71, higher than 71 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.29.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.04 — most of the change happening immediately, then levelling off.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 22 s · en · vprof_vc

k 3d_a -0.404d_b 0.288step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c0955track emolia_c0955total 21.6slevel spread 1.8 dBmax seam 1.8 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, good recording, no background noise, moderate pitch range, light breath
(affection, awe, infatuation · brisk, energised, slightly tense, casual) Because you shine with a light I can't ignore, every part of you, past and present, is utterly captivating to me. You deserve all the love in the world, always and forever.
full caption & clip details
An adult masculine voice; delivery is energised, brisk, slightly tense, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, almost no disfluency, moderate pitch range, light breath; affect is mildly negative, slightly dominant, fairly guarded; reads as affection, awe, infatuation; style: casual, dramatic; good recording, no background noise; mildly explicit content; genuineness 0.9/6; vocal-burst blend 2.3/10; 8.1s, EN.
emolia_c0955__E__Infatuation__B__en.c017.k2 · in -23.1 dBFS · gain +3.1 dB · vprof_vc-00064
(longing, sexual lust, pain · brisk, normally alert, slightly relaxed, conversational) Your browser feels so tempting, I want to see what else it hides for me. (contented sigh) This system really makes my body crave your attention right now.
full caption & clip details
An adult masculine voice; delivery is normally alert, brisk, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, almost no disfluency, moderate pitch range, light breath; affect is mildly negative, slightly dominant, fairly guarded; reads as longing, sexual lust, pain; style: conversational; good recording, no background noise; genuineness 1.1/6; vocal-burst blend 1.0/10; 5.8s, EN.
emolia_c0955__E__Sexual_Lust__C__en.c036.k0 · in -24.3 dBFS · gain +4.3 dB · vprof_vc-00064
(fatigue exhaustion, relief, distress · normal-paced, normally alert, neutral tension, casual) Honestly, after all this, I'm so wiped out. (mournful wail) A quick shower and maybe some gentle chest massage might actually help me feel a little less drained.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as fatigue exhaustion, relief, distress; style: casual, storytelling; good recording, no background noise; genuineness 2.5/6; vocal-burst blend 3.1/10; 7.3s, EN.
emolia_c0955__E__Fatigue_Exhaustion__C__en.c032.k2 · in -22.4 dBFS · gain +2.5 dB · vprof_vc-00064
Affection ↓  /  Bitternessvp-AB2-k3 · #19

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Bitterness up — by at least 0.25 each.

The chain starts with Bitterness around average — 0.56, higher than 56 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.44.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.60 (higher than 60 % of clips in this corpus), a change of -0.40. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.19 — an even, steady climb — each clip carries about the same share.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 41 s · en · vprof_vc

k 3d_a -0.404d_b 0.438step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c1028track emolia_c1028total 41.1slevel spread 1.1 dBmax seam 0.7 dB
Script — 3 chunks, 1 with a non-speech sound
Unchanged across all 3 clips: a young adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, average clarity, light breath
(affection, contentment, pleasure ecstasy · normal-paced, normally alert, relaxed, casual) Because that necklace comes from a heart full of love, it shows a special blessing. Someone who wears it will surely be the most wonderful wife.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as affection, contentment, pleasure ecstasy; style: casual, monologue; good recording, no background noise; genuineness 1.2/6; vocal-burst blend 4.7/10; 9.1s, EN.
emolia_c1028__V__ATCK__extremely_low__en.c030.k2 · in -23.0 dBFS · gain +3.0 dB · vprof_vc-00072
(pain, infatuation, confusion · normal-paced, normally alert, neutral tension, casual) This E-mini Dow futures thing, like this weird buzzing in my head, it's like a tiny slice of the whole market. It feels so (ahem) strangely real right now.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as pain, infatuation, confusion; style: casual, playful; average recording, quiet background; mildly explicit content; genuineness 3.8/6; vocal-burst blend 5.3/10; 6.8s, EN.
emolia_c1028__E__Intoxication_Altered_States_of_Consciousness__D__en.c023.k0 · in -22.3 dBFS · gain +2.3 dB · vprof_vc-00072
(bitterness, sadness, longing · measured, subdued, slightly relaxed, monologue) Mit derselben Farbe und demselben Grundanstrich, wie es sich auf meiner Haut anfühlte, passte ich den Farbton der Schachtel an die Wand an, sodass es genau zu meinem Geschmack passte. Jeder Farbton fühlte sich an wie ein Versprechen, das ich konsumieren wollte.
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, no disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, vulnerable; reads as bitterness, sadness, longing; style: monologue, narration; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 3.9/10; 24.9s, DE.
emolia_c1028__E__Sexual_Lust__C__de.c031.k2 · in -21.9 dBFS · gain +1.9 dB · vprof_vc-00072
Affection ↓  /  Hope Enthusiasm Optimismvp-AB2-k3 · #20

This chain comes from the two-sided rule: it only counts if both emotions move — Affection down and Hope Enthusiasm Optimism up — by at least 0.25 each.

The chain starts with Hope Enthusiasm Optimism clearly present — 0.72, higher than 72 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.28.

At the same time Affection goes the other way, from 1.00 (virtually no clip in this corpus scores higher) to 0.75 (higher than 75 % of clips in this corpus), a change of -0.25. Both halves had to happen for this chain to qualify.

It takes 3 clips to get there. Clip to clip the moves are +0.25, then +0.03 — most of the change happening immediately, then levelling off.

No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.

Same speaker? Guaranteed, without needing a check: one voice profile is one cloned voice by construction, so every clip here is the same synthetic speaker.

Voice consistency: not an issue here — every clip in this chain is the same cloned voice by construction, so there is no speaker mismatch to hear.

3 clips · 22 s · en · vprof_vc

k 3d_a -0.252d_b 0.283step_a step_b min_cos_consec min_cos_anchor dataset vprof_vclang enspeaker emolia_c1028track emolia_c1028total 21.9slevel spread 5.5 dBmax seam 5.5 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: a young adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(affection, longing, contentment · neutral tension, moderately variable, little disfluency, storytelling) Those people, they just... (guffaw) they became family to me. I felt such a connection with the nurses, like they were my closest friends for so long.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as affection, longing, contentment; style: storytelling, narration; good recording, no background noise; genuineness 0.5/6; vocal-burst blend 2.5/10; 7.6s, EN.
emolia_c1028__E__Infatuation__C__en.c001.k1 · in -22.7 dBFS · gain +2.7 dB · vprof_vc-00072
(awe, pleasure ecstasy, contentment · slightly relaxed, fairly steady, almost no disfluency, casual) Now that that old shadow of pride has lifted, we can truly see the magnificent spark of the Divine within him. Because of that, God can work through him, and every beautiful plan unfolds just as it should.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as awe, pleasure ecstasy, contentment; style: casual, narration; good recording, no background noise; genuineness 1.7/6; vocal-burst blend 1.5/10; 8.2s, EN.
emolia_c1028__E__Hope_Enthusiasm_Optimism__D__en.c015.k3 · in -22.9 dBFS · gain +3.0 dB · vprof_vc-00072
(hope enthusiasm optimism, elation, disgust · slightly relaxed, fairly steady, some disfluency, casual) Dude, I'm seriously chuffed about this gig, like stoked to bits. It's gonna be sick, no cap, I'm totally hyped for this whole thing.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as hope enthusiasm optimism, elation, disgust; style: casual, storytelling; good recording, no background noise; genuineness 1.9/6; vocal-burst blend 1.3/10; 5.8s, EN.
emolia_c1028__E__Sexual_Lust__D__en.c005.k0 · in -17.4 dBFS · gain -2.6 dB · vprof_vc-00072