Rule.B1 — one-sided: emotion B rises by >=T; the other axis is unconstrained Source. trajectories_v5.parquet | Family. one corpus in isolation Sampled from 177,245 matching rows, without replacement across the family, so no two tiers reuse a chain.
How to read a Script. Each chunk is one line:
a short tag of what the models heard in that clip, then the words spoken.
(underlined, plain ·
delivery, style) — the tag before the words. Emotions first, then how
it is delivered. Underlined descriptors are the ones that change across this chain
— anything identical on every clip is pulled out and stated once above, because a
value that never moves says nothing about a trajectory.
(ahem) — brackets inside the words
are a different thing: a real non-speech sound, printed where it happens. Most clips have
none; about a quarter do.
The full generated caption for any clip is still there, under
“full caption & clip details”. Its perceived-gender and background-noise
clauses were re-rendered from the numeric buckets, because the versions stored in the
corpus index had those two ladders running backwards.
The emotion clause has been re-derived, and it used to be wrong. The 40 emotion heads are not on a common scale — Interest has a median of 2.08 and is never zero, while Infatuation is zero on 87.7 % of clips — and the caption named an emotion whenever its raw score cleared an absolute 1.0. Interest therefore appeared in 94.8 % of captions and Sadness in almost none: the clause was reporting the scale of the head, not the emotion of the clip. An emotion is now named only when it lands in the top 10 % for that emotion, against a pooled tie-aware mid-rank ECDF over 132,833,726 utterances spanning every dataset and language. Interest now appears in 6.2 %, all 40 emotions occur, and a clip that is ordinary on all 40 says “no dominant emotion” rather than being forced to pick one (21.8 % of clips). This is the same scale the trajectory miner selects on, so the caption and the mining now refer to the same quantity: the mined target emotion is named in the final clip's caption on 73 % of chains, up from 46 %.
This chain comes from the one-sided rule: only Fatigue Exhaustion had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Fatigue Exhaustion clearly present — 0.69, higher than 69 % of clips in this corpus — and ends with it at the very top of the corpus at 0.95, higher than 95 % of clips in this corpus. That is a total rise of 0.26.
Nothing was asked of the other axis, and in fact Teasing drifts down from 0.87 to 0.36 (-0.50), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.11, then +0.06, then -0.09, then +0.18 — not a clean run: step 3 moves back the other way by 0.09 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 19 s · snippets
hear it un-normalised (raw levels, max seam 0.8 dB)
k 5d_a -0.503d_b 0.262step_a 0.503step_b 0.183min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch255_part2_batch255_patrack batch255_part2_batch255_patotal 19.5slevel spread 0.8 dBmax seam 0.8 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, good recording, no background noise, normally alert, slightly relaxed, no disfluency
(measured, fairly steady, narration, formal)ten times more frequently than their European cousins.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, no disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; no dominant emotion; style: narration, formal; good recording, no background noise; genuineness 0.5/6; vocal-burst blend 2.7/10; 4.0s.
batch255_part2_batch255_part2_chunk_763_1_821515 · in -27.2 dBFS · gain +7.2 dB · snippets-00810
(normal-paced, fairly steady, formal, narration)But fewer than one in five daytime hunts are successful.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, narration; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 4.9/10; 3.4s.
batch255_part2_batch255_part2_chunk_763_1_821526 · in -27.4 dBFS · gain +7.5 dB · snippets-00810
(normal-paced, steady, narration, formal)will spend most of the night foraging in the dried lagoon beds.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: narration, formal; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 2.7/10; 4.1s.
batch255_part2_batch255_part2_chunk_763_1_821612 · in -27.7 dBFS · gain +7.7 dB · snippets-00810
(fear, emotional numbness·measured, fairly steady, formal, monologue)all the area, investigating any potential threats.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as fear, emotional numbness; style: formal, monologue; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 2.1/10; 3.5s.
batch255_part2_batch255_part2_chunk_763_1_821657 · in -27.8 dBFS · gain +7.8 dB · snippets-00810
(fatigue exhaustion, relief, longing· measured, fairly steady, narration, monologue)Party winding down. It's time to find a roost for the night.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as fatigue exhaustion, relief, longing; style: narration, monologue; good recording, no background noise; genuineness 0.6/6; vocal-burst blend 4.2/10; 3.8s.
batch255_part2_batch255_part2_chunk_763_1_821681 · in -27.0 dBFS · gain +7.0 dB · snippets-00810
This chain comes from the one-sided rule: only Longing had to get where it was going, by at least 0.25. The other emotion was left completely free.
The chain starts with Longing around average — 0.57, higher than 57 % of clips in this corpus — and ends with it at the very top of the corpus at 0.94, higher than 94 % of clips in this corpus. That is a total rise of 0.38.
Nothing was asked of the other axis, and in fact Relief barely moves at all, sitting near 0.81 throughout.
It takes 4 clips to get there. Clip to clip the moves are +0.18, then +0.20, then +0.00 — a plateau around step 3, where it barely moves.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 28 s · snippets
hear it un-normalised (raw levels, max seam 3.5 dB)
k 4d_a 0.014d_b 0.378step_a 0.706step_b 0.196min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch95_part0_batch95_parttrack batch95_part0_batch95_parttotal 27.8slevel spread 3.5 dBmax seam 3.5 dB
Script — 4 chunks, 4 with a non-speech sound
Unchanged across all 4 clips: an adult masculine voice · neutral-toned, slightly dark, average recording, measured, fairly steady
(normally alert, slightly relaxed, frequent disfluency, casual)(ahem) just about that moment but transferred from fighter command.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; no dominant emotion; style: casual, conversational; average recording, no background noise; genuineness 3.7/6; vocal-burst blend 3.9/10; 3.7s.
batch95_part0_batch95_part0_chunk_1851_1_1643973 · in -27.0 dBFS · gain +7.0 dB · snippets-01381
(astonishment surprise, relief, triumph· normally alert, neutral tension, frequent disfluency, casual)ted and uh (low mumble) anyway so here it was and I was going down I suppose I got down to about
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, neutral tension, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, normal breath; affect is neutral, neutral stance, neutral openness; reads as astonishment surprise, relief, triumph; style: casual, conversational; average recording, quiet background; genuineness 4.9/6; vocal-burst blend 4.9/10; 6.0s.
batch95_part0_batch95_part0_chunk_1851_1_1644044 · in -30.3 dBFS · gain +10.3 dB · snippets-01381
(fear, sadness, distress·very low-energy, neutral tension, some disfluency, casual)and the young man knew that they fuse had started burning and he had 18 seconds I think it was before the bomb went off. (low mumble) Um
full caption & clip details
An elderly masculine voice; delivery is very low-energy, measured, neutral tension, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, slightly thin; somewhat unclear, some disfluency, fairly narrow pitch, normal breath; affect is mildly negative, neutral stance, neutral openness; reads as fear, sadness, distress; style: casual, monologue; average recording, quiet background; genuineness 3.5/6; vocal-burst blend 4.5/10; 7.8s.
batch95_part0_batch95_part0_chunk_1851_1_1644101 · in -26.7 dBFS · gain +6.7 dB · snippets-01381
(longing·normally alert, neutral tension, some disfluency, conversational)And when I looked down, I, because it's difficult to see down in a Spitfire, a great long nose, you know, from the side, (ahem) unless you can put your head outside the cockpit, which I couldn't of course, so I couldn't get the hood off.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, neutral tension, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; average clarity, some disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as longing; style: conversational, casual; average recording, quiet background; genuineness 3.2/6; vocal-burst blend 3.8/10; 9.8s.
batch95_part0_batch95_part0_chunk_1851_1_1644141 · in -29.8 dBFS · gain +9.8 dB · snippets-01381
Embarrassment ↑ (unconstrained axis: Sexual Lust)c-snippets-B1 · #3
This chain comes from the one-sided rule: only Embarrassment had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Embarrassment clearly present — 0.74, higher than 74 % of clips in this corpus — and ends with it at the very top of the corpus at 0.97, higher than 97 % of clips in this corpus. That is a total rise of 0.23.
Nothing was asked of the other axis, and in fact Sexual Lust drifts down from 1.00 to 0.89 (-0.11), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are -0.03, then +0.18, then +0.08 — not a clean run: step 1 moves back the other way by 0.03 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 18 s · snippets
hear it un-normalised (raw levels, max seam 10.7 dB)
k 4d_a -0.111d_b 0.232step_a 0.064step_b 0.184min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch151_part1_batch151_patrack batch151_part1_batch151_patotal 18.0slevel spread 10.7 dBmax seam 10.7 dB
Script — 4 chunks, 3 with a non-speech sound
Unchanged across all 4 clips: a young adult masculine voice · neutral-toned, average recording, normally alert, some disfluency
(sexual lust, infatuation, pleasure ecstasy · normal-paced, neutral tension, moderately variable, casual)I tell you what I love about a sunset. And this is what I love about my power. Is we're one beer in and I'm having beautiful thoughts. Yeah. You definitely are. (low mumble) Um.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly negative, slightly dominant, neutral openness; reads as sexual lust, infatuation, pleasure ecstasy; style: casual, conversational; average recording, quiet background; mildly explicit content; genuineness 5.8/6; vocal-burst blend 5.3/10; 6.7s.
batch151_part1_batch151_part1_chunk_2357_1_2127424 · in -33.2 dBFS · gain +13.2 dB · snippets-00271
(infatuation, sexual lust, intoxication altered states of consciousness· normal-paced, slightly relaxed, fairly steady, casual)where we wake up like you do and smoke a blunt in the morning in bed.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, dark, slightly rough, thin; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is mildly negative, slightly dominant, neutral openness; reads as infatuation, sexual lust, intoxication altered states of consciousness; style: casual, conversational; average recording, quiet background; genuineness 4.1/6; vocal-burst blend 4.7/10; 3.6s.
batch151_part1_batch151_part1_chunk_2357_1_2127642 · in -29.8 dBFS · gain +9.8 dB · snippets-00271
(intoxication altered states of consciousness, fatigue exhaustion, sexual lust ·measured, fully relaxed, fairly steady, casual)And then I'm (low mumble) I'm up we lived Obi and I lived in an attic.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, fully relaxed, fairly steady; timbre is neutral-toned, dark, fairly smooth, thin; slurred, some disfluency, fairly narrow pitch, audible breath; affect is mildly negative, neutral stance, neutral openness; reads as intoxication altered states of consciousness, fatigue exhaustion, sexual lust; style: casual, conversational; average recording, no background noise; genuineness 4.5/6; vocal-burst blend 3.8/10; 3.2s.
batch151_part1_batch151_part1_chunk_2357_1_2127684 · in -36.0 dBFS · gain +16.0 dB · snippets-00271
(embarrassment, teasing·normal-paced, slightly relaxed, moderately variable, conversational)I don't know about that. Maybe you should (ahem) uh talk to Burt. Hold on a second. No, (low mumble) um, Cody.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is neutral, slightly dominant, neutral openness; reads as embarrassment, teasing; style: conversational, casual; average recording, quiet background; genuineness 4.7/6; vocal-burst blend 1.4/10; 4.0s.
batch151_part1_batch151_part1_chunk_2357_1_2127722 · in -25.2 dBFS · gain +5.2 dB · snippets-00271
This chain comes from the one-sided rule: only Helplessness had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Helplessness strongly present — 0.76, higher than 76 % of clips in this corpus — and ends with it at the very top of the corpus at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.23.
Nothing was asked of the other axis, and in fact Fear climbs from 0.88 to 0.98 (+0.10), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.22, then -0.03, then +0.03 — not a clean run: step 2 moves back the other way by 0.03 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 26 s · snippets
hear it un-normalised (raw levels, max seam 2.7 dB)
k 4d_a 0.104d_b 0.227step_a 0.834step_b 0.219min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch235_part3_batch235_patrack batch235_part3_batch235_patotal 25.6slevel spread 2.7 dBmax seam 2.7 dB
Script — 4 chunks, 1 with a non-speech sound
Unchanged across all 4 clips: an adult masculine voice · fairly smooth, quiet background, light breath
(brisk, energised, slightly tense, dramatic)it releases the same amount of cortisol based on how much you feel in your brain it is dangerous.
full caption & clip details
An adult masculine voice; delivery is energised, brisk, slightly tense, moderately variable; timbre is slightly cool, slightly bright, fairly smooth, full; clear, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; no dominant emotion; style: dramatic, authoritative; average recording, quiet background; genuineness 1.4/6; vocal-burst blend 3.0/10; 7.0s.
batch235_part3_batch235_part3_chunk_586_1_301413 · in -23.9 dBFS · gain +3.9 dB · snippets-00705
(doubt, helplessness, impatience and irritability·normal-paced, normally alert, neutral tension, casual)that I haven't had the time necessary to plan this fundraiser.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, neutral tension, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, thin; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as doubt, helplessness, impatience and irritability; style: casual, storytelling; good recording, quiet background; genuineness 1.8/6; vocal-burst blend 1.8/10; 4.1s.
batch235_part3_batch235_part3_chunk_586_1_301431 · in -26.0 dBFS · gain +6.0 dB · snippets-00705
(helplessness, longing, embarrassment· normal-paced, normally alert, relaxed, casual)Yeah, it's like I go from, I go from being overwhelmed with business to being, to being (low mumble) uh...
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, thin; somewhat unclear, frequent disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as helplessness, longing, embarrassment; style: casual, conversational; average recording, quiet background; genuineness 5.5/6; vocal-burst blend 8.3/10; 5.2s.
batch235_part3_batch235_part3_chunk_586_1_301440 · in -23.3 dBFS · gain +3.3 dB · snippets-00705
(helplessness, fear, distress·brisk, energised, neutral tension, casual)that I'm like, I need a mental break. And those mental breaks are important because you don't want to get burnt out, you don't want to get stressed, you don't want to run yourself into the ground.
full caption & clip details
A young adult masculine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, thin; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, neutral openness; reads as helplessness, fear, distress; style: casual, conversational; average recording, quiet background; mildly explicit content; genuineness 4.6/6; vocal-burst blend 10.0/10; 8.9s.
batch235_part3_batch235_part3_chunk_586_1_301530 · in -24.9 dBFS · gain +4.9 dB · snippets-00705
This chain comes from the one-sided rule: only Infatuation had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Infatuation below average — 0.27, lower than 73 % of clips in this corpus — and ends with it at the very top of the corpus at 0.96, higher than 96 % of clips in this corpus. That is a total rise of 0.69.
Nothing was asked of the other axis, and in fact Astonishment Surprise drifts down from 0.66 to 0.52 (-0.13), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.14, then +0.12, then +0.18, then +0.25 — a fairly even climb, though some clips carry more of the change than others.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 26 s · snippets
hear it un-normalised (raw levels, max seam 1.6 dB)
k 5d_a -0.135d_b 0.688step_a 0.199step_b 0.247min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch230_part3_batch230_patrack batch230_part3_batch230_patotal 26.1slevel spread 1.6 dBmax seam 1.6 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, good recording, no background noise, normal-paced, normally alert, slightly relaxed
(some disfluency, average clarity, casual, monologue)firing their engines once to break out of the halo orbit.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; no dominant emotion; style: casual, monologue; good recording, no background noise; genuineness 2.3/6; vocal-burst blend 1.9/10; 3.4s.
batch230_part3_batch230_part3_chunk_540_1_323909 · in -25.3 dBFS · gain +5.3 dB · snippets-00681
(no disfluency, clear, authoritative, formal)Three of its four cooling pumps needed replacing.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; no dominant emotion; style: authoritative, formal; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 1.3/10; 3.1s.
batch230_part3_batch230_part3_chunk_540_1_324000 · in -25.2 dBFS · gain +5.2 dB · snippets-00681
(fear·almost no disfluency, clear, conversational, narration)hearing the end of this journey the service module is released and the crew module is oriented heat shield first.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as fear; style: conversational, narration; good recording, no background noise; genuineness 1.5/6; vocal-burst blend 0.8/10; 7.3s.
batch230_part3_batch230_part3_chunk_540_1_324029 · in -25.4 dBFS · gain +5.4 dB · snippets-00681
(awe, elation, hope enthusiasm optimism·no disfluency, clear, formal, narration)In this episode of Tech Effect, a new mission to the moon sets the stage for the first crude flight to Mars.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, full; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as awe, elation, hope enthusiasm optimism; style: formal, narration; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.6/10; 7.4s.
batch230_part3_batch230_part3_chunk_540_1_324062 · in -24.8 dBFS · gain +4.8 dB · snippets-00681
(infatuation· no disfluency, clear, narration, formal)has the most sophisticated particle detector ever sent into space.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as infatuation; style: narration, formal; good recording, no background noise; genuineness 1.5/6; vocal-burst blend 0.8/10; 4.4s.
batch230_part3_batch230_part3_chunk_540_1_324078 · in -26.4 dBFS · gain +6.4 dB · snippets-00681
Impatience and Irritability ↑ (unconstrained axis: Intoxication Altered States of Consciousness)c-snippets-B1 · #6
This chain comes from the one-sided rule: only Impatience and Irritability had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Impatience and Irritability clearly present — 0.74, higher than 74 % of clips in this corpus — and ends with it at the very top of the corpus at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.25.
Nothing was asked of the other axis, and in fact Intoxication Altered States of Consciousness drifts down from 1.00 to 0.89 (-0.11), which the rule did not require.
It takes 3 clips to get there. Clip to clip the moves are +0.20, then +0.05 — most of the change happening immediately, then levelling off.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
3 clips · 15 s · snippets
k 3d_a -0.111d_b 0.248step_a 0.083step_b 0.203min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch157_part3_batch157_patrack batch157_part3_batch157_patotal 14.8slevel spread 3.0 dBmax seam 3.0 dB
Script — 3 chunks, 2 with a non-speech sound
Unchanged across all 3 clips: a young adult masculine voice · moderately variable, some disfluency, average clarity, wide pitch range
(intoxication altered states of consciousness, embarrassment, triumph · fast, energised, neutral tension, casual)we did we do we did we did do a podcast. I told I told I told I told (ahem) Nadav.
full caption & clip details
A young adult masculine voice; delivery is energised, fast, neutral tension, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly dominant, neutral openness; reads as intoxication altered states of consciousness, embarrassment, triumph; style: casual, conversational; average recording, quiet background; genuineness 5.0/6; vocal-burst blend 6.5/10; 4.9s.
batch157_part3_batch157_part3_chunk_2413_1_2240230 · in -22.5 dBFS · gain +2.5 dB · snippets-00303
(pain, sadness, helplessness·normal-paced, normally alert, slightly relaxed, conversational)dying. Yeah, dying. And I'm just kind of on the couch like, (low mumble) uh, when do you say, when your dog is leaving?
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, neutral openness; reads as pain, sadness, helplessness; style: conversational, casual; good recording, no background noise; genuineness 4.3/6; vocal-burst blend 1.4/10; 5.4s.
batch157_part3_batch157_part3_chunk_2413_1_2240344 · in -23.8 dBFS · gain +3.8 dB · snippets-00303
(impatience and irritability, sexual lust, amusement·brisk, highly aroused, slightly tense, ranting)Let me finish my story. No, no, let me tell my side of the story. No, but I'm not done with mine!
full caption & clip details
A young adult somewhat masculine voice; delivery is highly aroused, brisk, slightly tense, moderately variable; timbre is slightly cool, slightly bright, very rough, thin; average clarity, some disfluency, wide pitch range, normal breath; affect is elated, slightly dominant, guarded; reads as impatience and irritability, sexual lust, amusement; style: ranting, casual; below-average recording, some background noise; genuineness 3.6/6; vocal-burst blend 3.6/10; 4.1s.
batch157_part3_batch157_part3_chunk_2413_1_2240420 · in -20.8 dBFS · gain +0.8 dB · snippets-00303
This chain comes from the one-sided rule: only Contentment had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Contentment clearly present — 0.68, higher than 68 % of clips in this corpus — and ends with it strongly present at 0.90, higher than 90 % of clips in this corpus. That is a total rise of 0.22.
Nothing was asked of the other axis, and in fact Sourness drifts down from 0.93 to 0.09 (-0.84), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.07, then +0.08, then +0.06 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 42 s · snippets
k 4d_a -0.842d_b 0.221step_a 0.934step_b 0.080min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch137_part4_batch137_patrack batch137_part4_batch137_patotal 42.3slevel spread 7.5 dBmax seam 5.4 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: a middle-aged masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, normally alert, moderate pitch range, light breath
(sourness, bitterness, pleasure ecstasy · measured, slightly relaxed, fairly steady, narration)Back to the woman he had just been talking with. It was a revelation in the light of which he already saw she would become more interesting.
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; reads as sourness, bitterness, pleasure ecstasy; style: narration, formal; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 1.5/10; 7.6s.
batch137_part4_batch137_part4_chunk_2226_1_2426184 · in -24.2 dBFS · gain +4.2 dB · snippets-00195
(infatuation, sexual lust, embarrassment·normal-paced, slightly relaxed, moderately variable, casual)And I have a particular memory of them doing a cover of I can't get no satisfaction by the Rolling Stones on Saturday Night Live wearing these yellow hazmat suits, you know, like they were really weird. Weird. Yes.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as infatuation, sexual lust, embarrassment; style: casual, conversational; good recording, quiet background; genuineness 3.5/6; vocal-burst blend 8.3/10; 12.8s.
batch137_part4_batch137_part4_chunk_2226_1_2426288 · in -21.9 dBFS · gain +1.9 dB · snippets-00195
(amusement, embarrassment, astonishment surprise· normal-paced, neutral tension, moderately variable, casual)But as a kid you I thought it was like just another religion, you know. Okay, well I'm going she picked me up from church school. She was very upset that I yeah I was baptized. I mean the whole thing.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as amusement, embarrassment, astonishment surprise; style: casual, conversational; good recording, no background noise; genuineness 4.6/6; vocal-burst blend 7.4/10; 10.1s.
batch137_part4_batch137_part4_chunk_2226_1_2426530 · in -22.1 dBFS · gain +2.1 dB · snippets-00195
(normal-paced, slightly relaxed, fairly steady, casual)So Neva, as you probably remember on all of our prior bridge episodes, we're going to ask you three trivia questions. The first is going to be a callback to our most recent full-length episode of Hit Parade.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; no dominant emotion; style: casual, monologue; good recording, no background noise; genuineness 1.4/6; vocal-burst blend 1.6/10; 11.4s.
batch137_part4_batch137_part4_chunk_2226_1_2426628 · in -16.7 dBFS · gain -3.3 dB · snippets-00195
This chain comes from the one-sided rule: only Doubt had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Doubt clearly present — 0.66, higher than 66 % of clips in this corpus — and ends with it at the very top of the corpus at 0.97, higher than 97 % of clips in this corpus. That is a total rise of 0.30.
Nothing was asked of the other axis, and in fact Concentration drifts down from 0.92 to 0.75 (-0.17), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.22, then -0.13, then +0.22 — not a clean run: step 2 moves back the other way by 0.13 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 31 s · snippets
k 4d_a -0.172d_b 0.304step_a 0.526step_b 0.220min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch213_part3_batch213_patrack batch213_part3_batch213_patotal 30.9slevel spread 2.7 dBmax seam 2.7 dB
Script — 4 chunks, 1 with a non-speech sound
Unchanged across all 4 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, normally alert, slightly relaxed, light breath
(concentration · normal-paced, fairly steady, some disfluency, casual)Initially the concept was like, oh, okay, let's harden, let's improve their cyber security. Let's harden (low mumble) um, uh voter registration databases and election night reporting and the systems that are used throughout the process.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, slightly guarded; reads as concentration; style: casual, conversational; average recording, quiet background; genuineness 3.1/6; vocal-burst blend 4.1/10; 13.6s.
batch213_part3_batch213_part3_chunk_387_1_150305 · in -16.6 dBFS · gain -3.5 dB · snippets-00588
(normal-paced, fairly steady, little disfluency, casual)telling us that their constituents would expect us to do even more.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: casual, monologue; good recording, no background noise; genuineness 2.4/6; vocal-burst blend 1.9/10; 3.6s.
batch213_part3_batch213_part3_chunk_387_1_150393 · in -19.3 dBFS · gain -0.7 dB · snippets-00588
(measured, fairly steady, no disfluency, monologue)when when they go and serve there. If
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, neutral openness; no dominant emotion; style: monologue, casual; good recording, no background noise; genuineness 1.3/6; vocal-burst blend 2.1/10; 3.0s.
batch213_part3_batch213_part3_chunk_387_1_150452 · in -16.9 dBFS · gain -3.1 dB · snippets-00588
(doubt, disappointment, bitterness·brisk, moderately variable, little disfluency, conversational)Michael, what about the adversarial relationship between this White House and the intelligence agencies? What are the implications of a president who has been so openly critical of the agency?
full caption & clip details
An adult feminine voice; delivery is normally alert, brisk, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, neutral openness; reads as doubt, disappointment, bitterness; style: conversational, playful; good recording, no background noise; genuineness 1.4/6; vocal-burst blend 0.9/10; 10.3s.
batch213_part3_batch213_part3_chunk_387_1_150516 · in -18.4 dBFS · gain -1.6 dB · snippets-00588
This chain comes from the one-sided rule: only Bitterness had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Bitterness barely there — 0.20, lower than 80 % of clips in this corpus — and ends with it clearly present at 0.69, higher than 69 % of clips in this corpus. That is a total rise of 0.49.
Nothing was asked of the other axis, and in fact Sadness drifts down from 0.93 to 0.43 (-0.50), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.15, then +0.14, then +0.16, then +0.04 — a plateau around step 4, where it barely moves.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 24 s · snippets
k 5d_a -0.497d_b 0.494step_a 0.566step_b 0.161min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch7_part2_batch7_part2_track batch7_part2_batch7_part2_total 23.6slevel spread 2.3 dBmax seam 2.0 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, good recording, no background noise, normally alert, light breath
(sadness, emotional numbness, distress · normal-paced, slightly relaxed, fairly steady, formal)The zero hunger program ensures that all school children in the country get to eat at least one meal a day.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as sadness, emotional numbness, distress; style: formal, narration; good recording, no background noise; genuineness 0.4/6; vocal-burst blend 1.7/10; 5.7s.
batch7_part2_batch7_part2_chunk_1061_1_1073191 · in -24.8 dBFS · gain +4.8 dB · snippets-01300
(fear, distress, sadness ·brisk, neutral tension, moderately variable, storytelling)Sometimes they come home very late, and I always am afraid then that something has happened to them on the river.
full caption & clip details
An adult feminine voice; delivery is normally alert, brisk, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, wide pitch range, light breath; affect is mildly negative, neutral stance, vulnerable; reads as fear, distress, sadness; style: storytelling, narration; good recording, no background noise; genuineness 1.0/6; vocal-burst blend 1.0/10; 6.6s.
batch7_part2_batch7_part2_chunk_1061_1_1073250 · in -26.8 dBFS · gain +6.8 dB · snippets-01300
(infatuation, longing, sexual lust·normal-paced, slightly relaxed, fairly steady, storytelling)When I grow up, I want to study and become a teacher.
full caption & clip details
A child feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, no disfluency, moderate pitch range, light breath; affect is positive, slightly submissive, neutral openness; reads as infatuation, longing, sexual lust; style: storytelling, casual; good recording, no background noise; genuineness 1.3/6; vocal-burst blend 2.2/10; 3.4s.
batch7_part2_batch7_part2_chunk_1061_1_1073274 · in -25.6 dBFS · gain +5.6 dB · snippets-01300
(relief·measured, slightly relaxed, fairly steady, narration)though they have just set off, water is already swamping the boat.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, very full; clear, no disfluency, fairly narrow pitch, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as relief; style: narration, storytelling; good recording, no background noise; genuineness 0.4/6; vocal-burst blend 3.2/10; 4.0s.
batch7_part2_batch7_part2_chunk_1061_1_1073379 · in -26.4 dBFS · gain +6.4 dB · snippets-01300
(normal-paced, slightly relaxed, fairly steady, formal)About 30% of people in the rural areas of Nicaragua.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, monologue; good recording, no background noise; genuineness 0.9/6; vocal-burst blend 1.4/10; 3.4s.
batch7_part2_batch7_part2_chunk_1061_1_1073443 · in -24.5 dBFS · gain +4.5 dB · snippets-01300
Pride ↑ (unconstrained axis: Disappointment)c-snippets-B1 · #10
This chain comes from the one-sided rule: only Pride had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Pride clearly present — 0.67, higher than 67 % of clips in this corpus — and ends with it at the very top of the corpus at 0.98, higher than 98 % of clips in this corpus. That is a total rise of 0.31.
Nothing was asked of the other axis, and in fact Disappointment drifts down from 0.99 to 0.39 (-0.60), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.18, then -0.03, then +0.16 — not a clean run: step 2 moves back the other way by 0.03 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 48 s · snippets
k 4d_a -0.599d_b 0.313step_a 0.599step_b 0.184min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch66_part2_batch66_parttrack batch66_part2_batch66_parttotal 48.0slevel spread 4.2 dBmax seam 3.4 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: an adult masculine voice · neutral-toned, light breath
(disappointment, astonishment surprise, jealousy and envy · normal-paced, normally alert, slightly relaxed, casual)There are 600,000 food items in America. 80% of them have added sugar. Your brain lights up with sugar just like it does with cocaine or heroin. You're going to become an addict. You end up with one of the great public health epidemics of our time. This talk is actually quite popular and largely focused on sugar and childhood obesity. It's sad to hear parents saying it's cheaper to buy prepared food over fresh.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as disappointment, astonishment surprise, jealousy and envy; style: casual, conversational; good recording, quiet background; genuineness 2.4/6; vocal-burst blend 7.8/10; 22.5s.
batch66_part2_batch66_part2_chunk_1587_1_1883309 · in -25.6 dBFS · gain +5.6 dB · snippets-01226
(affection, pleasure ecstasy, elation· normal-paced, normally alert, relaxed, casual)So, I love films and I love documentaries, like love. Love, love. Love, love, love.
full caption & clip details
A child masculine voice; delivery is normally alert, normal-paced, relaxed, moderately variable; timbre is neutral-toned, slightly bright, slightly rough, balanced body; very clear, frequent disfluency, wide pitch range, light breath; affect is positive, slightly dominant, neutral openness; reads as affection, pleasure ecstasy, elation; style: casual, conversational; good recording, quiet background; genuineness 1.5/6; vocal-burst blend 2.6/10; 6.4s.
batch66_part2_batch66_part2_chunk_1587_1_1883541 · in -22.2 dBFS · gain +2.2 dB · snippets-01226
(interest, hope enthusiasm optimism·brisk, energised, slightly relaxed, casual)gut reaction. Could what we eat be the ideal ingredient for treating many medical conditions? makes a compelling case for gut health, eating upwards of 50 grams of fiber, and is a great intro in understanding the microbiome and its importance.
full caption & clip details
An adult masculine voice; delivery is energised, brisk, slightly relaxed, fairly steady; timbre is neutral-toned, slightly bright, fairly smooth, full; average clarity, almost no disfluency, wide pitch range, light breath; affect is positive, slightly dominant, slightly guarded; reads as interest, hope enthusiasm optimism; style: casual, playful; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 1.3/10; 14.3s.
batch66_part2_batch66_part2_chunk_1587_1_1883563 · in -21.4 dBFS · gain +1.4 dB · snippets-01226
(pride, amusement, teasing· brisk, energised, slightly tense, storytelling)Fasting. If the last was the granddaddy then this must be the grandmammy.
full caption & clip details
A young adult masculine voice; delivery is energised, brisk, slightly tense, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, thin; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly dominant, fairly guarded; reads as pride, amusement, teasing; style: storytelling, casual; average recording, quiet background; genuineness 2.3/6; vocal-burst blend 2.7/10; 4.2s.
batch66_part2_batch66_part2_chunk_1587_1_1883597 · in -22.4 dBFS · gain +2.4 dB · snippets-01226
This chain comes from the one-sided rule: only Concentration had to get where it was going, by at least 0.25. The other emotion was left completely free.
The chain starts with Concentration clearly present — 0.63, higher than 63 % of clips in this corpus — and ends with it at the very top of the corpus at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.36.
Nothing was asked of the other axis, and in fact Helplessness drifts down from 0.97 to 0.76 (-0.21), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.09, then +0.25, then +0.03 — a plateau around step 3, where it barely moves.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 119 s · snippets
k 4d_a -0.212d_b 0.362step_a 0.565step_b 0.246min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch203_part1_batch203_patrack batch203_part1_batch203_patotal 118.7slevel spread 7.5 dBmax seam 7.5 dB
Script — 4 chunks, 4 with a non-speech sound
Unchanged across all 4 clips: a young adult masculine voice · neutral-toned, fairly smooth, average recording, quiet background, normal-paced, slightly relaxed, fairly steady, moderate pitch range
(helplessness, sadness, distress · normally alert, frequent disfluency, somewhat unclear, casual)(low mumble) eh, en un futuro (low mumble) eh, a la resiliencia a la resiliencia del sistema de agua potable
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, normal breath; affect is neutral, neutral stance, slightly guarded; reads as helplessness, sadness, distress; style: casual; average recording, quiet background; genuineness 4.2/6; vocal-burst blend 3.4/10; 5.9s.
batch203_part1_batch203_part1_chunk_294_1_176281 · in -22.7 dBFS · gain +2.7 dB · snippets-00539
(thankfulness gratitude, relief, interest·very low-energy, frequent disfluency, somewhat unclear, casual)uh (ahem) springing up to fill in the gaps and respond. And we're seeing that in Ghana, we saw that in Puerto Rico, we saw that, we're seeing that in Bangladesh (ahem) um and elsewhere. So I'll let other panelists (ahem) uh respond if they have other comments.
full caption & clip details
A young adult feminine voice; delivery is very low-energy, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, slightly thin; somewhat unclear, frequent disfluency, moderate pitch range, normal breath; affect is mildly positive, neutral stance, neutral openness; reads as thankfulness gratitude, relief, interest; style: casual, monologue; average recording, quiet background; genuineness 4.5/6; vocal-burst blend 3.3/10; 15.3s.
batch203_part1_batch203_part1_chunk_294_1_176435 · in -30.2 dBFS · gain +10.2 dB · snippets-00539
(contemplation, interest, concentration·subdued, frequent disfluency, average clarity, casual)That's really interesting and I think those are some of the key core kind of tensions or questions that our project explores. (ahem) Um, as we've seen today, I mean in in certain cases some want to relocate and can't because of a lack of resources, others do not want to relocate. (low mumble) Um, and so and some are forced to to survive. (ahem) Um, so there's kind of every every scenario even in very discrete locations. (low mumble) (low mumble) Um, and so there isn't really a one size fits all solution. (low mumble) Um, and uh, (low mumble) my response to like my my quick response, actually, I I see a question also in the chat about what congressional actions or state department initiatives are addressing these issues. I want to bring up a report that just
full caption & clip details
A young adult feminine voice; delivery is subdued, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as contemplation, interest, concentration; style: casual, monologue; average recording, quiet background; genuineness 2.6/6; vocal-burst blend 0.7/10; 48.6s.
batch203_part1_batch203_part1_chunk_294_1_176489 · in -28.0 dBFS · gain +8.0 dB · snippets-00539
(concentration, hope enthusiasm optimism, fear· subdued, some disfluency, average clarity, casual)the military, (low mumble) um, and, you know, national security arms, (ahem) um, it's a little concerning to me, although it is couched in a lot of humanitarian language. And so the US position, (ahem) um, there are some initiatives, there are a lot, (ahem) um, congressional members who are pushing for allocations or protections for climate migrants, especially within the United States. And the US has allocated some funding for relocation, such as in Southern Louisiana. (ahem) (ahem) Um, however, this this this looking outward to the rest of the world and this position of the United States to say, we need to allocate aid and resources to even to manage or even in the language of the report to prevent migrations, (ahem) um, is a little bit, I think
full caption & clip details
A young adult feminine voice; delivery is subdued, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as concentration, hope enthusiasm optimism, fear; style: casual, monologue; average recording, quiet background; genuineness 1.8/6; vocal-burst blend 2.9/10; 48.5s.
batch203_part1_batch203_part1_chunk_294_1_176642 · in -27.1 dBFS · gain +7.1 dB · snippets-00539
Pleasure Ecstasy ↑ (unconstrained axis: Impatience and Irritability)c-snippets-B1 · #12
This chain comes from the one-sided rule: only Pleasure Ecstasy had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Pleasure Ecstasy clearly present — 0.73, higher than 73 % of clips in this corpus — and ends with it at the very top of the corpus at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.26.
Nothing was asked of the other axis, and in fact Impatience and Irritability climbs from 0.86 to 0.94 (+0.08), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.22, then -0.15, then +0.18 — not a clean run: step 2 moves back the other way by 0.15 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 34 s · snippets
k 4d_a 0.081d_b 0.259step_a 0.077step_b 0.225min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch38_part3_batch38_parttrack batch38_part3_batch38_parttotal 34.5slevel spread 3.3 dBmax seam 3.3 dB
Script — 4 chunks, 1 with a non-speech sound
Unchanged across all 4 clips: a young adult feminine voice · some disfluency
(normal-paced, normally alert, slightly relaxed, casual)And a lot of people said, oh, should have won, so this time I need to make it clear.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, dark, fairly smooth, thin; slurred, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; no dominant emotion; style: casual, conversational; average recording, quiet background; genuineness 4.0/6; vocal-burst blend 5.5/10; 3.6s.
batch38_part3_batch38_part3_chunk_1341_1_1041934 · in -30.6 dBFS · gain +10.6 dB · snippets-01088
(intoxication altered states of consciousness, amusement, fatigue exhaustion·brisk, energised, neutral tension, casual)brain and ensure that we're not sneaking that block and get high as fuck. Oh yeah, I've got a video from just around here, yeah, these crackheads.
full caption & clip details
A young adult feminine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, slightly bright, fairly smooth, thin; somewhat unclear, some disfluency, wide pitch range, normal breath; affect is positive, slightly dominant, guarded; reads as intoxication altered states of consciousness, amusement, fatigue exhaustion; style: casual, dramatic; below-average recording, some background noise; mildly explicit content; genuineness 4.3/6; vocal-burst blend 7.4/10; 7.7s.
batch38_part3_batch38_part3_chunk_1341_1_1041969 · in -27.2 dBFS · gain +7.2 dB · snippets-01088
(intoxication altered states of consciousness, confusion· brisk, highly aroused, slightly tense, casual)It begins now! The beginning, oh wait, wait, wait. The beginning of my peak is now!
full caption & clip details
A child somewhat masculine voice; delivery is highly aroused, brisk, slightly tense, moderately variable; timbre is slightly cool, slightly bright, rough, thin; slurred, some disfluency, wide pitch range, normal breath; affect is positive, slightly dominant, guarded; reads as intoxication altered states of consciousness, confusion; style: casual, cartoonish; below-average recording, some background noise; genuineness 2.9/6; vocal-burst blend 4.6/10; 5.7s.
batch38_part3_batch38_part3_chunk_1341_1_1042002 · in -27.8 dBFS · gain +7.8 dB · snippets-01088
(contentment, pleasure ecstasy, fatigue exhaustion· brisk, energised, neutral tension, casual)Yeah, my life's changed so much since I used to live here, like it's crazy. Like I've got a flat now up in, (low mumble) um, up in a completely different area. This area's literally crackhead central in Coventry, I'm telling you now, mate. I'm surprised, yeah, that there's not a group of fucking smokers stood here.
full caption & clip details
A young adult feminine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, slightly bright, fairly smooth, slightly thin; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly dominant, slightly guarded; reads as contentment, pleasure ecstasy, fatigue exhaustion; style: casual, playful; average recording, quiet background; genuineness 3.9/6; vocal-burst blend 8.5/10; 17.0s.
batch38_part3_batch38_part3_chunk_1341_1_1042017 · in -28.5 dBFS · gain +8.5 dB · snippets-01088
This chain comes from the one-sided rule: only Pain had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Pain clearly present — 0.72, higher than 72 % of clips in this corpus — and ends with it at the very top of the corpus at 0.96, higher than 96 % of clips in this corpus. That is a total rise of 0.24.
Nothing was asked of the other axis, and in fact Longing drifts down from 0.95 to 0.73 (-0.22), which the rule did not require.
It takes 2 clips to get there. Clip to clip the moves are +0.24 — a single step, so there is no internal shape to speak of.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 9 s · snippets
k 2d_a -0.224d_b 0.243step_a 0.224step_b 0.243min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch266_part3_batch266_patrack batch266_part3_batch266_patotal 8.9slevel spread 2.1 dBmax seam 2.1 dB
Script — 2 chunks, 0 with a non-speech sound
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, good recording, no background noise, normal-paced, normally alert, slightly relaxed
(longing · steady, no disfluency, fairly narrow pitch, narration)gone were the days when the kingdom of Leon Castile could lord over the Taifa Emirates.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, very full; clear, no disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as longing; style: narration, formal; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 1.9/10; 4.5s.
batch266_part3_batch266_part3_chunk_860_1_881724 · in -24.1 dBFS · gain +4.1 dB · snippets-00866
(pain, sadness·fairly steady, almost no disfluency, moderate pitch range, monologue)These territories were small and existed within the heart of Muslim Spain.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as pain, sadness; style: monologue, narration; good recording, no background noise; genuineness 1.0/6; vocal-burst blend 3.7/10; 4.3s.
batch266_part3_batch266_part3_chunk_860_1_881733 · in -22.0 dBFS · gain +2.0 dB · snippets-00866
This chain comes from the one-sided rule: only Interest had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Interest around average — 0.43, lower than 57 % of clips in this corpus — and ends with it at the very top of the corpus at 0.96, higher than 96 % of clips in this corpus. That is a total rise of 0.53.
Nothing was asked of the other axis, and in fact Fatigue Exhaustion drifts down from 0.92 to 0.22 (-0.70), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are -0.06, then +0.14, then +0.24, then +0.21 — not a clean run: step 1 moves back the other way by 0.06 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 52 s · snippets
k 5d_a -0.697d_b 0.528step_a 0.702step_b 0.241min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch281_part3_batch281_patrack batch281_part3_batch281_patotal 52.2slevel spread 8.6 dBmax seam 6.2 dB
Script — 5 chunks, 2 with a non-speech sound
Unchanged across all 5 clips: a young adult somewhat masculine voice
(fatigue exhaustion · measured, very low-energy, relaxed, casual)familiar with, so, we'll start with him. (surprised gasp) (low mumble) Uh, he does kind of air on the scary side of his stories if
full caption & clip details
A young adult somewhat masculine voice; delivery is very low-energy, measured, relaxed, moderately variable; timbre is neutral-toned, dark, fairly smooth, slightly thin; somewhat unclear, frequent disfluency, fairly narrow pitch, normal breath; affect is mildly negative, submissive, neutral openness; reads as fatigue exhaustion; style: casual, monologue; below-average recording, quiet background; genuineness 4.4/6; vocal-burst blend 1.7/10; 10.8s.
batch281_part3_batch281_part3_chunk_9_1_282957 · in -27.8 dBFS · gain +7.8 dB · snippets-00950
(contemplation, emotional numbness·slow, very low-energy, relaxed, whispered)listening to text to speech can be somewhat monotonous, but, (ahem) uh, over the last couple years.
full caption & clip details
An adult masculine voice; delivery is very low-energy, slow, relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, slightly thin; slurred, frequent disfluency, narrow pitch range, normal breath; affect is mildly negative, submissive, neutral openness; reads as contemplation, emotional numbness; style: whispered, casual; average recording, quiet background; genuineness 3.6/6; vocal-burst blend 0.8/10; 9.1s.
batch281_part3_batch281_part3_chunk_9_1_283153 · in -31.2 dBFS · gain +11.2 dB · snippets-00950
(teasing, amusement, malevolence malice·measured, normally alert, slightly relaxed, storytelling)Wait, I know. I'm going to trick or treat while it's COVID-19. Ha ha ha ha ha ha ha ha. But first, I'm going to dress as evil Marcelo.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, moderately variable; timbre is neutral-toned, slightly dark, slightly rough, balanced body; slurred, almost no disfluency, wide pitch range, minimal breath; affect is neutral, neutral stance, neutral openness; reads as teasing, amusement, malevolence malice; style: storytelling, playful; good recording, no background noise; genuineness 1.8/6; vocal-burst blend 0.0/10; 9.8s.
batch281_part3_batch281_part3_chunk_9_1_283393 · in -26.9 dBFS · gain +6.9 dB · snippets-00950
(concentration, fatigue exhaustion·slow, very low-energy, relaxed, whispered)figure out where the holes need to be. I'm just going to go grab some clamps. So I will make sure that these are lined up how I want them to be.
full caption & clip details
A child feminine voice; delivery is very low-energy, slow, relaxed, moderately variable; timbre is slightly cool, slightly bright, smooth, slightly thin; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly positive, slightly submissive, neutral openness; reads as concentration, fatigue exhaustion; style: whispered, ASMR; average recording, quiet background; genuineness 2.2/6; vocal-burst blend 0.0/10; 12.8s.
batch281_part3_batch281_part3_chunk_9_1_283612 · in -33.1 dBFS · gain +13.1 dB · snippets-00950
(interest, thankfulness gratitude, concentration ·normal-paced, normally alert, slightly relaxed, whispered)you'll be able to answer questions like, how would I rate the risk of bias and the quality of this research? And with this effect, how confident I am in applying the findings into practice.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as interest, thankfulness gratitude, concentration; style: whispered, formal; good recording, no background noise; genuineness 0.3/6; vocal-burst blend 0.7/10; 9.2s.
batch281_part3_batch281_part3_chunk_9_1_284380 · in -35.5 dBFS · gain +15.5 dB · snippets-00950
This chain comes from the one-sided rule: only Thankfulness Gratitude had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Thankfulness Gratitude around average — 0.57, higher than 57 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.43.
Nothing was asked of the other axis, and in fact Affection climbs from 0.77 to 0.98 (+0.21), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.16, then -0.02, then +0.13, then +0.16 — not a clean run: step 2 moves back the other way by 0.02 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 43 s · snippets
k 5d_a 0.208d_b 0.428step_a 0.380step_b 0.161min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch218_part3_batch218_patrack batch218_part3_batch218_patotal 43.1slevel spread 7.7 dBmax seam 3.2 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an adult feminine voice · normally alert, fairly steady, moderate pitch range
(measured, slightly relaxed, almost no disfluency, whispered)talk to some of the 2000 plus New Zealanders who might be into sex.
full caption & clip details
An adult feminine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is slightly cool, neutral-bright, fairly smooth, thin; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: whispered, storytelling; good recording, no background noise; genuineness 1.3/6; vocal-burst blend 1.5/10; 5.0s.
batch218_part3_batch218_part3_chunk_429_1_78372 · in -25.8 dBFS · gain +5.8 dB · snippets-00612
(contempt, interest, concentration·normal-paced, slightly relaxed, some disfluency, monologue)Money's theory was that gender was mostly about nurture, about the way you're raised, that it's not about nature, whatever happened to you prenatally. He was using intersex to sort of prove that theory, and he was trying to show that as long as you made a child look typically male mostly or typically female mostly, that you could end up with a straight boy or a straight girl who had no gender issues and no sexual orientation unusualness.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as contempt, interest, concentration; style: monologue, casual; good recording, no background noise; genuineness 1.0/6; vocal-burst blend 2.0/10; 23.5s.
batch218_part3_batch218_part3_chunk_429_1_78443 · in -27.6 dBFS · gain +7.6 dB · snippets-00612
(doubt, contemplation, longing·measured, slightly relaxed, some disfluency, casual)of that photography was that difference about me?
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as doubt, contemplation, longing; style: casual, monologue; average recording, no background noise; genuineness 2.8/6; vocal-burst blend 1.1/10; 3.8s.
batch218_part3_batch218_part3_chunk_429_1_78588 · in -29.9 dBFS · gain +9.8 dB · snippets-00612
(awe, infatuation, sexual lust· measured, neutral tension, some disfluency, casual)gorgeous brown eyes. Next thing I knew is sitting next to me we're making out in a circle.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, neutral tension, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, some disfluency, moderate pitch range, normal breath; affect is mildly negative, slightly dominant, neutral openness; reads as awe, infatuation, sexual lust; style: casual, conversational; average recording, quiet background; mildly explicit content; genuineness 3.3/6; vocal-burst blend 2.2/10; 6.5s.
batch218_part3_batch218_part3_chunk_429_1_78793 · in -30.2 dBFS · gain +10.2 dB · snippets-00612
(thankfulness gratitude, sadness, longing·normal-paced, slightly relaxed, frequent disfluency, casual)I think it's like the greatest gift my parents ever gave me.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, thin; average clarity, frequent disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as thankfulness gratitude, sadness, longing; style: casual, whispered; good recording, no background noise; genuineness 3.2/6; vocal-burst blend 3.1/10; 3.7s.
batch218_part3_batch218_part3_chunk_429_1_78840 · in -33.4 dBFS · gain +13.4 dB · snippets-00612
Sexual Lust ↑ (unconstrained axis: Teasing)c-snippets-B1 · #16
This chain comes from the one-sided rule: only Sexual Lust had to get where it was going, by at least 0.25. The other emotion was left completely free.
The chain starts with Sexual Lust clearly present — 0.68, higher than 68 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.32.
Nothing was asked of the other axis, and in fact Teasing climbs from 0.83 to 0.95 (+0.13), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.21, then +0.11, then -0.00 — not a clean run: step 3 moves back the other way by 0.00 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 15 s · snippets
k 4d_a 0.127d_b 0.323step_a 0.463step_b 0.209min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch119_part2_batch119_patrack batch119_part2_batch119_patotal 14.9slevel spread 3.8 dBmax seam 3.8 dB
Script — 4 chunks, 1 with a non-speech sound
Unchanged across all 4 clips: an adult masculine voice · neutral-toned, fairly smooth, no background noise, light breath
(normal-paced, normally alert, slightly relaxed, casual)We work black pride. Okay, send it to New York right away.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; no dominant emotion; style: casual, conversational; good recording, no background noise; genuineness 2.4/6; vocal-burst blend 0.8/10; 3.0s.
batch119_part2_batch119_part2_chunk_2068_1_2013568 · in -29.5 dBFS · gain +9.5 dB · snippets-00100
(sourness, infatuation, jealousy and envy· normal-paced, normally alert, slightly relaxed, casual)gates, ghost sites, who's cooperating with us abroad, every operation.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, thin; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as sourness, infatuation, jealousy and envy; style: casual, storytelling; good recording, no background noise; genuineness 1.9/6; vocal-burst blend 1.5/10; 4.1s.
batch119_part2_batch119_part2_chunk_2068_1_2013602 · in -32.3 dBFS · gain +12.3 dB · snippets-00100
(sexual lust, infatuation, longing· normal-paced, normally alert, slightly relaxed, casual)I like to fuck somewhere in their new place.
full caption & clip details
A child feminine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, frequent disfluency, wide pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as sexual lust, infatuation, longing; style: casual, storytelling; good recording, no background noise; mildly explicit content; genuineness 3.0/6; vocal-burst blend 4.5/10; 4.0s.
batch119_part2_batch119_part2_chunk_2068_1_2013655 · in -28.5 dBFS · gain +8.5 dB · snippets-00100
(sexual lust, intoxication altered states of consciousness, fatigue exhaustion·slow, very low-energy, fully relaxed, storytelling)Call it a finishing touch. (wistful sigh)
full caption & clip details
A child feminine voice; delivery is very low-energy, slow, fully relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; slurred, frequent disfluency, wide pitch range, light breath; affect is mildly negative, slightly submissive, neutral openness; reads as sexual lust, intoxication altered states of consciousness, fatigue exhaustion; style: storytelling, casual; average recording, no background noise; genuineness 3.0/6; vocal-burst blend 3.8/10; 3.3s.
batch119_part2_batch119_part2_chunk_2068_1_2013666 · in -29.2 dBFS · gain +9.2 dB · snippets-00100
This chain comes from the one-sided rule: only Hope Enthusiasm Optimism had to get where it was going, by at least 0.25. The other emotion was left completely free.
The chain starts with Hope Enthusiasm Optimism barely there — 0.20, lower than 80 % of clips in this corpus — and ends with it strongly present at 0.87, higher than 87 % of clips in this corpus. That is a total rise of 0.67.
Nothing was asked of the other axis, and in fact Emotional Numbness drifts down from 0.95 to 0.05 (-0.90), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.24, then +0.03, then +0.22, then +0.18 — most of the change happening immediately, then levelling off.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 39 s · snippets
k 5d_a -0.900d_b 0.669step_a 0.633step_b 0.237min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch95_part3_batch95_parttrack batch95_part3_batch95_parttotal 38.9slevel spread 1.8 dBmax seam 1.8 dB
Script — 5 chunks, 4 with a non-speech sound
Unchanged across all 5 clips: a young adult feminine voice · fairly smooth, balanced body, moderately variable, wide pitch range
(emotional numbness · normal-paced, normally alert, slightly relaxed, casual)(ahem) uh by pimps, by procurers, by by the men who are
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, wide pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as emotional numbness; style: casual, storytelling; good recording, quiet background; genuineness 2.7/6; vocal-burst blend 2.3/10; 4.1s.
batch95_part3_batch95_part3_chunk_1856_1_1704030 · in -19.9 dBFS · gain -0.1 dB · snippets-01384
(astonishment surprise·brisk, energised, slightly relaxed, dramatic)That seems odd, but it shows the type of influence that these women had within the city.
full caption & clip details
An adult feminine voice; delivery is energised, brisk, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; clear, little disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, neutral openness; reads as astonishment surprise; style: dramatic, conversational; good recording, no background noise; genuineness 1.8/6; vocal-burst blend 2.7/10; 5.4s.
batch95_part3_batch95_part3_chunk_1856_1_1704043 · in -21.7 dBFS · gain +1.7 dB · snippets-01384
(jealousy and envy, amusement·normal-paced, normally alert, slightly relaxed, conversational)She would have a lot of times local businessmen, (low mumble) um, E.B. Daggett.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as jealousy and envy, amusement; style: conversational, casual; good recording, no background noise; genuineness 2.7/6; vocal-burst blend 4.3/10; 4.6s.
batch95_part3_batch95_part3_chunk_1856_1_1704152 · in -21.1 dBFS · gain +1.1 dB · snippets-01384
(impatience and irritability, triumph, disappointment·brisk, energised, neutral tension, casual)These essentially acted as indirect operating licenses. (ahem) Um, they weren't on the books, but everybody kind of knew what was going on. And I have so many letters to the editor in the Fort Worth newspapers from citizens going, "We're not stupid. We get what you're doing, stop it," which of course, the city of Fort Worth didn't do.
full caption & clip details
A young adult feminine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, slightly bright, fairly smooth, balanced body; clear, some disfluency, wide pitch range, normal breath; affect is mildly positive, slightly dominant, slightly guarded; reads as impatience and irritability, triumph, disappointment; style: casual, playful; average recording, quiet background; genuineness 2.3/6; vocal-burst blend 2.0/10; 16.6s.
batch95_part3_batch95_part3_chunk_1856_1_1704262 · in -20.6 dBFS · gain +0.6 dB · snippets-01384
(brisk, normally alert, neutral tension, dramatic)(ahem) Um, and a good income, which were both requirements if they wanted to start their own brothel. So I have a really good example.
full caption & clip details
An adult feminine voice; delivery is normally alert, brisk, neutral tension, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; clear, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; no dominant emotion; style: dramatic, authoritative; good recording, no background noise; genuineness 2.3/6; vocal-burst blend 0.8/10; 7.5s.
batch95_part3_batch95_part3_chunk_1856_1_1704286 · in -20.4 dBFS · gain +0.4 dB · snippets-01384
This chain comes from the one-sided rule: only Helplessness had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Helplessness strongly present — 0.76, higher than 76 % of clips in this corpus — and ends with it at the very top of the corpus at 1.00, virtually no clip in this corpus scores higher. That is a total rise of 0.23.
Nothing was asked of the other axis, and in fact Pride drifts down from 1.00 to 0.78 (-0.22), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.22, then -0.04, then +0.05 — not a clean run: step 2 moves back the other way by 0.04 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 19 s · snippets
k 4d_a -0.216d_b 0.234step_a 0.759step_b 0.225min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch105_part1_batch105_patrack batch105_part1_batch105_patotal 18.9slevel spread 23.5 dBmax seam 13.8 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: an adult strongly masculine voice
(pride, triumph, shame · slow, highly aroused, slightly relaxed, storytelling)I am the one who has earned the right to wear this crown.
full caption & clip details
An adult strongly masculine voice; delivery is highly aroused, slow, slightly relaxed, steady; timbre is warm, dark, very rough, very full; very clear, no disfluency, very wide pitch range, light breath; affect is negative, dominant, fairly guarded; reads as pride, triumph, shame; style: storytelling, dramatic; good recording, no background noise; genuineness 0.3/6; vocal-burst blend 0.4/10; 3.4s.
batch105_part1_batch105_part1_chunk_1944_1_1620771 · in -20.9 dBFS · gain +0.9 dB · snippets-00024
(impatience and irritability, distress, anger·brisk, highly aroused, slightly tense, ranting)Can't you see I'm trying to fix this? You're always bothering me, distracting me, asking me stupid questions!
full caption & clip details
An adult somewhat masculine voice; delivery is highly aroused, brisk, slightly tense, variable; timbre is slightly cool, slightly bright, very rough, thin; very clear, almost no disfluency, very wide pitch range, normal breath; affect is deeply negative, very dominant, guarded; reads as impatience and irritability, distress, anger; style: ranting, cartoonish; average recording, quiet background; mildly explicit content; genuineness 0.3/6; vocal-burst blend 0.1/10; 6.4s.
batch105_part1_batch105_part1_chunk_1944_1_1621041 · in -30.6 dBFS · gain +10.6 dB · snippets-00024
(pride, jealousy and envy, triumph·measured, very low-energy, slightly relaxed, whispered)Queen of this industry, no? My face, my body, my talent, all perfection, all divine.
full caption & clip details
A young adult somewhat feminine voice; delivery is very low-energy, measured, slightly relaxed, steady; timbre is slightly warm, neutral-bright, fairly smooth, thin; clear, little disfluency, fairly narrow pitch, light breath; affect is mildly negative, neutral stance, vulnerable; reads as pride, jealousy and envy, triumph; style: whispered, ASMR; good recording, no background noise; genuineness 1.8/6; vocal-burst blend 1.3/10; 4.8s.
batch105_part1_batch105_part1_chunk_1944_1_1621168 · in -44.4 dBFS · gain +24.4 dB · snippets-00024
(shame, helplessness, disappointment·fast, energised, neutral tension, casual)Darling, I know I shouldn't be saying this, but I just can't help myself.
full caption & clip details
A young adult feminine voice; delivery is energised, fast, neutral tension, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, thin; average clarity, some disfluency, wide pitch range, normal breath; affect is mildly negative, slightly dominant, vulnerable; reads as shame, helplessness, disappointment; style: casual, storytelling; average recording, quiet background; genuineness 3.5/6; vocal-burst blend 4.1/10; 3.8s.
batch105_part1_batch105_part1_chunk_1944_1_1621304 · in -31.1 dBFS · gain +11.1 dB · snippets-00024
Astonishment Surprise ↑ (unconstrained axis: Sexual Lust)c-snippets-B1 · #19
This chain comes from the one-sided rule: only Astonishment Surprise had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Astonishment Surprise around average — 0.52, higher than 52 % of clips in this corpus — and ends with it at the very top of the corpus at 0.96, higher than 96 % of clips in this corpus. That is a total rise of 0.44.
Nothing was asked of the other axis, and in fact Sexual Lust drifts down from 1.00 to 0.88 (-0.12), which the rule did not require.
It takes 3 clips to get there. Clip to clip the moves are +0.20, then +0.24 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
3 clips · 18 s · snippets
k 3d_a -0.118d_b 0.440step_a 0.091step_b 0.242min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch115_part1_batch115_patrack batch115_part1_batch115_patotal 18.3slevel spread 4.7 dBmax seam 4.7 dB
Script — 3 chunks, 1 with a non-speech sound
Unchanged across all 3 clips: an adult masculine voice · average recording, normally alert
(sexual lust, infatuation, malevolence malice · measured, slightly relaxed, fairly steady, whispered)especially before I send my recommendation to look whether for your shoes.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is slightly warm, dark, very rough, very full; somewhat unclear, no disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, fairly guarded; reads as sexual lust, infatuation, malevolence malice; style: whispered, monologue; average recording, no background noise; genuineness 2.0/6; vocal-burst blend 5.4/10; 3.3s.
batch115_part1_batch115_part1_chunk_2027_1_1466512 · in -36.2 dBFS · gain +16.2 dB · snippets-00079
(impatience and irritability, jealousy and envy, anger· measured, neutral tension, moderately variable, storytelling)Yes, this is personal, and it should be personal for you too. That's my friend. I'm asking you to do this for me. Sorry, no. Well, fuck you! Girl went missing last night.
full caption & clip details
An elderly masculine voice; delivery is normally alert, measured, neutral tension, moderately variable; timbre is slightly cool, slightly dark, slightly rough, balanced body; average clarity, almost no disfluency, wide pitch range, audible breath; affect is negative, slightly dominant, vulnerable; reads as impatience and irritability, jealousy and envy, anger; style: storytelling, narration; average recording, quiet background; genuineness 1.3/6; vocal-burst blend 1.7/10; 8.1s.
batch115_part1_batch115_part1_chunk_2027_1_1466562 · in -33.9 dBFS · gain +13.9 dB · snippets-00079
(astonishment surprise, embarrassment, teasing·normal-paced, neutral tension, moderately variable, conversational)Yeah, yeah, I was just (ahem) um, picking up some things from evidence, my wallet, watch, chapstick. I could have brought that to you. Oh, it's cool, it's cool, I got it.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as astonishment surprise, embarrassment, teasing; style: conversational, casual; average recording, quiet background; genuineness 4.6/6; vocal-burst blend 2.2/10; 6.6s.
batch115_part1_batch115_part1_chunk_2027_1_1466610 · in -38.6 dBFS · gain +18.6 dB · snippets-00079
This chain comes from the one-sided rule: only Infatuation had to get where it was going, by at least 0.20. The other emotion was left completely free.
The chain starts with Infatuation around average — 0.57, higher than 57 % of clips in this corpus — and ends with it at the very top of the corpus at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.43.
Nothing was asked of the other axis, and in fact Disappointment drifts down from 0.91 to 0.39 (-0.52), which the rule did not require.
It takes 3 clips to get there. Clip to clip the moves are +0.24, then +0.18 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the snippets clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
3 clips · 26 s · snippets
k 3d_a -0.520d_b 0.426step_a 0.520step_b 0.245min_cos_consec —min_cos_anchor —dataset snippetslang ?speaker batch140_part2_batch140_patrack batch140_part2_batch140_patotal 26.1slevel spread 15.9 dBmax seam 14.1 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: an adult feminine voice · neutral-toned
(disappointment · normal-paced, normally alert, slightly relaxed, monologue)The welfare to work legislative reforms of the now previous government have seen major changes in what's available to us as taxpayers if we're disabled by serious illness.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as disappointment; style: monologue, formal; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.4/10; 11.3s.
batch140_part2_batch140_part2_chunk_2260_1_2110601 · in -21.8 dBFS · gain +1.8 dB · snippets-00213
(fear·measured, normally alert, slightly relaxed, whispered)If they attempted to go down, they would be swamped by the meeting of the waves. If they attempted to come up.
full caption & clip details
A middle-aged somewhat feminine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as fear; style: whispered, monologue; good recording, no background noise; genuineness 1.1/6; vocal-burst blend 0.5/10; 6.8s.
batch140_part2_batch140_part2_chunk_2260_1_2110630 · in -23.6 dBFS · gain +3.6 dB · snippets-00213
(infatuation, sexual lust, awe·slow, lethargic, relaxed, whispered)delivering all to executors pale the lazy yawning drone. I this infer.
full caption & clip details
An elderly masculine voice; delivery is lethargic, slow, relaxed, steady; timbre is neutral-toned, very dark, rough, thin; slurred, frequent disfluency, narrow pitch range, audible breath; affect is neutral, submissive, neutral openness; reads as infatuation, sexual lust, awe; style: whispered, ASMR; below-average recording, quiet background; genuineness 1.0/6; vocal-burst blend 0.2/10; 7.7s.
batch140_part2_batch140_part2_chunk_2260_1_2110782 · in -37.8 dBFS · gain +17.8 dB · snippets-00213