Manifest tier. emotion, rule B1, T=0.5, step cap 0.25. Population 383,880 chains (4,595 h) over 6 corpora. The SHAREABLE variant of this tier (podcast and evasnippets excluded) holds 318,030.
Rule.B1 — one-sided: emotion B rises by >=T; the other axis is unconstrained Source. trajectories_v5.parquet | Family. the tier owner's manifest tiers -- the exact subsets used for training Manifest tier.emotion__B1__T0.50__C0.25__INTERNAL — population 383,880 chains (4,595 h). SHAREABLE variant: 318,030. Filter.rule=='B1' and T==0.2 and C==0.25 and dataset in ['mls', 'eurospeech', 'emolia', 'podcast', 'snippets', 'evasnippets'] and abs(d_b)>=0.5 and step_b<=0.25 Sampled from 59,271 matching rows, without replacement across the family, so no two tiers reuse a chain.
How to read a Script. Each chunk is one line:
a short tag of what the models heard in that clip, then the words spoken.
(underlined, plain ·
delivery, style) — the tag before the words. Emotions first, then how
it is delivered. Underlined descriptors are the ones that change across this chain
— anything identical on every clip is pulled out and stated once above, because a
value that never moves says nothing about a trajectory.
(ahem) — brackets inside the words
are a different thing: a real non-speech sound, printed where it happens. Most clips have
none; about a quarter do.
The full generated caption for any clip is still there, under
“full caption & clip details”. Its perceived-gender and background-noise
clauses were re-rendered from the numeric buckets, because the versions stored in the
corpus index had those two ladders running backwards.
The emotion clause has been re-derived, and it used to be wrong. The 40 emotion heads are not on a common scale — Interest has a median of 2.08 and is never zero, while Infatuation is zero on 87.7 % of clips — and the caption named an emotion whenever its raw score cleared an absolute 1.0. Interest therefore appeared in 94.8 % of captions and Sadness in almost none: the clause was reporting the scale of the head, not the emotion of the clip. An emotion is now named only when it lands in the top 10 % for that emotion, against a pooled tie-aware mid-rank ECDF over 132,833,726 utterances spanning every dataset and language. Interest now appears in 6.2 %, all 40 emotions occur, and a clip that is ordinary on all 40 says “no dominant emotion” rather than being forced to pick one (21.8 % of clips). This is the same scale the trajectory miner selects on, so the caption and the mining now refer to the same quantity: the mined target emotion is named in the final clip's caption on 73 % of chains, up from 46 %.
This chain comes from the one-sided rule: only Hope Enthusiasm Optimism had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Hope Enthusiasm Optimism around average — 0.45, lower than 55 % of clips in this corpus — and ends with it at the very top of the corpus at 0.96, higher than 96 % of clips in this corpus. That is a total rise of 0.52.
Nothing was asked of the other axis, and in fact Astonishment Surprise drifts down from 0.80 to 0.08 (-0.71), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.18, then +0.22, then +0.15, then -0.03 — not a clean run: step 4 moves back the other way by 0.03 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 61 s · zh · emolia
hear it un-normalised (raw levels, max seam 1.4 dB)
k 5d_a -0.714d_b 0.516step_a 0.522step_b 0.223min_cos_consec —min_cos_anchor —dataset emolialang zhspeaker ZH_B00066_S00718track ZH_B00066_S00718total 61.1slevel spread 2.3 dBmax seam 1.4 dB
Script — 5 chunks, 4 with a non-speech sound
Unchanged across all 5 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, normal-paced, normally alert, moderate pitch range
(relaxed, fairly steady, frequent disfluency, casual)So something a second train in that first book that comes back and bints them (ahem) on the bomb.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; no dominant emotion; style: casual, conversational; average recording, quiet background; genuineness 4.3/6; vocal-burst blend 1.4/10; 7.5s, ZH.
ZH_B00066_S00718_W000075 · in -19.2 dBFS · gain -0.8 dB · emolia-03940
(relief, shame, embarrassment·neutral tension, moderately variable, some disfluency, conversational)(low mumble) Uh what funn ough gh finfinish the first drafts (low mumble) i know so all this stuffone talking about. So all one's grgrsand, as in it happens, s it pretty will get taken out. (ahem) Uh.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; reads as relief, shame, embarrassment; style: conversational, casual; average recording, quiet background; genuineness 4.9/6; vocal-burst blend 5.8/10; 9.0s, ZH.
ZH_B00066_S00718_W000076 · in -18.4 dBFS · gain -1.6 dB · emolia-03940
(interest, affection, embarrassment ·slightly relaxed, fairly steady, some disfluency, conversational)It's really interesting that isn't because i didn't want to write a comedy book, and you need a kids. I'm in i love with tencer. I'm going an absolutely do do with tencer call. I mean like crazy, (low mumble) but you know, that's a pure comic book, (low mumble) and i love some one like michael frain.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, slightly guarded; reads as interest, affection, embarrassment; style: conversational, casual; average recording, quiet background; genuineness 3.4/6; vocal-burst blend 5.6/10; 15.0s, ZH.
ZH_B00066_S00718_W000077 · in -19.8 dBFS · gain -0.2 dB · emolia-03940
(affection, hope enthusiasm optimism, infatuation·neutral tension, moderately variable, some disfluency, casual)I don't i do stuff because i love popular culture, and i love mainstream culture, and i want to a be in the middle of it. And that's why why i've always always been telly because i wanna be in lele's living room. Give give them that, i think theyygonna and give give something their kids are gna like. That's what i like in terms of write in a book.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as affection, hope enthusiasm optimism, infatuation; style: casual, conversational; good recording, quiet background; genuineness 4.4/6; vocal-burst blend 10.0/10; 16.3s, ZH.
ZH_B00066_S00718_W000078 · in -20.7 dBFS · gain +0.7 dB · emolia-03940
(hope enthusiasm optimism, pleasure ecstasy, interest· neutral tension, fairly steady, some disfluency, conversational)I would write whether i would sit down and (ahem) do呃,like like a definitively comic novel. I don't know, because i you know, i love the idea. I love crime, books and thrillers (low mumble) uh because they are popular.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, neutral tension, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as hope enthusiasm optimism, pleasure ecstasy, interest; style: conversational, casual; good recording, quiet background; genuineness 4.9/6; vocal-burst blend 9.0/10; 12.8s, ZH.
ZH_B00066_S00718_W000079 · in -19.9 dBFS · gain -0.1 dB · emolia-03940
This chain comes from the one-sided rule: only Contemplation had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Contemplation below average — 0.33, lower than 67 % of clips in this corpus — and ends with it at the very top of the corpus at 0.93, higher than 93 % of clips in this corpus. That is a total rise of 0.61.
Nothing was asked of the other axis, and in fact Infatuation drifts down from 0.91 to 0.51 (-0.41), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.25, then +0.15, then +0.21 — a fairly even climb, though some clips carry more of the change than others.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? Not directly measured. What does exist is a timbre similarity of 0.62 against the first clip, which describes how alike the voices sound rather than whether they are the same person. It sits on a different scale from the identity check (corpus-wide the timbre numbers run much higher), so it cannot be read against the 0.80 identity threshold and is given here without a pass or fail.
Voice consistency: these clips are separate recordings joined together. Speaker identity was not measured for this sample; the available timbre similarity of 0.62 says the voices sound broadly alike but is not a same-person check. You may notice the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 17 s · zh · emolia
hear it un-normalised (raw levels, max seam 2.4 dB)
k 4d_a -0.407d_b 0.605step_a 0.866step_b 0.246min_cos_consec —min_cos_anchor —dataset emolialang zhspeaker ZH_B00010_S01926track ZH_B00010_S01926total 16.6slevel spread 2.4 dBmax seam 2.4 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: a young adult masculine voice · neutral-toned, fairly smooth, balanced body, normally alert, slightly relaxed, fairly steady, moderate pitch range, light breath
(infatuation · normal-paced, some disfluency, average clarity, casual)我为疫情过来的概率可能为零了。就今年。
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, dark, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as infatuation; style: casual, conversational; average recording, quiet background; genuineness 3.4/6; vocal-burst blend 4.4/10; 3.0s, ZH.
ZH_B00010_S01926_W000031 · in -16.4 dBFS · gain -3.5 dB · emolia-03378
(measured, some disfluency, somewhat unclear, monologue)到冬天的时候过来的概率,但是人们对于这种疫情的这种恐慌啊。
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: monologue, formal; average recording, no background noise; genuineness 1.3/6; vocal-burst blend 2.5/10; 5.9s, ZH.
ZH_B00010_S01926_W000032 · in -15.4 dBFS · gain -4.6 dB · emolia-03378
(normal-paced, no disfluency, clear, casual)我们在现场不只会有儿童的客户过来,也会有婚纱的客户过来。
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: casual, monologue; good recording, no background noise; genuineness 2.0/6; vocal-burst blend 2.9/10; 3.7s, ZH.
ZH_B00010_S01926_W000033 · in -17.8 dBFS · gain -2.2 dB · emolia-03378
(contemplation· normal-paced, some disfluency, average clarity, casual)我们不止现在做了儿童加盟,我们也做了婚纱加盟。所以说呢。
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as contemplation; style: casual, formal; average recording, no background noise; genuineness 3.3/6; vocal-burst blend 3.0/10; 3.5s, ZH.
ZH_B00010_S01926_W000034 · in -17.1 dBFS · gain -2.9 dB · emolia-03378
This chain comes from the one-sided rule: only Elation had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Elation around average — 0.46, lower than 54 % of clips in this corpus — and ends with it at the very top of the corpus at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.53.
Nothing was asked of the other axis, and in fact Disgust drifts down from 0.91 to 0.03 (-0.87), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.22, then +0.23, then -0.16, then +0.23 — not a clean run: step 3 moves back the other way by 0.16 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? The least similar clip scores 0.43 against the first clip, where 1.00 would mean an identical voice. That is below the 0.80 threshold the mining used — treat the “same speaker” claim here with caution. Neighbouring clips score at worst 0.50 against each other.
Voice consistency: these clips are separate recordings joined together and the match is loose (0.43, under the 0.80 threshold), so the voice may audibly change between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 63 s · en · podcast
hear it un-normalised (raw levels, max seam 1.6 dB)
Unchanged across all 5 clips: an adult masculine voice · neutral-bright, fairly smooth, energised, some disfluency, light breath
(disgust · brisk, neutral tension, fairly steady, casual)but don is rice bowl. That's a there's a whole kanji that just means rice bowl. So the kind that (low mumble) uh matsua is famous for is gyudong. So again, the wa gu, the gyu means cow, gyu don
full caption & clip details
An adult masculine voice; delivery is energised, brisk, neutral tension, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as disgust; style: casual, authoritative; good recording, quiet background; genuineness 2.6/6; vocal-burst blend 4.9/10; 10.4s, EN.
325346_00192336 · in -15.7 dBFS · gain -4.3 dB · podcast-04385
(malevolence malice, disgust, contempt·normal-paced, slightly relaxed, steady, monologue)beef rice bowl, right? (low mumble) Um but in addition to that, they have all kinds of beef curry. So that same era that Japan got katsu from France, they got curry,
full caption & clip details
An adult masculine voice; delivery is energised, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, some disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as malevolence malice, disgust, contempt; style: monologue, didactic; good recording, no background noise; genuineness 0.9/6; vocal-burst blend 0.6/10; 10.7s, EN.
325346_00193368 · in -16.2 dBFS · gain -3.8 dB · podcast-04404
(interest, hope enthusiasm optimism, elation· normal-paced, neutral tension, moderately variable, casual)(ahem) uh, I believe through Britain through India, yeah. Kind of across that line. (ahem) Uh it's very different from Indian curry. Very different, but it's delicious. It's absurdly delicious. And so the thing about matsuya is it kind of feels like you're in a waffle house. Um
full caption & clip details
A young adult masculine voice; delivery is energised, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly dominant, slightly guarded; reads as interest, hope enthusiasm optimism, elation; style: casual, conversational; average recording, quiet background; mildly explicit content; genuineness 3.8/6; vocal-burst blend 5.6/10; 13.2s, EN.
325346_00194432 · in -14.6 dBFS · gain -5.5 dB · podcast-00081
(teasing, astonishment surprise, doubt· normal-paced, fully relaxed, moderately variable, casual)That's a good point. Wouldn't you like waffle house steak for some reason? Waffle
full caption & clip details
A young adult masculine voice; delivery is energised, normal-paced, fully relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, thin; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly dominant, neutral openness; reads as teasing, astonishment surprise, doubt; style: casual, conversational; average recording, some background noise; genuineness 3.9/6; vocal-burst blend 2.1/10; 3.5s, EN.
325346_00196880 · in -14.1 dBFS · gain -5.9 dB · podcast-00171
(elation, hope enthusiasm optimism, interest·brisk, neutral tension, moderately variable, casual)I'm not cutting this thing. It's real. And it's good. But the point standing. Matsya curry, as well as the other Don places, (low mumble) are set up a lot, like Waffle House. There's there's sort of (ahem) um a counter around a central cooking area and these stools, and you order everything on a vending machine once again. (ahem) Uh, and you take it and set it down, and moments later they bring you your curry. And we would finish almost every night there. About 350 yen, so about three dollars. It was absurdly cheap.
full caption & clip details
A young adult masculine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly dominant, slightly guarded; reads as elation, hope enthusiasm optimism, interest; style: casual, playful; average recording, some background noise; mildly explicit content; genuineness 3.8/6; vocal-burst blend 7.4/10; 24.5s, EN.
325346_00200376 · in -14.4 dBFS · gain -5.6 dB · podcast-00090
This chain comes from the one-sided rule: only Concentration had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Concentration around average — 0.46, lower than 54 % of clips in this corpus — and ends with it at the very top of the corpus at 0.97, higher than 97 % of clips in this corpus. That is a total rise of 0.52.
Nothing was asked of the other axis, and in fact Relief drifts down from 0.77 to 0.63 (-0.14), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.25, then +0.03, then +0.24 — most of the change happening immediately, then levelling off.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? Not directly measured. What does exist is a timbre similarity of 0.84 against the first clip, which describes how alike the voices sound rather than whether they are the same person. It sits on a different scale from the identity check (corpus-wide the timbre numbers run much higher), so it cannot be read against the 0.80 identity threshold and is given here without a pass or fail.
Voice consistency: these clips are separate recordings joined together. Speaker identity was not measured for this sample; the available timbre similarity of 0.84 says the voices sound broadly alike but is not a same-person check. You may notice the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 38 s · en · emolia
hear it un-normalised (raw levels, max seam 0.9 dB)
k 4d_a -0.142d_b 0.516step_a 0.499step_b 0.246min_cos_consec —min_cos_anchor —dataset emolialang enspeaker EN_8IEkdUHW4Eutrack EN_8IEkdUHW4Eutotal 37.6slevel spread 1.3 dBmax seam 0.9 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(fairly steady, formal, monologue)Residence is often defined for individuals as presence in the country for more than 183 days
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, monologue; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.4/10; 5.9s, EN.
EN_8IEkdUHW4Eu_W000076 · in -14.8 dBFS · gain -5.2 dB · emolia-02023
(fear·steady, formal, authoritative)Most countries base residence of entities on either place of organization or place of management and control. The United Kingdom has three levels of residence
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as fear; style: formal, authoritative; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.0/10; 9.3s, EN.
EN_8IEkdUHW4Eu_W000077 · in -13.9 dBFS · gain -6.1 dB · emolia-02023
(steady, formal, newsreading)Most systems define income subject to tax broadly for residents, but tax nonresidents only on specific types of income
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, newsreading; good recording, no background noise; genuineness 0.3/6; vocal-burst blend 0.2/10; 7.3s, EN.
EN_8IEkdUHW4Eu_W000079 · in -14.8 dBFS · gain -5.2 dB · emolia-02023
(concentration, emotional numbness· steady, newsreading, formal)Income generally includes most types of receipts that enrich the taxpayer, including compensation for services, gain from sale of goods or other property, interest, dividends, rents, royalties, annuities, pensions, and all manner of other items
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as concentration, emotional numbness; style: newsreading, formal; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.0/10; 14.6s, EN.
EN_8IEkdUHW4Eu_W000081 · in -15.2 dBFS · gain -4.8 dB · emolia-02023
This chain comes from the one-sided rule: only Emotional Numbness had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Emotional Numbness around average — 0.47, lower than 53 % of clips in this corpus — and ends with it at the very top of the corpus at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.52.
Nothing was asked of the other axis, and in fact Interest drifts down from 0.97 to 0.09 (-0.88), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.17, then +0.07, then +0.18, then +0.09 — an uneven climb, but always in the same direction.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? Not directly measured. What does exist is a timbre similarity of 0.67 against the first clip, which describes how alike the voices sound rather than whether they are the same person. It sits on a different scale from the identity check (corpus-wide the timbre numbers run much higher), so it cannot be read against the 0.80 identity threshold and is given here without a pass or fail.
Voice consistency: these clips are separate recordings joined together. Speaker identity was not measured for this sample; the available timbre similarity of 0.67 says the voices sound broadly alike but is not a same-person check. You may notice the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 57 s · en · emolia
hear it un-normalised (raw levels, max seam 2.3 dB)
k 5d_a -0.877d_b 0.516step_a 0.420step_b 0.182min_cos_consec —min_cos_anchor —dataset emolialang enspeaker EN_dZGGFLTJ9EMtrack EN_dZGGFLTJ9EMtotal 57.0slevel spread 4.3 dBmax seam 2.3 dB
Script — 5 chunks, 2 with a non-speech sound
Unchanged across all 5 clips: a young adult masculine voice · neutral-toned, fairly smooth, balanced body, quiet background, normally alert, slightly relaxed, fairly steady, moderate pitch range
(interest, concentration, contemplation · normal-paced, some disfluency, average clarity, didactic)Representing economics as a whole, discipline. It's really interesting that you can sum it up in four lines. So what's in a price? When you hear the word price, what do you think of? A price is just a mutually agreed upon
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as interest, concentration, contemplation; style: didactic, monologue; average recording, quiet background; genuineness 2.6/6; vocal-burst blend 0.8/10; 13.9s, EN.
EN_dZGGFLTJ9EM_W000022 · in -22.1 dBFS · gain +2.1 dB · emolia-01462
(concentration · normal-paced, some disfluency, average clarity, monologue)There's money to make transactions easier. We don't have to agree on the value of a cow versus a chicken. We have dollars or other currencies that we can agree upon the value of based on its relative price to other things. So in economics, everything starts with two axes.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as concentration; style: monologue, casual; average recording, quiet background; genuineness 1.9/6; vocal-burst blend 0.1/10; 18.5s, EN.
EN_dZGGFLTJ9EM_W000024 · in -20.2 dBFS · gain +0.2 dB · emolia-01462
(concentration · normal-paced, some disfluency, average clarity, didactic)(low mumble) Uh, (ahem) the price axis is based on like price per unit of things that you're selling and the quantity is the number of things that you're selling. When we graph supply, supply curves are always up and to the right.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as concentration; style: didactic, monologue; average recording, quiet background; genuineness 3.1/6; vocal-burst blend 0.1/10; 12.5s, EN.
EN_dZGGFLTJ9EM_W000026 · in -19.8 dBFS · gain -0.2 dB · emolia-01462
(normal-paced, some disfluency, average clarity, authoritative)They start at a point where you have to cover your costs unless you're doing some sort of weird pricing strategy where you're taking a loss on a product.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; no dominant emotion; style: authoritative, casual; good recording, quiet background; genuineness 1.2/6; vocal-burst blend 1.1/10; 7.2s, EN.
EN_dZGGFLTJ9EM_W000027 · in -20.1 dBFS · gain +0.1 dB · emolia-01462
(emotional numbness, doubt·measured, frequent disfluency, somewhat unclear, casual)(ahem) Uh, so it starts above the, above the, (low mumble) uhm, end of the...
full caption & clip details
A young adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; reads as emotional numbness, doubt; style: casual, monologue; average recording, quiet background; genuineness 3.3/6; vocal-burst blend 1.0/10; 4.4s, EN.
EN_dZGGFLTJ9EM_W000028 · in -17.9 dBFS · gain -2.1 dB · emolia-01462
This chain comes from the one-sided rule: only Concentration had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Concentration below average — 0.41, lower than 59 % of clips in this corpus — and ends with it at the very top of the corpus at 0.98, higher than 98 % of clips in this corpus. That is a total rise of 0.58.
Nothing was asked of the other axis, and in fact Interest drifts down from 0.96 to 0.91 (-0.05), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.08, then +0.24, then +0.14, then +0.12 — an uneven climb, but always in the same direction.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? The least similar clip scores 0.36 against the first clip, where 1.00 would mean an identical voice. That is below the 0.80 threshold the mining used — treat the “same speaker” claim here with caution. Neighbouring clips score at worst 0.36 against each other.
Voice consistency: these clips are separate recordings joined together and the match is loose (0.36, under the 0.80 threshold), so the voice may audibly change between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 5 clips: a young adult masculine voice · neutral-toned, fairly smooth, balanced body, normal-paced, normally alert, some disfluency, average clarity, moderate pitch range
(interest, hope enthusiasm optimism, disgust · neutral tension, moderately variable, casual)I guess they're technically building a convent, and it's all like non-violent, and (low mumble) uh the Bray's character has a violent past. He has been a fighter and killer, he even admits this kind of stuff. Worth noting that the they apparently wrote a two-page speech for Ian McShane's character here that they cut more or less from it. So I think maybe there would have been a little bit more of an explicit reference to the broken man within that.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as interest, hope enthusiasm optimism, disgust; style: casual; average recording, quiet background; genuineness 3.8/6; vocal-burst blend 6.2/10; 23.3s, EN.
805544_00062576 · in -28.5 dBFS · gain +8.5 dB · podcast-06108
(doubt, infatuation·slightly relaxed, fairly steady, casual, conversational)That he said, and seems like that was maybe one of the big reasons they wanted to hire a (low mumble) an actor of stature for this.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; reads as doubt, infatuation; style: casual, conversational; good recording, no background noise; genuineness 4.4/6; vocal-burst blend 3.1/10; 5.5s, EN.
805544_00064967 · in -29.8 dBFS · gain +9.8 dB · podcast-04466
(disappointment, emotional numbness· slightly relaxed, fairly steady, casual, dramatic)Right. Well, and there's even (contented sigh) the show s can't I mean it it doesn't have the same perspective as George R. Martin, because Brother Ray even says violence is a disease, you can't cure it by spreading it,
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as disappointment, emotional numbness; style: casual, dramatic; good recording, no background noise; genuineness 3.5/6; vocal-burst blend 1.8/10; 10.4s, EN.
805544_00065552 · in -23.6 dBFS · gain +3.6 dB · podcast-04470
(contemplation, fear, relief· slightly relaxed, moderately variable, casual, conversational)I think the pacifist might say, you do actually. You do cure it by dying. Like it may it may result in your own death, but at least you don't meet violence with violence.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as contemplation, fear, relief; style: casual, conversational; good recording, no background noise; genuineness 3.9/6; vocal-burst blend 5.0/10; 9.3s, EN.
805544_00066952 · in -24.4 dBFS · gain +4.4 dB · podcast-04477
(concentration, contemplation, doubt· slightly relaxed, fairly steady, casual, conversational)Right. And this episode very clearly sets up not necessarily like I don't think it's saying overarchingly that this is one right thing versus one wrong, but the hound certainly comes to the conclusion that you need violence to combat violence.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as concentration, contemplation, doubt; style: casual, conversational; good recording, quiet background; genuineness 4.1/6; vocal-burst blend 3.7/10; 13.9s, EN.
805544_00067920 · in -29.7 dBFS · gain +9.7 dB · podcast-04462
This chain comes from the one-sided rule: only Malevolence Malice had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Malevolence Malice barely there — 0.22, lower than 78 % of clips in this corpus — and ends with it at the very top of the corpus at 0.95, higher than 95 % of clips in this corpus. That is a total rise of 0.72.
Nothing was asked of the other axis, and in fact Disgust drifts down from 0.96 to 0.53 (-0.43), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.25, then +0.23, then +0.24 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the eurospeech clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 54 s · hr · eurospeech
k 4d_a -0.433d_b 0.723step_a 0.338step_b 0.249min_cos_consec —min_cos_anchor —dataset eurospeechlang hrspeaker croatia_20080424175417-345track croatia_20080424175417-345total 54.2slevel spread 3.5 dBmax seam 3.5 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: an adult masculine voice · slightly cool, neutral-bright, average recording, brisk, energised, neutral tension, moderately variable, wide pitch range
(disgust, contempt, pride · some disfluency, average clarity, normal breath, cartoonish)da bi ih mogli nastaviti proizvoditi ili uvoziti deseci tisuća proizvođača i uvoznika morat će već ove godine predregistrirati kemikalije uključujući kiseline, metale, razređivače, površinsko aktivne tvari, ljepila i
full caption & clip details
An adult masculine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, slightly rough, balanced body; average clarity, some disfluency, wide pitch range, normal breath; affect is neutral, slightly dominant, fairly guarded; reads as disgust, contempt, pride; style: cartoonish, ranting; average recording, quiet background; genuineness 2.0/6; vocal-burst blend 6.8/10; 15.1s, HR.
croatia_20080424175417-3458_1383904_1399008 · in -23.9 dBFS · gain +3.9 dB · eurospeech-01242
(thankfulness gratitude·almost no disfluency, average clarity, normal breath, authoritative)Europska kemijska agencija iz Helsinkija će u slijedećih jedanaest godina staviti registar tih tridesetak tisuća kemikalija. Registracija kemikalija vršit će se u nekoliko faza,
full caption & clip details
An adult masculine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, slightly rough, balanced body; average clarity, almost no disfluency, wide pitch range, normal breath; affect is neutral, slightly dominant, fairly guarded; reads as thankfulness gratitude; style: authoritative, dramatic; average recording, some background noise; genuineness 0.7/6; vocal-burst blend 2.3/10; 11.9s, HR.
croatia_20080424175417-3458_1399008_1410864 · in -22.7 dBFS · gain +2.7 dB · eurospeech-01242
(thankfulness gratitude, shame· almost no disfluency, somewhat unclear, audible breath, dramatic)ove godine predregistracija, 2010. registracija kemikalija proizvedenih i uvezenih u količinama preko tisuću tona, 2013. registracija količina od 100 do tisuću tona a 2008. registracija količina kemikalija iznad jedne
full caption & clip details
A young adult masculine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, thin; somewhat unclear, almost no disfluency, wide pitch range, audible breath; affect is neutral, slightly dominant, fairly guarded; reads as thankfulness gratitude, shame; style: dramatic, authoritative; average recording, some background noise; genuineness 0.9/6; vocal-burst blend 6.4/10; 16.5s, HR.
croatia_20080424175417-3458_1410864_1427376 · in -25.8 dBFS · gain +5.8 dB · eurospeech-01242
(malevolence malice, concentration·some disfluency, average clarity, normal breath, cartoonish)Troškovi registracije uključujući i testiranja procjenjuju se na 2,3 milijarde eura kroz jedanaest godina a efekti
full caption & clip details
An adult masculine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, slightly rough, balanced body; average clarity, some disfluency, wide pitch range, normal breath; affect is neutral, slightly dominant, fairly guarded; reads as malevolence malice, concentration; style: cartoonish, authoritative; average recording, quiet background; genuineness 2.6/6; vocal-burst blend 2.5/10; 10.2s, HR.
croatia_20080424175417-3458_1427376_1437616 · in -22.3 dBFS · gain +2.3 dB · eurospeech-01242
This chain comes from the one-sided rule: only Fatigue Exhaustion had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Fatigue Exhaustion around average — 0.42, lower than 58 % of clips in this corpus — and ends with it at the very top of the corpus at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.57.
Nothing was asked of the other axis, and in fact Contemplation drifts down from 0.96 to 0.52 (-0.44), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.22, then +0.14, then +0.00, then +0.21 — most of the change happening immediately, then levelling off.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? The least similar clip scores 0.36 against the first clip, where 1.00 would mean an identical voice. That is below the 0.80 threshold the mining used — treat the “same speaker” claim here with caution. Neighbouring clips score at worst 0.36 against each other.
Voice consistency: these clips are separate recordings joined together and the match is loose (0.36, under the 0.80 threshold), so the voice may audibly change between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 5 clips: a young adult masculine voice · neutral-toned, fairly smooth, balanced body, normal-paced, normally alert, average clarity, light breath
(contemplation, sexual lust, interest · relaxed, moderately variable, frequent disfluency, casual)I I get why she needs it, maybe like some depth to the storyline of maybe other people that can help. I get it. It was a terrible episode
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as contemplation, sexual lust, interest; style: casual, conversational; good recording, no background noise; genuineness 4.2/6; vocal-burst blend 7.6/10; 8.5s, EN.
562076_00195328 · in -22.1 dBFS · gain +2.1 dB · podcast-01314
(emotional numbness, amusement, teasing·neutral tension, moderately variable, some disfluency, casual)been. The other ten care the other ten before eleven could have just like literally died off in
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as emotional numbness, amusement, teasing; style: casual, conversational; average recording, quiet background; mildly explicit content; genuineness 5.2/6; vocal-burst blend 4.3/10; 4.8s, EN.
562076_00196728 · in -21.3 dBFS · gain +1.3 dB · podcast-01304
(slightly relaxed, fairly steady, some disfluency, casual)like so in experiments in which is what everybody thought.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; no dominant emotion; style: casual, conversational; good recording, no background noise; genuineness 4.0/6; vocal-burst blend 4.9/10; 3.1s, EN.
562076_00197216 · in -21.3 dBFS · gain +1.3 dB · podcast-00665
(embarrassment, shame, pleasure ecstasy·relaxed, moderately variable, some disfluency, casual)Yeah, and then I didn't I didn't. Yeah. It was just so unlikable. Like you could have done your side story, but have likable characters. Yeah, I have to rewatch that episode again. Like
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly negative, slightly submissive, neutral openness; reads as embarrassment, shame, pleasure ecstasy; style: casual, conversational; average recording, quiet background; mildly explicit content; genuineness 4.7/6; vocal-burst blend 10.0/10; 15.1s, EN.
562076_00197552 · in -21.8 dBFS · gain +1.8 dB · podcast-01300
(embarrassment, fatigue exhaustion, shame ·slightly relaxed, fairly steady, some disfluency, casual)and then just remind myself why I hated it so much. (low mumble) It's like a sit through.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as embarrassment, fatigue exhaustion, shame; style: casual, conversational; good recording, quiet background; genuineness 4.1/6; vocal-burst blend 4.4/10; 4.2s, EN.
562076_00199088 · in -21.8 dBFS · gain +1.8 dB · podcast-01303
This chain comes from the one-sided rule: only Embarrassment had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Embarrassment below average — 0.38, lower than 62 % of clips in this corpus — and ends with it at the very top of the corpus at 0.94, higher than 94 % of clips in this corpus. That is a total rise of 0.56.
Nothing was asked of the other axis, and in fact Thankfulness Gratitude drifts down from 0.99 to 0.55 (-0.45), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.24, then +0.12, then -0.00, then +0.20 — not a clean run: step 3 moves back the other way by 0.00 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? Not directly measured. What does exist is a timbre similarity of 0.67 against the first clip, which describes how alike the voices sound rather than whether they are the same person. It sits on a different scale from the identity check (corpus-wide the timbre numbers run much higher), so it cannot be read against the 0.80 identity threshold and is given here without a pass or fail.
Voice consistency: these clips are separate recordings joined together. Speaker identity was not measured for this sample; the available timbre similarity of 0.67 says the voices sound broadly alike but is not a same-person check. You may notice the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 42 s · ja · emolia
k 5d_a -0.446d_b 0.556step_a 0.877step_b 0.238min_cos_consec —min_cos_anchor —dataset emolialang jaspeaker JA_B00004_S01331track JA_B00004_S01331total 41.7slevel spread 7.0 dBmax seam 7.0 dB
Script — 5 chunks, 2 with a non-speech sound
Unchanged across all 5 clips: a young adult masculine voice · neutral-toned, fairly smooth, balanced body, average recording, normally alert, fairly steady, moderate pitch range, light breath
(thankfulness gratitude, affection · fast, slightly relaxed, some disfluency, conversational)他の産業を見てみれば、例えばですね、航空会社だったりとか、アナとか、ジャルとかだったりとかもね、
full caption & clip details
A young adult masculine voice; delivery is normally alert, fast, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as thankfulness gratitude, affection; style: conversational, casual; average recording, quiet background; genuineness 3.6/6; vocal-burst blend 2.2/10; 5.2s, JA.
JA_B00004_S01331_W000002 · in -22.7 dBFS · gain +2.7 dB · emolia-02994
(contemplation, confusion, doubt·measured, slightly relaxed, some disfluency, casual)こういう風の時代とかを乗り切るというふうな時に身につけて考え方が、あの、
full caption & clip details
A young adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; reads as contemplation, confusion, doubt; style: casual, monologue; average recording, quiet background; genuineness 4.0/6; vocal-burst blend 2.5/10; 5.4s, JA.
JA_B00004_S01331_W000003 · in -15.7 dBFS · gain -4.3 dB · emolia-02994
(infatuation, longing, sexual lust·normal-paced, slightly relaxed, some disfluency, casual)(low mumble) てことをやって、え、このですね、期間を生き延びるってことをお勧めしようかなと思っております。はい。もう一つ、私の方でお勧めさせてもらうことは何なのかっていうと、もちろんですけど、コロナの助成金とか保障ですね。
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as infatuation, longing, sexual lust; style: casual, conversational; average recording, quiet background; genuineness 4.4/6; vocal-burst blend 7.9/10; 12.2s, JA.
JA_B00004_S01331_W000004 · in -21.6 dBFS · gain +1.6 dB · emolia-02994
(pleasure ecstasy, infatuation, sexual lust ·fast, slightly relaxed, some disfluency, monologue)それが果たしてですね、節税の役割になっているのかどうかってことを考えたときに、実はあまり肥料体効果の高い投資でないと気づかれる方がいらっしゃるのであれば、
full caption & clip details
A young adult masculine voice; delivery is normally alert, fast, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as pleasure ecstasy, infatuation, sexual lust; style: monologue, casual; average recording, no background noise; genuineness 2.5/6; vocal-burst blend 6.5/10; 9.6s, JA.
JA_B00004_S01331_W000005 · in -19.8 dBFS · gain -0.2 dB · emolia-02994
This chain comes from the one-sided rule: only Concentration had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Concentration below average — 0.39, lower than 61 % of clips in this corpus — and ends with it at the very top of the corpus at 0.93, higher than 93 % of clips in this corpus. That is a total rise of 0.54.
Nothing was asked of the other axis, and in fact Emotional Numbness drifts down from 0.91 to 0.23 (-0.68), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.24, then -0.09, then +0.19, then +0.19 — not a clean run: step 2 moves back the other way by 0.09 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? Not directly measured. What does exist is a timbre similarity of 0.85 against the first clip, which describes how alike the voices sound rather than whether they are the same person. It sits on a different scale from the identity check (corpus-wide the timbre numbers run much higher), so it cannot be read against the 0.80 identity threshold and is given here without a pass or fail.
Voice consistency: these clips are separate recordings joined together. Speaker identity was not measured for this sample; the available timbre similarity of 0.85 says the voices sound broadly alike but is not a same-person check. You may notice the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 35 s · en · emolia
k 5d_a -0.679d_b 0.536step_a 0.709step_b 0.244min_cos_consec —min_cos_anchor —dataset emolialang enspeaker EN_yh__sq2RM8Qtrack EN_yh__sq2RM8Qtotal 34.6slevel spread 2.3 dBmax seam 1.5 dB
Script — 5 chunks, 3 with a non-speech sound
Unchanged across all 5 clips: a young adult feminine voice · average recording, slightly relaxed, somewhat unclear
(emotional numbness · measured, normally alert, fairly steady, monologue)There is no scope creep and uh, there is no certificate problems.
full caption & clip details
A young adult feminine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, thin; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as emotional numbness; style: monologue, formal; average recording, no background noise; genuineness 3.2/6; vocal-burst blend 1.3/10; 3.8s, EN.
EN_yh__sq2RM8Q_W000065 · in -16.6 dBFS · gain -3.4 dB · emolia-01813
(measured, normally alert, fairly steady, monologue)So we have another browsing traction which we can go through again and select an account.
full caption & clip details
A young adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: monologue, didactic; average recording, no background noise; genuineness 2.2/6; vocal-burst blend 2.0/10; 5.0s, EN.
EN_yh__sq2RM8Q_W000066 · in -16.5 dBFS · gain -3.5 dB · emolia-01813
(doubt·normal-paced, normally alert, fairly steady, monologue)As I said, this is running locally, so some of the tests do not (low mumble) validate out as we are not integrated with the OBI.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as doubt; style: monologue, didactic; average recording, quiet background; genuineness 3.2/6; vocal-burst blend 2.2/10; 6.8s, EN.
EN_yh__sq2RM8Q_W000068 · in -15.3 dBFS · gain -4.7 dB · emolia-01813
(slow, normally alert, fairly steady, monologue)That is the basic process of (low mumble) running the test suite. And in this case, (low mumble) for
full caption & clip details
A young adult feminine voice; delivery is normally alert, slow, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; no dominant emotion; style: monologue, casual; average recording, quiet background; genuineness 2.9/6; vocal-burst blend 0.9/10; 5.7s, EN.
EN_yh__sq2RM8Q_W000069 · in -14.3 dBFS · gain -5.7 dB · emolia-01813
(concentration·measured, subdued, steady, whispered)(low mumble) Or gaining conformance and (ahem) getting the accreditation from OBIE, you generally just run the test suite and then export, (low mumble) uh, export the test suite logs and send it to OBIE.
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is slightly cool, slightly dark, slightly rough, balanced body; somewhat unclear, frequent disfluency, fairly narrow pitch, audible breath; affect is neutral, neutral stance, slightly guarded; reads as concentration; style: whispered, monologue; average recording, quiet background; genuineness 2.8/6; vocal-burst blend 1.2/10; 12.8s, EN.
EN_yh__sq2RM8Q_W000070 · in -15.8 dBFS · gain -4.2 dB · emolia-01813
This chain comes from the one-sided rule: only Longing had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Longing barely there — 0.23, lower than 77 % of clips in this corpus — and ends with it strongly present at 0.84, higher than 84 % of clips in this corpus. That is a total rise of 0.60.
Nothing was asked of the other axis, and in fact Triumph drifts down from 0.93 to 0.25 (-0.68), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.22, then +0.20, then +0.18 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? Not directly measured. What does exist is a timbre similarity of 0.92 against the first clip, which describes how alike the voices sound rather than whether they are the same person. It sits on a different scale from the identity check (corpus-wide the timbre numbers run much higher), so it cannot be read against the 0.80 identity threshold and is given here without a pass or fail.
Voice consistency: these clips are separate recordings joined together. Speaker identity was not measured for this sample; the available timbre similarity of 0.92 says the voices sound broadly alike but is not a same-person check. You may notice the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 46 s · en · emolia
k 4d_a -0.681d_b 0.604step_a 0.483step_b 0.221min_cos_consec —min_cos_anchor —dataset emolialang enspeaker EN_9k40tiLrziMtrack EN_9k40tiLrziMtotal 46.3slevel spread 0.6 dBmax seam 0.5 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: a young adult feminine voice · neutral-toned, fairly smooth, good recording, no background noise, normally alert, slightly relaxed, no disfluency, clear
(triumph · normal-paced, fairly steady, light breath, formal)Their vacuum oil company began operating in Australia in 1895, introducing its plume brand of petrol in 1916
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as triumph; style: formal, monologue; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.0/10; 8.4s, EN.
EN_9k40tiLrziM_W000076 · in -14.7 dBFS · gain -5.3 dB · emolia-00314
(measured, steady, light breath, formal)The Flying Red Horse Pegasus logo was introduced in 1939, and in 1954 the plume brand was replaced by Mobilgas
full caption & clip details
A young adult feminine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, monologue; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.0/10; 9.2s, EN.
EN_9k40tiLrziM_W000077 · in -14.3 dBFS · gain -5.7 dB · emolia-00314
(normal-paced, steady, no audible breath, newsreading)Mobile Australia's corporate office is in Melbourne. In 1946 Mobile commenced construction of a refinery at Altona in Melbourne's western suburbs, which originally produced lubricating oils and bitumen, before producing motor vehicle fuels in 1956.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, no audible breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: newsreading, formal; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.0/10; 17.8s, EN.
EN_9k40tiLrziM_W000078 · in -14.8 dBFS · gain -5.2 dB · emolia-00314
(measured, steady, light breath, formal)It is still in use. A second refinery at Port Stanvac, south of Adelaide, came on stream in 1963, but was closed in 2003
full caption & clip details
A young adult feminine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, thin; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, monologue; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.0/10; 10.4s, EN.
EN_9k40tiLrziM_W000079 · in -14.9 dBFS · gain -5.1 dB · emolia-00314
This chain comes from the one-sided rule: only Disgust had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Disgust below average — 0.33, lower than 67 % of clips in this corpus — and ends with it at the very top of the corpus at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.66.
Nothing was asked of the other axis, and in fact Pride drifts down from 0.85 to 0.53 (-0.32), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.20, then +0.25, then -0.04, then +0.24 — not a clean run: step 3 moves back the other way by 0.04 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 40 s · zh · emolia
k 5d_a -0.320d_b 0.658step_a 0.320step_b 0.247min_cos_consec —min_cos_anchor —dataset emolialang zhspeaker ZH_B00066_S04994track ZH_B00066_S04994total 39.9slevel spread 2.2 dBmax seam 1.7 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normally alert, slightly relaxed
(normal-paced, fairly steady, almost no disfluency, formal)Jack wills was one of the three most important movie producers in hollywood owner of his own studio, with dozens of stars under contract.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, monologue; good recording, no background noise; genuineness 0.4/6; vocal-burst blend 0.3/10; 7.9s, ZH.
ZH_B00066_S04994_W000071 · in -17.3 dBFS · gain -2.7 dB · emolia-03936
(normal-paced, steady, no disfluency, formal)He was on the president of the united states advisory council for war information.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, authoritative; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.3/10; 3.9s, ZH.
ZH_B00066_S04994_W000072 · in -16.4 dBFS · gain -3.5 dB · emolia-03936
(measured, fairly steady, almost no disfluency, narration)Cinematic division, which meant simply that he helped propaganda movies he'd had dinner at the white house. He had entertain jared gurhover in his hollywood home, but none of this was as impressive as it sound.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; no dominant emotion; style: narration, formal; good recording, no background noise; genuineness 1.2/6; vocal-burst blend 3.0/10; 10.4s, ZH.
ZH_B00066_S04994_W000073 · in -18.2 dBFS · gain -1.8 dB · emolia-03936
(bitterness, malevolence malice, contempt· measured, steady, almost no disfluency, narration)They were all official relationships walls, didn't have any personal political car, mainly because he wasn't extreme reactionary.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as bitterness, malevolence malice, contempt; style: narration, formal; good recording, no background noise; mildly explicit content; genuineness 1.0/6; vocal-burst blend 0.3/10; 7.6s, ZH.
ZH_B00066_S04994_W000074 · in -17.4 dBFS · gain -2.6 dB · emolia-03936
(disgust, malevolence malice, awe·normal-paced, fairly steady, almost no disfluency, narration)Partly because he was a mclleniac who loved to wield power wildly without regard to the fact that by so doing, legions of enemies sprang up out of the ground.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as disgust, malevolence malice, awe; style: narration, monologue; good recording, no background noise; genuineness 0.6/6; vocal-burst blend 1.6/10; 9.4s, ZH.
ZH_B00066_S04994_W000075 · in -16.0 dBFS · gain -4.0 dB · emolia-03936
This chain comes from the one-sided rule: only Pain had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Pain below average — 0.35, lower than 65 % of clips in this corpus — and ends with it at the very top of the corpus at 0.93, higher than 93 % of clips in this corpus. That is a total rise of 0.58.
Nothing was asked of the other axis, and in fact Fear drifts down from 0.99 to 0.35 (-0.64), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.20, then +0.17, then +0.21 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? The least similar clip scores 0.66 against the first clip, where 1.00 would mean an identical voice. That is below the 0.80 threshold the mining used — treat the “same speaker” claim here with caution. Neighbouring clips score at worst 0.66 against each other.
Voice consistency: these clips are separate recordings joined together and the match is loose (0.66, under the 0.80 threshold), so the voice may audibly change between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 4 clips: a young adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, some disfluency, average clarity, light breath
(fear, teasing, amusement · normal-paced, energised, fully relaxed, casual)It's certainly been scary and intimidating. There's no doubt about that.
full caption & clip details
A young adult masculine voice; delivery is energised, normal-paced, fully relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly submissive, neutral openness; reads as fear, teasing, amusement; style: casual, playful; average recording, quiet background; explicit content; genuineness 5.4/6; vocal-burst blend 2.5/10; 3.1s, EN.
623066_00319440 · in -18.4 dBFS · gain -1.6 dB · podcast-00064
(hope enthusiasm optimism, interest, doubt· normal-paced, normally alert, slightly relaxed, casual)If people wanted to become live coders themselves, like where would they get started?
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as hope enthusiasm optimism, interest, doubt; style: casual, dramatic; good recording, quiet background; genuineness 1.5/6; vocal-burst blend 5.5/10; 30.0s, EN.
623066_00319784 · in -19.6 dBFS · gain -0.4 dB · podcast-06146
(concentration· normal-paced, normally alert, slightly relaxed, conversational)Yeah, that must be a challenging part.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as concentration; style: conversational, casual; good recording, no background noise; genuineness 3.0/6; vocal-burst blend 5.5/10; 19.7s, EN.
623066_00326547 · in -19.7 dBFS · gain -0.3 dB · podcast-06143
(pain·measured, normally alert, slightly relaxed, casual)(ahem) One of the streams I did that was really helpful for
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as pain; style: casual, conversational; good recording, no background noise; genuineness 2.3/6; vocal-burst blend 3.9/10; 3.6s, EN.
623066_00328644 · in -19.2 dBFS · gain -0.8 dB · podcast-05792
Embarrassment ↑ (unconstrained axis: Intoxication Altered States of Consciousness)emotion__B1__T0.50__C0.25__INTERNAL · #14
This chain comes from the one-sided rule: only Embarrassment had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Embarrassment below average — 0.42, lower than 58 % of clips in this corpus — and ends with it at the very top of the corpus at 0.97, higher than 97 % of clips in this corpus. That is a total rise of 0.55.
Nothing was asked of the other axis, and in fact Intoxication Altered States of Consciousness barely moves at all, sitting near 0.95 throughout.
It takes 4 clips to get there. Clip to clip the moves are +0.21, then +0.21, then +0.13 — a fairly even climb, though some clips carry more of the change than others.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 29 s · zh · emolia
k 4d_a -0.025d_b 0.547step_a 0.205step_b 0.210min_cos_consec —min_cos_anchor —dataset emolialang zhspeaker ZH_B00015_S02227track ZH_B00015_S02227total 29.2slevel spread 11.9 dBmax seam 11.9 dB
Script — 4 chunks, 2 with a non-speech sound
Unchanged across all 4 clips: a young adult masculine voice · quiet background, normally alert
(intoxication altered states of consciousness, contempt, teasing · fast, neutral tension, moderately variable, casual)有到底听还是不听,好像有听有不听的,挺大争议。
full caption & clip details
A young adult masculine voice; delivery is normally alert, fast, neutral tension, moderately variable; timbre is neutral-toned, dark, very rough, thin; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly dominant, neutral openness; reads as intoxication altered states of consciousness, contempt, teasing; style: casual, conversational; average recording, quiet background; explicit content; genuineness 4.5/6; vocal-burst blend 4.6/10; 3.5s, ZH.
ZH_B00015_S02227_W000099 · in -28.9 dBFS · gain +8.9 dB · emolia-03429
This chain comes from the one-sided rule: only Hope Enthusiasm Optimism had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Hope Enthusiasm Optimism barely there — 0.20, lower than 80 % of clips in this corpus — and ends with it strongly present at 0.83, higher than 83 % of clips in this corpus. That is a total rise of 0.62.
Nothing was asked of the other axis, and in fact Longing drifts down from 0.82 to 0.04 (-0.78), which the rule did not require.
It takes 4 clips to get there. Clip to clip the moves are +0.24, then +0.16, then +0.22 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 24 s · zh · emolia
k 4d_a -0.779d_b 0.624step_a 0.801step_b 0.244min_cos_consec —min_cos_anchor —dataset emolialang zhspeaker ZH_B00057_S05209track ZH_B00057_S05209total 23.6slevel spread 2.4 dBmax seam 1.7 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, average recording, no background noise, normally alert, slightly relaxed
(measured, no disfluency, authoritative, didactic)也就是只关注做更好的产品,还是固定在右侧象限,关注企业系统。
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: authoritative, didactic; average recording, no background noise; genuineness 1.8/6; vocal-burst blend 2.8/10; 6.1s, ZH.
ZH_B00057_S05209_W000047 · in -19.3 dBFS · gain -0.7 dB · emolia-03851
(normal-paced, some disfluency, authoritative, monologue)实际上这个世界上有很多的人能提供比那些跨国公司更好的产品和服务。
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: authoritative, monologue; average recording, no background noise; genuineness 1.8/6; vocal-burst blend 3.6/10; 5.7s, ZH.
ZH_B00057_S05209_W000048 · in -19.1 dBFS · gain -0.9 dB · emolia-03851
(thankfulness gratitude, pain·fast, no disfluency, formal, authoritative)就像有几十亿人,能做出比麦当劳汉堡包更好的汉堡业。
full caption & clip details
An adult masculine voice; delivery is normally alert, fast, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as thankfulness gratitude, pain; style: formal, authoritative; average recording, no background noise; genuineness 2.4/6; vocal-burst blend 3.8/10; 4.4s, ZH.
ZH_B00057_S05209_W000049 · in -18.6 dBFS · gain -1.4 dB · emolia-03851
(measured, no disfluency, monologue, authoritative)但是呢只有麦当劳拥有能够提供几十亿人吃的汉堡包的企业系统。
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: monologue, authoritative; average recording, no background noise; genuineness 1.2/6; vocal-burst blend 1.9/10; 6.9s, ZH.
ZH_B00057_S05209_W000050 · in -16.9 dBFS · gain -3.1 dB · emolia-03851
This chain comes from the one-sided rule: only Relief had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Relief below average — 0.32, lower than 68 % of clips in this corpus — and ends with it strongly present at 0.85, higher than 85 % of clips in this corpus. That is a total rise of 0.53.
Nothing was asked of the other axis, and in fact Triumph drifts down from 0.84 to 0.38 (-0.46), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.21, then +0.13, then +0.09, then +0.11 — a fairly even climb, though some clips carry more of the change than others.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 54 s · zh · emolia
k 5d_a -0.457d_b 0.530step_a 0.646step_b 0.209min_cos_consec —min_cos_anchor —dataset emolialang zhspeaker ZH_B00058_S02568track ZH_B00058_S02568total 54.1slevel spread 1.7 dBmax seam 1.7 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: a young adult masculine voice · neutral-toned, fairly smooth, balanced body, average recording, measured, normally alert, slightly relaxed, fairly steady
(no disfluency, clear, moderate pitch range, authoritative)免得他这个产品稍微一比价之后,就觉得我们价格高原先的产品也不合作了。
full caption & clip details
A young adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: authoritative, monologue; average recording, no background noise; genuineness 1.6/6; vocal-burst blend 4.1/10; 6.9s, ZH.
ZH_B00058_S02568_W000028 · in -19.4 dBFS · gain -0.6 dB · emolia-03860
This chain comes from the one-sided rule: only Pain had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Pain barely there — 0.20, lower than 80 % of clips in this corpus — and ends with it at the very top of the corpus at 0.91, higher than 91 % of clips in this corpus. That is a total rise of 0.72.
Nothing was asked of the other axis, and in fact Bitterness drifts down from 0.72 to 0.62 (-0.10), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.15, then +0.20, then +0.17, then +0.20 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? Not directly measured. What does exist is a timbre similarity of 0.86 against the first clip, which describes how alike the voices sound rather than whether they are the same person. It sits on a different scale from the identity check (corpus-wide the timbre numbers run much higher), so it cannot be read against the 0.80 identity threshold and is given here without a pass or fail.
Voice consistency: these clips are separate recordings joined together. Speaker identity was not measured for this sample; the available timbre similarity of 0.86 says the voices sound broadly alike but is not a same-person check. You may notice the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 44 s · en · emolia
k 5d_a -0.100d_b 0.719step_a 0.699step_b 0.198min_cos_consec —min_cos_anchor —dataset emolialang enspeaker EN_Mox4tKp9okktrack EN_Mox4tKp9okktotal 44.4slevel spread 4.5 dBmax seam 4.5 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an adult masculine voice · neutral-toned, fairly smooth, balanced body, good recording, no background noise, normally alert, slightly relaxed, clear
(measured, steady, little disfluency, monologue)Marx later refers to this method as real subsumption.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; no dominant emotion; style: monologue, formal; good recording, no background noise; genuineness 0.6/6; vocal-burst blend 1.7/10; 4.2s, EN.
EN_Mox4tKp9okk_W000009 · in -16.7 dBFS · gain -3.3 dB · emolia-01967
(concentration·normal-paced, steady, little disfluency, formal)Under this method, relative surplus value is produced by increasing labor's productivity and intensity, devaluing their labor power via the cheapening of commodities needed for their own reproduction, thus increasing the amount of time they are producing surplus value for the capitalists.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as concentration; style: formal, monologue; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.0/10; 18.8s, EN.
EN_Mox4tKp9okk_W000010 · in -21.1 dBFS · gain +1.1 dB · emolia-01967
(concentration, contemplation· normal-paced, fairly steady, little disfluency, monologue)The concept of a productive worker therefore implies not merely a relation between the activity of work and its useful effect between the worker and the product of his work.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as concentration, contemplation; style: monologue, formal; good recording, no background noise; genuineness 1.5/6; vocal-burst blend 1.0/10; 10.3s, EN.
EN_Mox4tKp9okk_W000011 · in -16.6 dBFS · gain -3.4 dB · emolia-01967
(emotional numbness·measured, steady, little disfluency, whispered)But also a specifically social relation of production. A relation with a historical origin.
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as emotional numbness; style: whispered, monologue; good recording, no background noise; genuineness 1.7/6; vocal-burst blend 0.7/10; 6.1s, EN.
EN_Mox4tKp9okk_W000012 · in -17.4 dBFS · gain -2.6 dB · emolia-01967
(pain·normal-paced, steady, almost no disfluency, monologue)Which stamps the worker as capital's direct means of valorisation.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as pain; style: monologue, formal; good recording, no background noise; genuineness 1.2/6; vocal-burst blend 1.3/10; 4.3s, EN.
EN_Mox4tKp9okk_W000013 · in -16.6 dBFS · gain -3.4 dB · emolia-01967
This chain comes from the one-sided rule: only Infatuation had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Infatuation below average — 0.27, lower than 73 % of clips in this corpus — and ends with it strongly present at 0.89, higher than 89 % of clips in this corpus. That is a total rise of 0.61.
Nothing was asked of the other axis, and in fact Relief drifts down from 0.82 to 0.55 (-0.27), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.17, then +0.18, then +0.20, then +0.06 — an uneven climb, but always in the same direction.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? Not directly measured. What does exist is a timbre similarity of 0.91 against the first clip, which describes how alike the voices sound rather than whether they are the same person. It sits on a different scale from the identity check (corpus-wide the timbre numbers run much higher), so it cannot be read against the 0.80 identity threshold and is given here without a pass or fail.
Voice consistency: these clips are separate recordings joined together. Speaker identity was not measured for this sample; the available timbre similarity of 0.91 says the voices sound broadly alike but is not a same-person check. You may notice the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 61 s · en · emolia
k 5d_a -0.274d_b 0.611step_a 0.866step_b 0.199min_cos_consec —min_cos_anchor —dataset emolialang enspeaker EN_tMSiF3yP9-4track EN_tMSiF3yP9-4total 61.1slevel spread 4.2 dBmax seam 3.6 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an adult masculine voice · neutral-toned, neutral-bright, balanced body, normally alert, slightly relaxed
(normal-paced, fairly steady, little disfluency, formal)The water amount can be regulated directly at the panel, supported by a water pump.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, monologue; good recording, no background noise; genuineness 1.4/6; vocal-burst blend 0.7/10; 5.0s, EN.
EN_tMSiF3yP9-4_W000007 · in -14.6 dBFS · gain -5.4 dB · emolia-00006
(concentration· normal-paced, steady, almost no disfluency, formal)For even more economical use of the machine, the ecoefficiency mode reduces detergent and energy consumption and prolongs the running time of the machine.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, almost no disfluency, fairly narrow pitch, minimal breath; affect is neutral, neutral stance, slightly guarded; reads as concentration; style: formal, didactic; average recording, no background noise; genuineness 0.6/6; vocal-burst blend 0.6/10; 11.2s, EN.
EN_tMSiF3yP9-4_W000008 · in -15.1 dBFS · gain -4.9 dB · emolia-00006
(concentration · normal-paced, fairly steady, some disfluency, didactic)The dose system for a precise detergent dosage or for switching detergent off. Saving money and protecting the environment are the main topics of the system. Thanks to a newly developed exhaust air channel, we reduce the noise level to a minimum.
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; somewhat unclear, some disfluency, fairly narrow pitch, minimal breath; affect is mildly positive, neutral stance, slightly guarded; reads as concentration; style: didactic, monologue; average recording, quiet background; genuineness 0.9/6; vocal-burst blend 1.5/10; 16.4s, EN.
EN_tMSiF3yP9-4_W000009 · in -18.7 dBFS · gain -1.3 dB · emolia-00006
(concentration ·measured, steady, almost no disfluency, formal)Which makes the machine ideal for hospitals, bigger hotels, et cetera. The easy to understand color coding simplifies the operation and maintenance of the machine. Yellow parts are for operation, gray parts are for maintenance.
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, almost no disfluency, fairly narrow pitch, minimal breath; affect is neutral, neutral stance, slightly guarded; reads as concentration; style: formal, monologue; average recording, quiet background; genuineness 0.0/6; vocal-burst blend 1.0/10; 15.9s, EN.
EN_tMSiF3yP9-4_W000010 · in -16.0 dBFS · gain -4.0 dB · emolia-00006
(measured, steady, almost no disfluency, formal)Very often, manual cleaning tools like mops or scrapers are needed to get into corners and edges where the machine cannot reach. The home base kit is the helpful solution.
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, fairly narrow pitch, minimal breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, didactic; average recording, quiet background; genuineness 0.2/6; vocal-burst blend 0.6/10; 12.0s, EN.
EN_tMSiF3yP9-4_W000011 · in -15.3 dBFS · gain -4.7 dB · emolia-00006
Intoxication Altered States of Consciousness ↑ (unconstrained axis: Impatience and Irritability)emotion__B1__T0.50__C0.25__INTERNAL · #19
This chain comes from the one-sided rule: only Intoxication Altered States of Consciousness had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Intoxication Altered States of Consciousness below average — 0.40, lower than 60 % of clips in this corpus — and ends with it at the very top of the corpus at 0.98, higher than 98 % of clips in this corpus. That is a total rise of 0.59.
Nothing was asked of the other axis, and in fact Impatience and Irritability drifts down from 0.97 to 0.37 (-0.60), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.15, then +0.12, then +0.19, then +0.13 — a fairly even climb, though some clips carry more of the change than others.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? Not directly measured. What does exist is a timbre similarity of 0.37 against the first clip, which describes how alike the voices sound rather than whether they are the same person. It sits on a different scale from the identity check (corpus-wide the timbre numbers run much higher), so it cannot be read against the 0.80 identity threshold and is given here without a pass or fail.
Voice consistency: these clips are separate recordings joined together. Speaker identity was not measured for this sample; the available timbre similarity of 0.37 says the voices sound broadly alike but is not a same-person check. You may notice the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 36 s · en · emolia
k 5d_a -0.597d_b 0.587step_a 0.945step_b 0.189min_cos_consec —min_cos_anchor —dataset emolialang enspeaker EN_gPGxA5FzJXutrack EN_gPGxA5FzJXutotal 36.5slevel spread 6.4 dBmax seam 4.4 dB
Script — 5 chunks, 3 with a non-speech sound
Unchanged across all 5 clips: a young adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, normal-paced, normally alert, average clarity, light breath
(impatience and irritability, astonishment surprise, sexual lust · slightly relaxed, fairly steady, some disfluency, casual)It's a mindfuck, Canada, actually. It's a massive mindfuck.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as impatience and irritability, astonishment surprise, sexual lust; style: casual, conversational; good recording, no background noise; genuineness 2.5/6; vocal-burst blend 3.0/10; 3.1s, EN.
EN_gPGxA5FzJXu_W000184 · in -19.5 dBFS · gain -0.5 dB · emolia-00304
(awe, astonishment surprise · slightly relaxed, fairly steady, some disfluency, casual)Yeah. We have the second biggest landmass on the planet for a nation state and we have 35 million people. That's amazing.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as awe, astonishment surprise; style: casual, monologue; average recording, quiet background; genuineness 3.5/6; vocal-burst blend 3.4/10; 6.9s, EN.
EN_gPGxA5FzJXu_W000185 · in -21.7 dBFS · gain +1.7 dB · emolia-00304
(astonishment surprise, pleasure ecstasy, elation·relaxed, moderately variable, some disfluency, casual)That's why (ahem) today, it's like, yeah, today I posted it on my story, just holding up my Canadian passport. I just got one today. Thanks, literally.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as astonishment surprise, pleasure ecstasy, elation; style: casual, conversational; average recording, some background noise; mildly explicit content; genuineness 5.9/6; vocal-burst blend 2.2/10; 7.5s, EN.
EN_gPGxA5FzJXu_W000186 · in -17.9 dBFS · gain -2.1 dB · emolia-00304
(infatuation, astonishment surprise, embarrassment· relaxed, moderately variable, frequent disfluency, casual)My, my buddy, uh, (low mumble) yeah, uh, (low mumble) he does, (ahem) uh, that's amazing. I was just, I was just saying hi.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, wide pitch range, light breath; affect is positive, slightly submissive, neutral openness; reads as infatuation, astonishment surprise, embarrassment; style: casual, conversational; average recording, quiet background; genuineness 5.4/6; vocal-burst blend 3.8/10; 4.7s, EN.
EN_gPGxA5FzJXu_W000187 · in -15.3 dBFS · gain -4.7 dB · emolia-00304
(intoxication altered states of consciousness, elation, hope enthusiasm optimism· relaxed, fairly steady, frequent disfluency, casual)I was just, I was just saying how I think to, (ahem) uh, I'm really grateful for Canada. Canada is great. I think long-term, Canada is like one of the best places to be at or, or have (low mumble) like that connection with. Yeah.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as intoxication altered states of consciousness, elation, hope enthusiasm optimism; style: casual, conversational; average recording, quiet background; genuineness 5.7/6; vocal-burst blend 6.3/10; 13.7s, EN.
EN_gPGxA5FzJXu_W000188 · in -19.7 dBFS · gain -0.3 dB · emolia-00304
This chain comes from the one-sided rule: only Triumph had to get where it was going, by at least 0.50. The other emotion was left completely free.
The chain starts with Triumph below average — 0.38, lower than 62 % of clips in this corpus — and ends with it at the very top of the corpus at 0.97, higher than 97 % of clips in this corpus. That is a total rise of 0.58.
Nothing was asked of the other axis, and in fact Fear drifts down from 0.98 to 0.92 (-0.06), which the rule did not require.
It takes 5 clips to get there. Clip to clip the moves are +0.22, then +0.24, then -0.02, then +0.14 — not a clean run: step 3 moves back the other way by 0.02 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? The least similar clip scores 0.63 against the first clip, where 1.00 would mean an identical voice. That is below the 0.80 threshold the mining used — treat the “same speaker” claim here with caution. Neighbouring clips score at worst 0.54 against each other.
Voice consistency: these clips are separate recordings joined together and the match is loose (0.63, under the 0.80 threshold), so the voice may audibly change between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 5 clips: an elderly feminine voice · average recording, quiet background, wide pitch range
(fear, bitterness, distress · normal-paced, normally alert, neutral tension, casual)And then Joe Biden came back and he said, Hey, (low mumble) um, let's talk about the loss of lives. Now, guys, the number is rising every single day. Every single day, there are more and more people dying from coronavirus because of the lack of action from your current president. Over 211,000 people
full caption & clip details
An elderly feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, slightly rough, thin; average clarity, frequent disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, neutral openness; reads as fear, bitterness, distress; style: casual, monologue; average recording, quiet background; genuineness 5.0/6; vocal-burst blend 6.8/10; 25.6s, EN.
938466_00105348 · in -20.4 dBFS · gain +0.4 dB · podcast-02933
(malevolence malice, affection, bitterness ·brisk, energised, neutral tension, casual)The only thing he knows is hate. So everything that it comes from him is hate. It comes from him talking about himself. Oh, well, I did this and I did that, taking credit for things that the Obama administration was able to do.
full caption & clip details
An elderly feminine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, slightly bright, slightly rough, thin; somewhat unclear, some disfluency, wide pitch range, normal breath; affect is negative, slightly dominant, slightly guarded; reads as malevolence malice, affection, bitterness; style: casual, monologue; average recording, quiet background; genuineness 4.6/6; vocal-burst blend 7.5/10; 17.2s, EN.
938466_00109356 · in -21.1 dBFS · gain +1.1 dB · podcast-02931
(contempt, malevolence malice, anger· brisk, normally alert, neutral tension, casual)When they talked about Obamacare, he just won. Oh, and I wanted to take out take out, (low mumble) um, take it out. Excuse me. I just wanted to just go ahead and get rid of it. And then, but what is your alternative? Trump care? Is that your real alternative? How is it going to be different?
full caption & clip details
A young adult feminine voice; delivery is normally alert, brisk, neutral tension, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; somewhat unclear, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as contempt, malevolence malice, anger; style: casual, conversational; average recording, quiet background; genuineness 5.1/6; vocal-burst blend 10.0/10; 19.8s, EN.
938466_00111068 · in -20.5 dBFS · gain +0.5 dB · podcast-02931
(contempt, sourness, disgust·slow, highly aroused, slightly relaxed, casual)Guys, think about it like this.
full caption & clip details
A child masculine voice; delivery is highly aroused, slow, slightly relaxed, fairly steady; timbre is slightly cool, neutral-bright, rough, thin; slurred, frequent disfluency, wide pitch range, no audible breath; affect is neutral, slightly dominant, slightly guarded; reads as contempt, sourness, disgust; style: casual; average recording, quiet background; genuineness 3.5/6; vocal-burst blend 3.8/10; 3.8s, EN.
938466_00113052 · in -24.1 dBFS · gain +4.1 dB · podcast-02960
(triumph, hope enthusiasm optimism, malevolence malice·brisk, energised, neutral tension, casual)God forbid you catch coronavirus, or even if you are listening to this and you have it, somebody, and then you were able to be cured of it, not necessarily cured, but you got over it and you're still living and you're trying to get back to the semblance of (low mumble) um surviving and thriving in life like you were before you came down with coronavirus. Guys, now you have a pre existing condition
full caption & clip details
A young adult feminine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as triumph, hope enthusiasm optimism, malevolence malice; style: casual, conversational; average recording, quiet background; genuineness 3.1/6; vocal-burst blend 9.2/10; 26.7s, EN.
938466_00113428 · in -20.0 dBFS · gain +0.0 dB · podcast-02946