Rule.VN1 — VoiceNet: one of 57 voice-descriptor dimensions sweeps by >=T, each consecutive step <=C Source. trajectories_v5.parquet | Family. one corpus in isolation Sampled from 72,000 matching rows, without replacement across the family, so no two tiers reuse a chain.
How to read a Script. Each chunk is one line:
a short tag of what the models heard in that clip, then the words spoken.
(underlined, plain ·
delivery, style) — the tag before the words. Emotions first, then how
it is delivered. Underlined descriptors are the ones that change across this chain
— anything identical on every clip is pulled out and stated once above, because a
value that never moves says nothing about a trajectory.
(ahem) — brackets inside the words
are a different thing: a real non-speech sound, printed where it happens. Most clips have
none; about a quarter do.
The full generated caption for any clip is still there, under
“full caption & clip details”. Its perceived-gender and background-noise
clauses were re-rendered from the numeric buckets, because the versions stored in the
corpus index had those two ladders running backwards.
The emotion clause has been re-derived, and it used to be wrong. The 40 emotion heads are not on a common scale — Interest has a median of 2.08 and is never zero, while Infatuation is zero on 87.7 % of clips — and the caption named an emotion whenever its raw score cleared an absolute 1.0. Interest therefore appeared in 94.8 % of captions and Sadness in almost none: the clause was reporting the scale of the head, not the emotion of the clip. An emotion is now named only when it lands in the top 10 % for that emotion, against a pooled tie-aware mid-rank ECDF over 132,833,726 utterances spanning every dataset and language. Interest now appears in 6.2 %, all 40 emotions occur, and a clip that is ordinary on all 40 says “no dominant emotion” rather than being forced to pick one (21.8 % of clips). This is the same scale the trajectory miner selects on, so the caption and the mining now refer to the same quantity: the mined target emotion is named in the final clip's caption on 73 % of chains, up from 46 %.
S_RANT — style: ranting ↓c-mls-VN1 · #1
This is a VoiceNet dimension, not an emotion: style: ranting (S_RANT) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with style: ranting (S_RANT) below average — 0.33, lower than 67 % of clips in this corpus — and works its way down to low at 0.11, lower than 89 % of clips in this corpus. That is a total fall of 0.21.
It takes 3 clips to get there. Clip to clip the moves are +0.00, then -0.21 — a slow start, with most of the change arriving in the final step.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
3 clips · 52 s · french · mls
hear it un-normalised (raw levels, max seam 0.3 dB)
k 3d_a -0.212d_b -0.212step_a 0.212step_b 0.212min_cos_consec —min_cos_anchor —dataset mlslang frenchspeaker 1840track 1840|derniermohicans_13_cototal 51.9slevel spread 0.3 dBmax seam 0.3 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: a middle-aged masculine voice · slightly dark, slightly rough, balanced body, average recording, quiet background, measured, slightly relaxed, steady
(bitterness, anger, sourness · subdued, some disfluency, somewhat unclear, monologue)vous savez ce que le major heyward vous a promis et moi-même magua secoua la tête et lui défendit de répéter des offres qu'il méprisait que voulez-vous donc demanda cora convaincue douloureusement que la franchise du trop généreux duncan s'était laissé abuser par la duplicité maligne d'un sauvage
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, some disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as bitterness, anger, sourness; style: monologue, narration; average recording, quiet background; genuineness 0.2/6; vocal-burst blend 2.6/10; 19.8s, FRENCH.
1840_10379_003743 · in -30.3 dBFS · gain +10.3 dB · mls-00044
(bitterness, anger, sourness · subdued, some disfluency, somewhat unclear, monologue)vous savez ce que le major heyward vous a promis et moi-même magua secoua la tête et lui défendit de répéter des offres qu'il méprisait que voulez-vous donc demanda cora convaincue douloureusement que la franchise du trop généreux duncan s'était laissé abuser par la duplicité maligne d'un sauvage
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, some disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as bitterness, anger, sourness; style: monologue, narration; average recording, quiet background; genuineness 0.2/6; vocal-burst blend 2.6/10; 19.8s, FRENCH.
1840_10379_003743 · in -30.3 dBFS · gain +10.3 dB · mls-00044
(malevolence malice, anger, disgust·very low-energy, frequent disfluency, slurred, monologue)ce que veut un huron est de rendre le bien pour le bien et le mal pour le mal vous voulez donc vous venger de l'insulte que vous a faite munro sur deux filles sans défense
full caption & clip details
A middle-aged masculine voice; delivery is very low-energy, measured, slightly relaxed, steady; timbre is warm, slightly dark, slightly rough, balanced body; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is neutral, neutral stance, fairly guarded; reads as malevolence malice, anger, disgust; style: monologue, narration; average recording, quiet background; genuineness 0.5/6; vocal-burst blend 0.6/10; 12.0s, FRENCH.
1840_10379_003705 · in -30.0 dBFS · gain +10.0 dB · mls-00044
ROUG — roughness ↑c-mls-VN1 · #2
This is a VoiceNet dimension, not an emotion: roughness (ROUG) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with roughness (ROUG) below average — 0.33, lower than 67 % of clips in this corpus — and ends with it above average at 0.71, higher than 71 % of clips in this corpus. That is a total rise of 0.38.
It takes 4 clips to get there. Clip to clip the moves are +0.11, then +0.09, then +0.19 — a fairly even climb, though some clips carry more of the change than others.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 53 s · portuguese · mls
hear it un-normalised (raw levels, max seam 0.5 dB)
k 4d_a 0.384d_b 0.384step_a 0.188step_b 0.188min_cos_consec —min_cos_anchor —dataset mlslang portuguesespeaker 5677track 5677|tristefimdepolicarpoqtotal 53.3slevel spread 0.5 dBmax seam 0.5 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: a young adult masculine voice · balanced body, average recording, quiet background, slightly relaxed, frequent disfluency, fairly narrow pitch
(measured, subdued, fairly steady, monologue)o seu pequeno aborrecimento é não poder de quando em quando soltar o peito o comandante do destacamento é quaresma que talvez consentisse
full caption & clip details
A young adult masculine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; somewhat unclear, frequent disfluency, fairly narrow pitch, normal breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: monologue, whispered; average recording, quiet background; genuineness 1.8/6; vocal-burst blend 0.6/10; 11.6s, PORTUGUESE.
5677_4807_001132 · in -26.5 dBFS · gain +6.5 dB · mls-00120
(emotional numbness, malevolence malice, fatigue exhaustion·slow, very low-energy, steady, whispered)o major está no interior da casa que serve de quartel lendo o seu estudo predileto é agora artilharia comprou compêndios
full caption & clip details
A middle-aged masculine voice; delivery is very low-energy, slow, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly negative, neutral stance, fairly guarded; reads as emotional numbness, malevolence malice, fatigue exhaustion; style: whispered, monologue; average recording, quiet background; genuineness 0.9/6; vocal-burst blend 0.5/10; 10.5s, PORTUGUESE.
5677_4807_001028 · in -26.5 dBFS · gain +6.5 dB · mls-00120
(concentration·measured, subdued, steady, monologue)mas como sua instrução é insuficiente da artilharia vai à balística da balística à mecânica da mecânica ao cálculo e à geometria analítica desce mais a escada vai à trigonometria à geometria e à álgebra e à aritmética
full caption & clip details
A child masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is slightly cool, dark, slightly rough, balanced body; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is neutral, neutral stance, slightly guarded; reads as concentration; style: monologue, whispered; average recording, quiet background; genuineness 2.4/6; vocal-burst blend 1.8/10; 18.3s, PORTUGUESE.
5677_4807_000791 · in -26.4 dBFS · gain +6.5 dB · mls-00120
(emotional numbness, concentration, malevolence malice· measured, subdued, steady, monologue)ele percorre essa cadeia de ciências entrelaçadas com uma fé de inventor aprende uma noção elementaríssima após um rosário de consultas de compêndio em compêndio
full caption & clip details
A young adult masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is neutral, neutral stance, slightly guarded; reads as emotional numbness, concentration, malevolence malice; style: monologue, whispered; average recording, quiet background; genuineness 2.0/6; vocal-burst blend 1.8/10; 12.4s, PORTUGUESE.
5677_4807_001239 · in -26.9 dBFS · gain +6.9 dB · mls-00120
TEMP — tempo ↓c-mls-VN1 · #3
This is a VoiceNet dimension, not an emotion: tempo (TEMP) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.50.
The chain starts with tempo (TEMP) above average — 0.71, higher than 71 % of clips in this corpus — and works its way down to low at 0.11, lower than 89 % of clips in this corpus. That is a total fall of 0.61.
It takes 5 clips to get there. Clip to clip the moves are -0.08, then -0.19, then -0.17, then -0.16 — an uneven climb, but always in the same direction.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 64 s · spanish · mls
hear it un-normalised (raw levels, max seam 1.8 dB)
k 5d_a -0.605d_b -0.605step_a 0.189step_b 0.189min_cos_consec —min_cos_anchor —dataset mlslang spanishspeaker 10246track 10246|canasybarro_05_blasctotal 63.7slevel spread 1.8 dBmax seam 1.8 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, normally alert, slightly relaxed, fairly steady
(sourness, disgust · normal-paced, almost no disfluency, narration, monologue)en el altar mostraba eu carita sonriente y bu falda hueca el n fio jeiúa patrón del pueblo una imagen que no levantaba más de un palmo pero á pesar de su pequenez sabía llenar de anguilas en las noches
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as sourness, disgust; style: narration, monologue; good recording, no background noise; genuineness 0.6/6; vocal-burst blend 2.5/10; 14.1s, SPANISH.
10246_11643_001321 · in -27.5 dBFS · gain +7.5 dB · mls-00036
(normal-paced, almost no disfluency, authoritative, monologue)las barcas de los que conseguían los mejores puestos con otros milagros no menos asombrosos que relataban las mujeres del palmar en las paredes se destacaban sobre el fondo blanco algunos cuadres procedentes de antiguos conventos
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: authoritative, monologue; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.9/10; 14.7s, SPANISH.
10246_11643_001554 · in -28.0 dBFS · gain +8.0 dB · mls-00036
(contempt, malevolence malice, anger· normal-paced, no disfluency, storytelling, monologue)tablas enormes con falanges de conde nados todos rojos como si acabasen de ser cocí dos y ángeles de plumaje de cotorras arreándolos con flamígeras espadas
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as contempt, malevolence malice, anger; style: storytelling, monologue; good recording, quiet background; genuineness 0.4/6; vocal-burst blend 0.2/10; 10.6s, SPANISH.
10246_11643_001419 · in -27.9 dBFS · gain +7.9 dB · mls-00036
(malevolence malice ·measured, no disfluency, authoritative, monologue)sobre la pila de agua bendita un cartelón con caracteres góticos rezaba así si por la ley del amor no es licito delinquir no se permite escupir en la casa del señor
full caption & clip details
A young adult feminine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as malevolence malice; style: authoritative, monologue; good recording, no background noise; genuineness 0.5/6; vocal-burst blend 1.9/10; 13.4s, SPANISH.
10246_11643_000903 · in -28.7 dBFS · gain +8.7 dB · mls-00036
(astonishment surprise, fear, relief·normal-paced, no disfluency, monologue, authoritative)no había en el palmar quien no admirase estos versos obra según el tío paloma de cierto vicario allá en los tiempos en que el barquero era mozo
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as astonishment surprise, fear, relief; style: monologue, authoritative; good recording, no background noise; genuineness 0.7/6; vocal-burst blend 0.9/10; 10.3s, SPANISH.
10246_11643_001017 · in -26.9 dBFS · gain +6.9 dB · mls-00036
R_NASL — resonance: nasal ↓c-mls-VN1 · #4
This is a VoiceNet dimension, not an emotion: resonance: nasal (R_NASL) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with resonance: nasal (R_NASL) high — 0.87, higher than 87 % of clips in this corpus — and works its way down to low at 0.17, lower than 83 % of clips in this corpus. That is a total fall of 0.70.
It takes 5 clips to get there. Clip to clip the moves are -0.17, then -0.25, then -0.20, then -0.08 — an uneven climb, but always in the same direction.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 80 s · french · mls
hear it un-normalised (raw levels, max seam 2.9 dB)
k 5d_a -0.697d_b -0.697step_a 0.250step_b 0.250min_cos_consec —min_cos_anchor —dataset mlslang frenchspeaker 13177track 13177|contesjouretnuit_03_total 80.2slevel spread 3.0 dBmax seam 2.9 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an elderly masculine voice · quiet background, steady
(emotional numbness, contemplation, longing · slow, very low-energy, slightly relaxed, whispered)chaque jour il se levait à la même heure suivait les mêmes rues passait par la même porte devant le même concierge entrait dans le même bureau s'asseyait sur le même siège et accomplissait la même besogne
full caption & clip details
An elderly masculine voice; delivery is very low-energy, slow, slightly relaxed, steady; timbre is warm, slightly dark, slightly rough, balanced body; slurred, frequent disfluency, narrow pitch range, audible breath; affect is neutral, neutral stance, slightly guarded; reads as emotional numbness, contemplation, longing; style: whispered, monologue; average recording, quiet background; genuineness 1.2/6; vocal-burst blend 0.3/10; 13.4s, FRENCH.
13177_14238_000082 · in -26.1 dBFS · gain +6.1 dB · mls-00058
(pain, helplessness, sadness· slow, very low-energy, relaxed, whispered)ii était seul au monde seul le jour au milieu de ses collègues indifférents seul la nuit dans on logement de garçon ii économisait cent francs par mois pour la vieillesse
full caption & clip details
An elderly masculine voice; delivery is very low-energy, slow, relaxed, steady; timbre is warm, dark, rough, thin; slurred, frequent disfluency, narrow pitch range, audible breath; affect is neutral, neutral stance, slightly guarded; reads as pain, helplessness, sadness; style: whispered, monologue; poor recording, quiet background; genuineness 0.9/6; vocal-burst blend 0.9/10; 13.3s, FRENCH.
13177_14238_000048 · in -24.4 dBFS · gain +4.5 dB · mls-00058
(jealousy and envy, malevolence malice, emotional numbness·measured, subdued, slightly relaxed, monologue)chaque dimanche il faisait un tour aux champs-elysées afin de regarder passer le monde élégant les équipages et les jolies femmes ii disait le lendemain à son compagnon de peine le retour du bois était fort
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, frequent disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as jealousy and envy, malevolence malice, emotional numbness; style: monologue, narration; average recording, quiet background; genuineness 0.6/6; vocal-burst blend 1.2/10; 15.6s, FRENCH.
13177_14238_000000 · in -23.8 dBFS · gain +3.8 dB · mls-00058
(sourness, malevolence malice, bitterness· measured, very low-energy, slightly relaxed, narration)alors il revint cherchant à la voir encore elle s'était assise maintenant le garçon demeurait très sage à son côté tandis que la fillette faisait des pâtés de terre c'était elle c'était bien elle elle avait un air sérieux de dame une toilette simple une allure assurée et digne
full caption & clip details
A middle-aged masculine voice; delivery is very low-energy, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; clear, some disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as sourness, malevolence malice, bitterness; style: narration, whispered; average recording, quiet background; genuineness 0.2/6; vocal-burst blend 1.6/10; 18.8s, FRENCH.
13177_14238_000024 · in -26.8 dBFS · gain +6.8 dB · mls-00058
(fear, disappointment, contentment· measured, subdued, slightly relaxed, whispered)ii la regardait de loin n'osant pas approcher le petit garçon leva la tête françois tessier se sentit trembler c'était son fils sans doute et il le considéra et il crut se re connaître lui-même tel qu'il était sur une photocrraphie faite autrefois
full caption & clip details
A middle-aged somewhat feminine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, frequent disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as fear, disappointment, contentment; style: whispered, monologue; average recording, quiet background; genuineness 0.2/6; vocal-burst blend 0.4/10; 18.5s, FRENCH.
13177_14238_000050 · in -26.9 dBFS · gain +6.9 dB · mls-00058
S_WHIS — style: whispered ↓c-mls-VN1 · #5
This is a VoiceNet dimension, not an emotion: style: whispered (S_WHIS) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with style: whispered (S_WHIS) high — 0.87, higher than 87 % of clips in this corpus — and works its way down to around average at 0.55, higher than 55 % of clips in this corpus. That is a total fall of 0.32.
It takes 5 clips to get there. Clip to clip the moves are -0.13, then +0.00, then -0.19, then +0.00 — a plateau around step 2, where it barely moves.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 73 s · dutch · mls
hear it un-normalised (raw levels, max seam 1.4 dB)
k 5d_a -0.320d_b -0.320step_a 0.187step_b 0.187min_cos_consec —min_cos_anchor —dataset mlslang dutchspeaker 2450track 2450|elisabethmusch_15_lentotal 73.5slevel spread 1.4 dBmax seam 1.4 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an elderly masculine voice · neutral-toned, slightly dark, average recording, quiet background, slightly relaxed
(infatuation, malevolence malice · slow, very low-energy, steady, monologue)wel dan hoe eer hoe beter maar zeide gourville morgen ochtend of morgen avond zoo gy wilt jacob van lennep elizabeth musch het zij gelijk gy t verkiest
full caption & clip details
An elderly masculine voice; delivery is very low-energy, slow, slightly relaxed, steady; timbre is neutral-toned, slightly dark, rough, thin; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is neutral, neutral stance, fairly guarded; reads as infatuation, malevolence malice; style: monologue, cartoonish; average recording, quiet background; genuineness 1.4/6; vocal-burst blend 0.3/10; 12.8s, DUTCH.
2450_10026_003807 · in -29.9 dBFS · gain +9.9 dB · mls-00085
(contemplation, sourness·measured, normally alert, fairly steady, cartoonish)en ik neem deze gelegenheid te gretiger aan om dat ik u toch wenschte te raadplegen over een zaak van eer waarin ik niet twijfel dat uw ondervinding my van dienst zal kunnen zijn
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; average clarity, some disfluency, fairly narrow pitch, audible breath; affect is neutral, neutral stance, slightly guarded; reads as contemplation, sourness; style: cartoonish, monologue; average recording, quiet background; genuineness 2.2/6; vocal-burst blend 0.6/10; 11.9s, DUTCH.
2450_10026_003127 · in -28.5 dBFS · gain +8.5 dB · mls-00084
(contemplation, sourness · measured, normally alert, fairly steady, cartoonish)en ik neem deze gelegenheid te gretiger aan om dat ik u toch wenschte te raadplegen over een zaak van eer waarin ik niet twijfel dat uw ondervinding my van dienst zal kunnen zijn
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; average clarity, some disfluency, fairly narrow pitch, audible breath; affect is neutral, neutral stance, slightly guarded; reads as contemplation, sourness; style: cartoonish, monologue; average recording, quiet background; genuineness 2.2/6; vocal-burst blend 0.6/10; 11.9s, DUTCH.
2450_10026_003127 · in -28.5 dBFS · gain +8.5 dB · mls-00084
(malevolence malice, disgust, anger· measured, normally alert, fairly steady, cartoonish)hier werd hun onderhoud gestoord door bromley die met drift kwam toegeloopen waar schuilt gy toch vroeg hy aan buat men zoekt u sints een geruimen tijd overal als een speld mademoiselle van beverweert heeft naar u gevraagd naar my vroeg buat
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; average clarity, some disfluency, wide pitch range, normal breath; affect is neutral, slightly dominant, slightly guarded; reads as malevolence malice, disgust, anger; style: cartoonish, storytelling; average recording, quiet background; genuineness 1.2/6; vocal-burst blend 1.0/10; 18.1s, DUTCH.
2450_10026_001201 · in -29.7 dBFS · gain +9.7 dB · mls-00084
(malevolence malice, disgust, anger · measured, normally alert, fairly steady, cartoonish)hier werd hun onderhoud gestoord door bromley die met drift kwam toegeloopen waar schuilt gy toch vroeg hy aan buat men zoekt u sints een geruimen tijd overal als een speld mademoiselle van beverweert heeft naar u gevraagd naar my vroeg buat
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; average clarity, some disfluency, wide pitch range, normal breath; affect is neutral, slightly dominant, slightly guarded; reads as malevolence malice, disgust, anger; style: cartoonish, storytelling; average recording, quiet background; genuineness 1.2/6; vocal-burst blend 1.0/10; 18.1s, DUTCH.
2450_10026_001201 · in -29.7 dBFS · gain +9.7 dB · mls-00084
VALS — valence stability ↓c-mls-VN1 · #6
This is a VoiceNet dimension, not an emotion: valence stability (VALS) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.50.
The chain starts with valence stability (VALS) high — 0.87, higher than 87 % of clips in this corpus — and works its way down to below average at 0.37, lower than 63 % of clips in this corpus. That is a total fall of 0.51.
It takes 4 clips to get there. Clip to clip the moves are -0.05, then -0.24, then -0.21 — a plateau around step 1, where it barely moves.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 71 s · dutch · mls
k 4d_a -0.506d_b -0.506step_a 0.240step_b 0.240min_cos_consec —min_cos_anchor —dataset mlslang dutchspeaker 2450track 2450|mauritslijnslager_1_0total 70.9slevel spread 1.4 dBmax seam 1.4 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: a middle-aged masculine voice · neutral-toned, slightly dark, slightly rough, balanced body, average recording, measured, subdued, slightly relaxed
(thankfulness gratitude, relief, sexual lust · somewhat unclear, whispered, monologue)maar dat daar en boven rijkelijk van allerlei soorten van wijn voorzien was hij gaf een knecht die bij hetzelve stond een wenk en deze schonk terstond een glas vol met den besten franschen wijn en hiermede heette van vliet zijnen jeugdigen gast op nieuw welkom
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, little disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as thankfulness gratitude, relief, sexual lust; style: whispered, monologue; average recording, no background noise; genuineness 1.1/6; vocal-burst blend 0.0/10; 17.5s, DUTCH.
2450_10565_000824 · in -28.7 dBFS · gain +8.7 dB · mls-00087
(sexual lust, malevolence malice, concentration· somewhat unclear, narration, monologue)liep ondertusschen het zangstuk af en lijnslager vervoegde zich bij het gezelschap eerst tot de huisvrouw van van vliet die door haar man onderrigt wie hij was hem op eene niet minder gastvrije wijze dan haar echtgenoot begroette
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, little disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as sexual lust, malevolence malice, concentration; style: narration, monologue; average recording, quiet background; genuineness 0.6/6; vocal-burst blend 0.0/10; 15.8s, DUTCH.
2450_10565_001103 · in -28.5 dBFS · gain +8.5 dB · mls-00088
(anger, concentration ·clear, narration, monologue)zoo deden ook de twee zonen die beide de instrumenten waarop zjj gespeeld hadden nederzetten maria de dochter ontving de pligtpleging van lijnslaadriaan loosjes pzn het leven van maurits lijnslager ger met ongemaakte vriendelijkheid terwijl de zedigheid der zuivere onschuld uit hare oogen straalde
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; clear, little disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as anger, concentration; style: narration, monologue; average recording, no background noise; genuineness 0.6/6; vocal-burst blend 0.1/10; 17.3s, DUTCH.
2450_10565_001915 · in -27.6 dBFS · gain +7.6 dB · mls-00088
(concentration, contemplation· clear, narration, monologue)en zijne zuster beantwoordden zijne pligtplegingen met de gewone beleefdheid die men aan een onbekenden bewijst lijnslager verzocht dringende dat zijne komst geene verstoring mogt geven in het gezelschap en dat men voort zou varen met zingen en spelen
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; clear, little disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as concentration, contemplation; style: narration, monologue; average recording, no background noise; genuineness 0.4/6; vocal-burst blend 0.0/10; 19.8s, DUTCH.
2450_10565_000704 · in -29.1 dBFS · gain +9.1 dB · mls-00087
VOLT — loudness / volume ↑c-mls-VN1 · #7
This is a VoiceNet dimension, not an emotion: loudness / volume (VOLT) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.50.
The chain starts with loudness / volume (VOLT) low — 0.16, lower than 84 % of clips in this corpus — and ends with it above average at 0.69, higher than 69 % of clips in this corpus. That is a total rise of 0.53.
It takes 5 clips to get there. Clip to clip the moves are +0.05, then +0.11, then +0.13, then +0.23 — a slow start, with most of the change arriving in the final step.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 81 s · portuguese · mls
k 5d_a 0.529d_b 0.529step_a 0.232step_b 0.232min_cos_consec —min_cos_anchor —dataset mlslang portuguesespeaker 2961track 2961|oateneu_04_pompeia_64total 80.9slevel spread 3.8 dBmax seam 3.8 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an elderly somewhat feminine voice · slightly cool, quiet background, frequent disfluency, audible breath
(malevolence malice, anger, bitterness · slow, very low-energy, slightly relaxed, whispered)no recreio andava só e calado como um monge depois do sanches não me aproximava de nenhum colega senão incidentemente por palavras indispensáveis rebelo tentou atrair me eu desviava
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, slightly relaxed, fairly steady; timbre is slightly cool, slightly dark, slightly rough, thin; clear, frequent disfluency, narrow pitch range, audible breath; affect is mildly positive, neutral stance, slightly guarded; reads as malevolence malice, anger, bitterness; style: whispered, ASMR; average recording, quiet background; genuineness 1.8/6; vocal-burst blend 0.2/10; 16.9s, PORTUGUESE.
2961_3702_001730 · in -33.4 dBFS · gain +13.4 dB · mls-00118
(malevolence malice, awe, bitterness · slow, very low-energy, relaxed, monologue)sanches rancoroso perseguia me como um demônio dizia coisas imundas deixa estar jurava entre dentes que ainda hei de tirar te a vergonha
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, relaxed, fairly steady; timbre is slightly cool, dark, rough, thin; slurred, frequent disfluency, narrow pitch range, audible breath; affect is mildly negative, submissive, neutral openness; reads as malevolence malice, awe, bitterness; style: monologue, whispered; below-average recording, quiet background; genuineness 3.0/6; vocal-burst blend 3.6/10; 12.4s, PORTUGUESE.
2961_3702_001621 · in -34.8 dBFS · gain +14.8 dB · mls-00118
(sourness, malevolence malice, disgust·measured, very low-energy, slightly relaxed, whispered)na qualidade de vigilante levava me brutalmente à espada eu tinha as pernas roxas dos golpes as canelas me incharam se barbalho se lembra de vingar a bofetada creio que me submetia à letra evangélica
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, measured, slightly relaxed, fairly steady; timbre is slightly cool, dark, rough, thin; slurred, frequent disfluency, narrow pitch range, audible breath; affect is mildly negative, neutral stance, slightly guarded; reads as sourness, malevolence malice, disgust; style: whispered, monologue; poor recording, quiet background; genuineness 3.2/6; vocal-burst blend 2.7/10; 16.9s, PORTUGUESE.
2961_3702_001544 · in -31.1 dBFS · gain +11.1 dB · mls-00118
(fear, shame, confusion· measured, subdued, slightly relaxed, whispered)durante este período de depressão contemplativa uma coisa apenas magoava me não tinha o ar angélico do ribas não cantava tão bem como ele que faria se morresse entre os anjos sem saber cantar
full caption & clip details
A middle-aged somewhat feminine voice; delivery is subdued, measured, slightly relaxed, moderately variable; timbre is slightly cool, neutral-bright, slightly rough, balanced body; average clarity, frequent disfluency, moderate pitch range, audible breath; affect is mildly positive, neutral stance, slightly guarded; reads as fear, shame, confusion; style: whispered, narration; average recording, quiet background; genuineness 2.2/6; vocal-burst blend 1.6/10; 16.4s, PORTUGUESE.
2961_3702_001716 · in -32.8 dBFS · gain +12.8 dB · mls-00118
(sourness, malevolence malice, disgust·slow, very low-energy, slightly relaxed, whispered)ribas quinze anos era feio magro linfático boca sem lábios de velha carpideira desenhada em angústia a súplica feita boca a prece perene rasgada em beiços sobre dentes
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, slightly relaxed, moderately variable; timbre is slightly cool, dark, rough, thin; slurred, frequent disfluency, narrow pitch range, audible breath; affect is negative, neutral stance, slightly guarded; reads as sourness, malevolence malice, disgust; style: whispered, cartoonish; poor recording, quiet background; genuineness 1.9/6; vocal-burst blend 0.2/10; 17.8s, PORTUGUESE.
2961_3702_001574 · in -32.7 dBFS · gain +12.7 dB · mls-00118
R_MASK — resonance: mask ↓c-mls-VN1 · #8
This is a VoiceNet dimension, not an emotion: resonance: mask (R_MASK) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with resonance: mask (R_MASK) around average — 0.49, lower than 51 % of clips in this corpus — and works its way down to below average at 0.28, lower than 72 % of clips in this corpus. That is a total fall of 0.20.
It takes 2 clips to get there. Clip to clip the moves are -0.20 — a single step, so there is no internal shape to speak of.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 36 s · german · mls
k 2d_a -0.204d_b -0.204step_a 0.204step_b 0.204min_cos_consec —min_cos_anchor —dataset mlslang germanspeaker 11927track 11927|volkssagen_08_temme_total 36.0slevel spread 0.6 dBmax seam 0.6 dB
Script — 2 chunks, 0 with a non-speech sound
Unchanged across all 2 clips: a middle-aged masculine voice · neutral-toned, neutral-bright, balanced body, average recording, quiet background, normally alert, slightly relaxed, fairly steady
(triumph · measured, some disfluency, moderate pitch range, monologue)der linke arm des götzen war in die seite gesetzt und bildete auf diese weise einen bogen der gott trug ein gewand das bis auf die schienbeine herabreichte mit den füßen stand er auf einem gestell das aber so tief in die erde hineingelassen oder hineingesunken war
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as triumph; style: monologue, didactic; average recording, quiet background; genuineness 1.1/6; vocal-burst blend 0.0/10; 19.3s, GERMAN.
11927_12408_000200 · in -28.9 dBFS · gain +8.9 dB · mls-00010
(emotional numbness, contentment·slow, frequent disfluency, fairly narrow pitch, didactic)daß man es nicht mehr sehen konnte nahe bei dem bilde hingen sattel zaum und schwert des gottes das schwert war von ungemeiner größe gefäß und scheide desselben waren von silber mit feiner eingelegter arbeit
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, slow, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; average clarity, frequent disfluency, fairly narrow pitch, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as emotional numbness, contentment; style: didactic, monologue; average recording, quiet background; genuineness 1.0/6; vocal-burst blend 0.0/10; 16.5s, GERMAN.
11927_12408_000102 · in -28.3 dBFS · gain +8.3 dB · mls-00010
DARC — darkness of timbre ↑c-mls-VN1 · #9
This is a VoiceNet dimension, not an emotion: darkness of timbre (DARC) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with darkness of timbre (DARC) around average — 0.51, right about the corpus median — and ends with it high at 0.75, higher than 75 % of clips in this corpus. That is a total rise of 0.24.
It takes 2 clips to get there. Clip to clip the moves are +0.24 — a single step, so there is no internal shape to speak of.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 31 s · italian · mls
k 2d_a 0.243d_b 0.243step_a 0.243step_b 0.243min_cos_consec —min_cos_anchor —dataset mlslang italianspeaker 2033track 2033|cuore_01_deamicis_64ktotal 31.4slevel spread 1.4 dBmax seam 1.4 dB
Script — 2 chunks, 0 with a non-speech sound
Unchanged across all 2 clips: a child masculine voice · slightly cool, neutral-bright, fairly smooth, average recording, quiet background, measured, slightly relaxed, frequent disfluency
(shame, pride, contemplation · subdued, steady, slurred, monologue)pensa ai ragazzi muti e ciechi che pure studiano e fino ai prigionieri che anch'essi imparano a leggere e a scrivere pensa la mattina quando esci che in quello stesso momento nella tua stessa città
full caption & clip details
A child masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is slightly cool, neutral-bright, fairly smooth, balanced body; slurred, frequent disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as shame, pride, contemplation; style: monologue, narration; average recording, quiet background; genuineness 1.3/6; vocal-burst blend 2.5/10; 15.6s, ITALIAN.
2033_1596_000752 · in -25.9 dBFS · gain +5.9 dB · mls-00068
(bitterness, anger, pain·normally alert, moderately variable, average clarity, monologue)altri trentamila ragazzi vanno come te a chiudersi per tre ore in una stanza a studiare ma che pensa agli innumerevoli ragazzi che presso a poco a quell'ora vanno a scuola in tutti i paesi
full caption & clip details
A child masculine voice; delivery is normally alert, measured, slightly relaxed, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, thin; average clarity, frequent disfluency, moderate pitch range, audible breath; affect is neutral, neutral stance, slightly guarded; reads as bitterness, anger, pain; style: monologue, cartoonish; average recording, quiet background; genuineness 3.0/6; vocal-burst blend 3.0/10; 15.6s, ITALIAN.
2033_1596_000432 · in -24.5 dBFS · gain +4.5 dB · mls-00068
S_WHIS — style: whispered ↑c-mls-VN1 · #10
This is a VoiceNet dimension, not an emotion: style: whispered (S_WHIS) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with style: whispered (S_WHIS) above average — 0.71, higher than 71 % of clips in this corpus — and ends with it at the very top of the range at 0.96, higher than 96 % of clips in this corpus. That is a total rise of 0.26.
It takes 4 clips to get there. Clip to clip the moves are +0.21, then -0.14, then +0.19 — not a clean run: step 2 moves back the other way by 0.14 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 58 s · french · mls
k 4d_a 0.256d_b 0.256step_a 0.209step_b 0.209min_cos_consec —min_cos_anchor —dataset mlslang frenchspeaker 12899track 12899|hommeoreillecassee_1total 58.5slevel spread 1.4 dBmax seam 0.8 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: an elderly feminine voice · slightly cool, neutral-bright, thin, average recording, quiet background, slightly relaxed, moderately variable, clear
(malevolence malice, sourness, contempt · measured, normally alert, some disfluency, cartoonish)les ronflements de ce gros homme avaient quelque chose de sinistre vous eussiez cru entendre les ophicléides du jugement dernier quelle ombre le visita dans cette heure de sommeil
full caption & clip details
An elderly feminine voice; delivery is normally alert, measured, slightly relaxed, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, thin; clear, some disfluency, moderate pitch range, audible breath; affect is neutral, slightly dominant, slightly guarded; reads as malevolence malice, sourness, contempt; style: cartoonish, dramatic; average recording, quiet background; genuineness 2.6/6; vocal-burst blend 2.5/10; 12.8s, FRENCH.
12899_13862_000973 · in -29.9 dBFS · gain +9.9 dB · mls-00056
(relief, malevolence malice, helplessness· measured, normally alert, frequent disfluency, cartoonish)nul étranger ne l'a jamais su car il gardait ses rêves pour lui comme tout ce qui lui appartenait mais entre deux stations le train étant lancé à toute vitesse il sentit distinctement deux mains énergiques qui le tiraient par les pieds
full caption & clip details
A child feminine voice; delivery is normally alert, measured, slightly relaxed, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, thin; clear, frequent disfluency, narrow pitch range, audible breath; affect is neutral, neutral stance, slightly guarded; reads as relief, malevolence malice, helplessness; style: cartoonish, whispered; average recording, quiet background; genuineness 1.5/6; vocal-burst blend 1.4/10; 17.7s, FRENCH.
12899_13862_001077 · in -29.7 dBFS · gain +9.7 dB · mls-00056
(longing, contentment, contemplation·slow, energised, frequent disfluency, cartoonish)sensation trop connue hélas et qui lui rappelait les plus mauvais souvenirs de sa vie il ouvrit les yeux avec épouvante et vit l'homme de la photographie dans le costume de la photographie
full caption & clip details
An elderly feminine voice; delivery is energised, slow, slightly relaxed, moderately variable; timbre is slightly cool, neutral-bright, rough, thin; clear, frequent disfluency, very wide pitch range, audible breath; affect is positive, slightly dominant, slightly guarded; reads as longing, contentment, contemplation; style: cartoonish, storytelling; average recording, quiet background; genuineness 2.3/6; vocal-burst blend 2.9/10; 13.9s, FRENCH.
12899_13862_000925 · in -28.9 dBFS · gain +8.9 dB · mls-00056
(disgust, sourness, jealousy and envy·measured, normally alert, frequent disfluency, whispered)ses cheveux se hérissèrent ses yeux s'arrondirent en boules de loto il poussa un grand cri et se jeta à corps perdu entre les deux banquettes dans les jambes de ses voisins
full caption & clip details
A child feminine voice; delivery is normally alert, measured, slightly relaxed, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, thin; clear, frequent disfluency, fairly narrow pitch, audible breath; affect is neutral, neutral stance, slightly guarded; reads as disgust, sourness, jealousy and envy; style: whispered, didactic; average recording, quiet background; genuineness 2.1/6; vocal-burst blend 1.5/10; 13.6s, FRENCH.
12899_13862_000067 · in -28.5 dBFS · gain +8.5 dB · mls-00056
GEND — perceived gender ↑c-mls-VN1 · #11
This is a VoiceNet dimension, not an emotion: perceived gender (GEND) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with perceived gender (GEND) around average — 0.57, higher than 57 % of clips in this corpus — and ends with it high at 0.88, higher than 88 % of clips in this corpus. That is a total rise of 0.32.
It takes 5 clips to get there. Clip to clip the moves are +0.02, then +0.14, then -0.07, then +0.22 — not a clean run: step 3 moves back the other way by 0.07 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 83 s · german · mls
k 5d_a 0.317d_b 0.317step_a 0.221step_b 0.221min_cos_consec —min_cos_anchor —dataset mlslang germanspeaker 5753track 5753|fraubovary_04_flaubertotal 83.5slevel spread 1.2 dBmax seam 0.5 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, average recording, slightly relaxed, light breath
(infatuation, anger, longing · normal-paced, normally alert, fairly steady, monologue)karl konnte seiner patienten wegen nicht länger verweilen vater rouault ließ das ehepaar in seinem wagen nach haus fahren und gab ihm persönlich bis vassonville das geleite beim abschied küßte er seine tochter noch einmal dann stieg er aus und machte sich zu fuß auf den rückweg
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as infatuation, anger, longing; style: monologue, narration; average recording, quiet background; genuineness 0.7/6; vocal-burst blend 0.4/10; 18.5s, GERMAN.
5753_10524_003279 · in -28.6 dBFS · gain +8.7 dB · mls-00006
(contemplation, relief, contentment·measured, subdued, steady, monologue)nachdem er hundert schritte gegangen war blieb er stehen um dem wagen nachzuschauen der die sandige straße dahinrollte dabei seufzte er tief auf er dachte zurück an seine eigne hochzeit an längstvergangne tage an die zeit der ersten mutterschaft seiner frau
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as contemplation, relief, contentment; style: monologue, narration; average recording, quiet background; genuineness 0.3/6; vocal-burst blend 0.8/10; 19.0s, GERMAN.
5753_10524_003309 · in -29.2 dBFS · gain +9.2 dB · mls-00006
(contemplation, relief, longing· measured, normally alert, steady, monologue)wie froh war er damals gewesen er erinnerte sich des tages wo er mit ihr das haus des schwiegervaters verlassen hatte auf dem ritt in das eigne heim durch den tiefen schnee da hatte er seine frau hinten auf die kruppe seines pferdes gesetzt
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as contemplation, relief, longing; style: monologue, newsreading; average recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.8/10; 18.1s, GERMAN.
5753_10524_003103 · in -29.6 dBFS · gain +9.6 dB · mls-00006
(fear, sadness, awe· measured, normally alert, steady, narration)es war so um weihnachten herum gewesen und die ganze gegend war verschneit mit der einen hand hatte sie sich an ihm festgehalten in der andern ihren korb getragen
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as fear, sadness, awe; style: narration, monologue; average recording, no background noise; genuineness 0.3/6; vocal-burst blend 0.6/10; 11.9s, GERMAN.
5753_10524_003286 · in -29.9 dBFS · gain +9.9 dB · mls-00006
(sourness, jealousy and envy, sadness ·normal-paced, normally alert, fairly steady, narration)die langen bänder ihres normannischen kopfputzes hatten im winde geflattert und manchmal waren sie ihm um die nase geflogen und wenn er sich umdrehte sah er über seine schulter weg ganz dicht hinter sich ihr niedliches rosiges gesicht
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as sourness, jealousy and envy, sadness; style: narration, monologue; average recording, no background noise; genuineness 0.9/6; vocal-burst blend 1.5/10; 15.5s, GERMAN.
5753_10524_003008 · in -29.5 dBFS · gain +9.5 dB · mls-00006
S_STRY — style: storytelling ↑c-mls-VN1 · #12
This is a VoiceNet dimension, not an emotion: style: storytelling (S_STRY) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with style: storytelling (S_STRY) above average — 0.75, higher than 75 % of clips in this corpus — and ends with it at the very top of the range at 0.99, higher than 99 % of clips in this corpus. That is a total rise of 0.24.
It takes 2 clips to get there. Clip to clip the moves are +0.24 — a single step, so there is no internal shape to speak of.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 21 s · italian · mls
k 2d_a 0.245d_b 0.245step_a 0.245step_b 0.245min_cos_consec —min_cos_anchor —dataset mlslang italianspeaker 4998track 4998|novelle10_12_pirandeltotal 21.2slevel spread 0.7 dBmax seam 0.7 dB
Script — 2 chunks, 0 with a non-speech sound
Unchanged across all 2 clips: a child masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, no background noise, measured, almost no disfluency
(affection, pain, infatuation · subdued, slightly relaxed, fairly steady, narration)tra titta e mauro poco dopo s'accese il diverbio mauro brandí una lepre e minacciò il fratello vennero alle mani
full caption & clip details
A child masculine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as affection, pain, infatuation; style: narration, monologue; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.2/10; 10.6s, ITALIAN.
4998_7720_000281 · in -30.0 dBFS · gain +10.0 dB · mls-00079
(pleasure ecstasy, pain, intoxication altered states of consciousness·energised, neutral tension, moderately variable, storytelling)mazurka mazurka esclamò in quella angelica udendo per fortuna i mandolini e la chitarra d'una serenata giú per la via
full caption & clip details
A child masculine voice; delivery is energised, measured, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, almost no disfluency, wide pitch range, normal breath; affect is positive, slightly dominant, slightly guarded; reads as pleasure ecstasy, pain, intoxication altered states of consciousness; style: storytelling, cartoonish; average recording, no background noise; genuineness 1.1/6; vocal-burst blend 0.8/10; 10.4s, ITALIAN.
4998_7720_000085 · in -29.2 dBFS · gain +9.2 dB · mls-00079
R_CHST — resonance: chest ↓c-mls-VN1 · #13
This is a VoiceNet dimension, not an emotion: resonance: chest (R_CHST) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.50.
The chain starts with resonance: chest (R_CHST) high — 0.83, higher than 83 % of clips in this corpus — and works its way down to below average at 0.29, lower than 71 % of clips in this corpus. That is a total fall of 0.54.
It takes 4 clips to get there. Clip to clip the moves are -0.17, then -0.12, then -0.25 — a fairly even climb, though some clips carry more of the change than others.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 80 s · spanish · mls
k 4d_a -0.540d_b -0.540step_a 0.246step_b 0.246min_cos_consec —min_cos_anchor —dataset mlslang spanishspeaker 8882track 8882|milnoches3_07_anonymototal 80.5slevel spread 1.4 dBmax seam 1.1 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: a middle-aged masculine voice · slightly rough, balanced body, average recording, quiet background, measured, slightly relaxed, frequent disfluency, slurred
(sourness, infatuation, awe · normally alert, moderately variable, wide pitch range, storytelling)es tal cosa para que inmediatamente le contestara tuya es si otro decía oh mi querido señor qué hermosa es esta anca inmediatamente le replicaba ali nur voy á mandar que la inscriban ahora
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; slurred, frequent disfluency, wide pitch range, audible breath; affect is positive, neutral stance, slightly guarded; reads as sourness, infatuation, awe; style: storytelling, cartoonish; average recording, quiet background; genuineness 2.2/6; vocal-burst blend 3.2/10; 20.0s, SPANISH.
8882_11802_000229 · in -24.6 dBFS · gain +4.6 dB · mls-00037
(contentment, affection, infatuation · normally alert, steady, fairly narrow pitch, monologue)tu nombre y mandaba traer el cálamo el tintero de cobre y el papel é inscribía la casa á nombre del amigo sellando el documento con su propio sello y así hizo durante todo un año y por la mañana daba un
full caption & clip details
An elderly masculine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is neutral, neutral stance, slightly guarded; reads as contentment, affection, infatuation; style: monologue, didactic; average recording, quiet background; genuineness 1.7/6; vocal-burst blend 2.7/10; 20.0s, SPANISH.
8882_11802_000069 · in -25.8 dBFS · gain +5.8 dB · mls-00037
(disgust, bitterness, contentment ·subdued, steady, fairly narrow pitch, whispered)á todos sus amigos y por la tarde les ofrecía otro al son de los instrumentos amenizándolo los mejores cantantes y las danzarinas más notables biblioteca valenciana generalitat valenciana las mil noches y una noche y ya no hacía caso de las advertencias de dulceamiga y hasta llegó
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; slurred, frequent disfluency, fairly narrow pitch, minimal breath; affect is neutral, neutral stance, slightly guarded; reads as disgust, bitterness, contentment; style: whispered, monologue; average recording, quiet background; genuineness 0.6/6; vocal-burst blend 2.3/10; 20.0s, SPANISH.
8882_11802_000426 · in -25.8 dBFS · gain +5.8 dB · mls-00037
(longing, affection, awe· subdued, fairly steady, fairly narrow pitch, whispered)tenerla olvidada pero ella no se quejaba nunca y se consolaba con la lectura de los libros de los poetas y un día que alí nur entró en su gabinete le dijo oh luz de mis ojos escucha estas estrofas cuanto más
full caption & clip details
An elderly masculine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is warm, slightly dark, slightly rough, balanced body; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly positive, neutral stance, slightly guarded; reads as longing, affection, awe; style: whispered, monologue; average recording, quiet background; genuineness 2.0/6; vocal-burst blend 4.4/10; 20.0s, SPANISH.
8882_11802_000628 · in -26.0 dBFS · gain +6.0 dB · mls-00037
SMTH — smoothness ↑c-mls-VN1 · #14
This is a VoiceNet dimension, not an emotion: smoothness (SMTH) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.50.
The chain starts with smoothness (SMTH) below average — 0.25, lower than 75 % of clips in this corpus — and ends with it high at 0.87, higher than 87 % of clips in this corpus. That is a total rise of 0.61.
It takes 4 clips to get there. Clip to clip the moves are +0.18, then +0.23, then +0.20 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 61 s · french · mls
k 4d_a 0.612d_b 0.612step_a 0.230step_b 0.230min_cos_consec —min_cos_anchor —dataset mlslang frenchspeaker 12709track 12709|elevegilles_06_lafontotal 60.6slevel spread 1.4 dBmax seam 1.1 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, normally alert, slightly relaxed, no disfluency
(relief · measured, steady, fairly narrow pitch, monologue)bereng la chemise pas sée il sautait au lit et tourné de mon côté pour éviter le rayonnement dp la veilleuse il attendait le sommeil qui ne tardait pas à le venir prendre
full caption & clip details
A young adult feminine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as relief; style: monologue, narration; good recording, no background noise; genuineness 0.3/6; vocal-burst blend 0.0/10; 11.6s, FRENCH.
12709_13548_000104 · in -28.9 dBFS · gain +8.9 dB · mls-00052
(confusion, fatigue exhaustion, longing· measured, steady, fairly narrow pitch, monologue)pendant ce temps je me déshabillais moi-même sans rapidité ne sachant pas dépouiller d'un seul coup caleçons bas et culotte et aussi parce qu'avant de retirer ma chemise je m'assurais que mon sca pulaire fût bien dissimulé
full caption & clip details
A young adult feminine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as confusion, fatigue exhaustion, longing; style: monologue, narration; good recording, quiet background; genuineness 0.1/6; vocal-burst blend 0.7/10; 14.4s, FRENCH.
12709_13548_000386 · in -28.6 dBFS · gain +8.6 dB · mls-00053
(longing, relief· measured, steady, fairly narrow pitch, narration)le réveil nous était signifié dès six heures par la cloche de la cour longuement sonnée un peu avant m laurin laissait l'alcôve dans laquelle nous l'entendions quelquefois s'habiller et faisait les cent pas dans le dortoir où l'habitude
full caption & clip details
A child feminine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; reads as longing, relief; style: narration, monologue; good recording, no background noise; genuineness 0.3/6; vocal-burst blend 0.9/10; 15.2s, FRENCH.
12709_13548_000376 · in -28.6 dBFS · gain +8.6 dB · mls-00053
(longing ·normal-paced, fairly steady, moderate pitch range, monologue)le jour naissant et le dernier passage du veilleur nous réveillaient à peu près tous il frappait trois coups au premier appel de la cloche et aussitôt chacun bondissait de son lit pour passer sa culotte et courir prendre place au lavabo près duquel le nombre insuffisant des robinets forçait les moins prompts à
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as longing; style: monologue, newsreading; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.5/10; 18.9s, FRENCH.
12709_13548_000409 · in -27.6 dBFS · gain +7.6 dB · mls-00053
R_NASL — resonance: nasal ↑c-mls-VN1 · #15
This is a VoiceNet dimension, not an emotion: resonance: nasal (R_NASL) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with resonance: nasal (R_NASL) low — 0.12, lower than 88 % of clips in this corpus — and ends with it around average at 0.46, lower than 54 % of clips in this corpus. That is a total rise of 0.34.
It takes 3 clips to get there. Clip to clip the moves are +0.19, then +0.15 — an even, steady climb — each clip carries about the same share.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
3 clips · 36 s · dutch · mls
k 3d_a 0.339d_b 0.339step_a 0.184step_b 0.184min_cos_consec —min_cos_anchor —dataset mlslang dutchspeaker 2450track 2450|devanderlindens_08_datotal 35.8slevel spread 1.1 dBmax seam 1.0 dB
Script — 3 chunks, 0 with a non-speech sound
Unchanged across all 3 clips: a middle-aged masculine voice · slightly dark, quiet background, steady, slurred, fairly narrow pitch, normal breath
(malevolence malice, disgust, fatigue exhaustion · measured, subdued, slightly relaxed, narration)wilt u den dokter spreken de vrouw keek haar scherp aan nam haar op van t hoofd tot de voeten en glimlachte op zonderlinge manier
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; slurred, some disfluency, fairly narrow pitch, normal breath; affect is neutral, neutral stance, slightly guarded; reads as malevolence malice, disgust, fatigue exhaustion; style: narration, storytelling; average recording, quiet background; genuineness 0.9/6; vocal-burst blend 0.0/10; 11.3s, DUTCH.
2450_11772_000960 · in -28.8 dBFS · gain +8.8 dB · mls-00091
(fatigue exhaustion, anger, relief·slow, very low-energy, relaxed, whispered)wilt u den dokter spreken herhaalde louise ongeduldig neen ik weet dat hij niet thuis is dus is het om mij te doen
full caption & clip details
A middle-aged masculine voice; delivery is very low-energy, slow, relaxed, steady; timbre is slightly warm, slightly dark, rough, slightly thin; slurred, frequent disfluency, fairly narrow pitch, normal breath; affect is neutral, neutral stance, fairly guarded; reads as fatigue exhaustion, anger, relief; style: whispered, narration; average recording, quiet background; genuineness 0.7/6; vocal-burst blend 0.0/10; 12.3s, DUTCH.
2450_11772_001045 · in -28.9 dBFS · gain +8.9 dB · mls-00091
(disgust, helplessness, fatigue exhaustion · slow, very low-energy, relaxed, whispered)ja om u hoe is uw naam als ik t vragen mag mijn naam ik heet donker ga zitten mevrouw donker
full caption & clip details
An elderly masculine voice; delivery is very low-energy, slow, relaxed, steady; timbre is neutral-toned, slightly dark, rough, slightly thin; slurred, frequent disfluency, fairly narrow pitch, normal breath; affect is negative, neutral stance, fairly guarded; reads as disgust, helplessness, fatigue exhaustion; style: whispered, monologue; below-average recording, quiet background; genuineness 1.2/6; vocal-burst blend 0.0/10; 11.9s, DUTCH.
2450_11772_000441 · in -29.9 dBFS · gain +9.9 dB · mls-00090
R_MASK — resonance: mask ↑c-mls-VN1 · #16
This is a VoiceNet dimension, not an emotion: resonance: mask (R_MASK) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.50.
The chain starts with resonance: mask (R_MASK) at the very bottom of the range — 0.03, lower than 97 % of clips in this corpus — and ends with it above average at 0.63, higher than 63 % of clips in this corpus. That is a total rise of 0.60.
It takes 5 clips to get there. Clip to clip the moves are +0.07, then +0.25, then +0.10, then +0.18 — an uneven climb, but always in the same direction.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 78 s · dutch · mls
k 5d_a 0.602d_b 0.602step_a 0.247step_b 0.247min_cos_consec —min_cos_anchor —dataset mlslang dutchspeaker 1724track 1724|roda_20_heimans_64kb.total 77.9slevel spread 3.0 dBmax seam 3.0 dB
Script — 5 chunks, 1 with a non-speech sound
Unchanged across all 5 clips: an elderly somewhat feminine voice
(disappointment, infatuation, jealousy and envy · slow, energised, neutral tension, storytelling)roda wat woon je hoog en wat een steile trap zeide omens hijgend van het klimmen mijn oude beenen zijn er niet meer voor geschikt en ik begrijp niet hoe jij minstens viermaal daags op en af kunt klimmen neen als je niet gelijkvloers gaat wonen kom ik niet meer bij je
full caption & clip details
An elderly somewhat feminine voice; delivery is energised, slow, neutral tension, moderately variable; timbre is neutral-toned, slightly dark, rough, slightly thin; somewhat unclear, some disfluency, wide pitch range, audible breath; affect is negative, neutral stance, vulnerable; reads as disappointment, infatuation, jealousy and envy; style: storytelling, monologue; average recording, quiet background; genuineness 2.3/6; vocal-burst blend 4.3/10; 17.3s, DUTCH.
1724_2649_002343 · in -23.2 dBFS · gain +3.2 dB · mls-00104
(anger, awe, sourness·measured, very low-energy, neutral tension, storytelling)wat voert je hier heen omens vroeg roda die met verbazing zijn vriend aanhoorde je weet toch dat ik straks naar dat kantoor moet wel juist daarom kom ik hier ik zal je brengen en je installeeren neen spreek niet tegen je bent klaar
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, measured, neutral tension, moderately variable; timbre is neutral-toned, slightly dark, slightly rough, slightly thin; clear, some disfluency, wide pitch range, audible breath; affect is mildly positive, neutral stance, neutral openness; reads as anger, awe, sourness; style: storytelling, narration; average recording, quiet background; genuineness 2.1/6; vocal-burst blend 4.7/10; 15.1s, DUTCH.
1724_2649_002222 · in -23.3 dBFS · gain +3.3 dB · mls-00103
(contempt, sourness, infatuation·fast, normally alert, neutral tension, casual)naar beneden dan ik breng je stel je aan de klerken voor en daarmee uit de bankier wil het zoo dat is een jon (low mumble) och een man bedoel ik die van die aardigheden houdt en wij oudjes moeten er ons maar aan onderwerpen voorwaarts goeden dag emilia goeden dag mevrouw
full caption & clip details
A child somewhat masculine voice; delivery is normally alert, fast, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, slightly thin; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as contempt, sourness, infatuation; style: casual, monologue; average recording, quiet background; genuineness 3.8/6; vocal-burst blend 6.3/10; 15.5s, DUTCH.
1724_2649_002135 · in -22.8 dBFS · gain +2.8 dB · mls-00103
(astonishment surprise, fatigue exhaustion, relief·measured, very low-energy, slightly relaxed, whispered)tot straks voegde hij er fluisterend bij en sloeg de deur snel achter zich dicht roda was reeds de trap af moeder riep emilia half schreiend half juichend en zich zelve steeds meer opwindend uit
full caption & clip details
A middle-aged somewhat feminine voice; delivery is very low-energy, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, slightly thin; clear, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly negative, neutral stance, slightly guarded; reads as astonishment surprise, fatigue exhaustion, relief; style: whispered, narration; good recording, quiet background; genuineness 1.6/6; vocal-burst blend 0.3/10; 13.7s, DUTCH.
1724_2649_002028 · in -25.8 dBFS · gain +5.8 dB · mls-00103
(confusion, fear, jealousy and envy·slow, energised, neutral tension, dramatic)moeder er is iets op til ik weet het ik voel het die opgezegde betrekking de nieuwe in ons huis die engelschman omens die vader komt halen de brief die uitblijft moeder willem komt hij is al hier met herman
full caption & clip details
An elderly feminine voice; delivery is energised, slow, neutral tension, variable; timbre is slightly cool, neutral-bright, rough, thin; clear, some disfluency, very wide pitch range, audible breath; affect is negative, submissive, vulnerable; reads as confusion, fear, jealousy and envy; style: dramatic, cartoonish; average recording, no background noise; genuineness 1.2/6; vocal-burst blend 3.1/10; 15.7s, DUTCH.
1724_2649_001850 · in -24.1 dBFS · gain +4.1 dB · mls-00103
EXPL — expressiveness ↑c-mls-VN1 · #17
This is a VoiceNet dimension, not an emotion: expressiveness (EXPL) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with expressiveness (EXPL) low — 0.15, lower than 85 % of clips in this corpus — and ends with it around average at 0.43, lower than 56 % of clips in this corpus. That is a total rise of 0.29.
It takes 5 clips to get there. Clip to clip the moves are +0.02, then -0.04, then +0.25, then +0.06 — not a clean run: step 2 moves back the other way by 0.04 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 73 s · dutch · mls
k 5d_a 0.287d_b 0.287step_a 0.249step_b 0.249min_cos_consec —min_cos_anchor —dataset mlslang dutchspeaker 1666track 1666|burgerhart_wolf-dekentotal 73.2slevel spread 1.6 dBmax seam 1.6 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, good recording, no background noise, slightly relaxed, fairly steady, clear
(disgust, jealousy and envy · measured, subdued, little disfluency, monologue)ik geloof niet dat zy iemand kan beminnen buiten haren schoothond zy is slordig en echter prachtig zy loopt veel uit ook in de kerk kent vele grote lieden wordt er dikwyls by verzogt leest geleerde autheuren in vier of vyf talen
full caption & clip details
A young adult feminine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, slightly thin; clear, little disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as disgust, jealousy and envy; style: monologue, narration; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 0.6/10; 15.3s, DUTCH.
1666_1548_000069 · in -31.0 dBFS · gain +11.0 dB · mls-00095
(relief, concentration·normal-paced, very low-energy, little disfluency, monologue)zy is ook ervaren in vele mag ik het zo eens noemen onvrouwlyke kundigheden en krygt veel brieven van geleerde mannen zy wil by ons zich eene meerderheid geven die wy haar gaarn laten
full caption & clip details
A young adult feminine voice; delivery is very low-energy, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as relief, concentration; style: monologue, narration; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 0.8/10; 11.6s, DUTCH.
1666_1548_000178 · in -31.4 dBFS · gain +11.4 dB · mls-00095
(normal-paced, normally alert, no disfluency, narration)met één woord juffr hartog is eene scavante van de alleronaangebetje wolff en aagje deken historie van mejuffrouw sara burgerhart naamste soort die alle dichters rymers en alle nederduitsche poezy vodden noemt
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, wide pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: narration, dramatic; good recording, no background noise; genuineness 0.8/6; vocal-burst blend 0.7/10; 10.1s, DUTCH.
1666_1548_000298 · in -29.8 dBFS · gain +9.8 dB · mls-00095
(jealousy and envy, anger· normal-paced, normally alert, almost no disfluency, narration)nu van juffrouw charlotte rien du tout zy is taamlyk mooi eene brunet groot en gezet een weinig ouder dan letje zy heeft geen karakter geen uur is zy dezelfde nu zit zy tot twaalf uuren ongekleed dan is zy voor dag en daauw in volle order en zit ons met het ontbyt te wagten want wy ontbyten met elkander
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as jealousy and envy, anger; style: narration, whispered; good recording, no background noise; genuineness 0.4/6; vocal-burst blend 2.8/10; 19.9s, DUTCH.
1666_1548_000796 · in -30.2 dBFS · gain +10.2 dB · mls-00095
(measured, subdued, little disfluency, monologue)behalven onze wysneus die het meeste geld verteert en daar voor veel grillen wil ingevolgt hebben lotje slaapt nu eens tot elf uuren dan zit zy den nagt over ofschoon zy niets doet of te doen heeft nu drinkt zy een glas water dan weer engelsch bier
full caption & clip details
A young adult feminine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, fairly narrow pitch, light breath; affect is mildly negative, neutral stance, slightly guarded; no dominant emotion; style: monologue, narration; good recording, no background noise; genuineness 0.7/6; vocal-burst blend 2.8/10; 15.6s, DUTCH.
1666_1548_000674 · in -31.3 dBFS · gain +11.3 dB · mls-00095
SMTH — smoothness ↑c-mls-VN1 · #18
This is a VoiceNet dimension, not an emotion: smoothness (SMTH) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with smoothness (SMTH) below average — 0.35, lower than 65 % of clips in this corpus — and ends with it above average at 0.58, higher than 58 % of clips in this corpus. That is a total rise of 0.23.
It takes 2 clips to get there. Clip to clip the moves are +0.23 — a single step, so there is no internal shape to speak of.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 33 s · dutch · mls
k 2d_a 0.233d_b 0.233step_a 0.233step_b 0.233min_cos_consec —min_cos_anchor —dataset mlslang dutchspeaker 1724track 1724|klaasjezevenster_2_17total 33.5slevel spread 0.1 dBmax seam 0.1 dB
Script — 2 chunks, 0 with a non-speech sound
Unchanged across all 2 clips: an elderly somewhat feminine voice · slightly rough, balanced body, measured, neutral tension, moderately variable, some disfluency
(relief, infatuation, contemplation · normally alert, clear, moderate pitch range, whispered)zoo mompelde de te leur gestelde vrijster nu ik moet zeggen t zijn wel dagen van geheimenissen en acob van ennep laasje evenster delen wonderen e vriend ylar heeft je zeker verteld wat er gisteren avond gebeurd is na ons vertrek een antwoordde ol is er iets van gewicht voorgevallen
full caption & clip details
An elderly somewhat feminine voice; delivery is normally alert, measured, neutral tension, moderately variable; timbre is slightly cool, slightly dark, slightly rough, balanced body; clear, some disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as relief, infatuation, contemplation; style: whispered, narration; average recording, quiet background; genuineness 2.6/6; vocal-burst blend 2.2/10; 18.1s, DUTCH.
1724_10268_000863 · in -23.8 dBFS · gain +3.8 dB · mls-00086
(astonishment surprise, contemplation, sourness·very low-energy, average clarity, wide pitch range, narration)u ik moet zeggen jijlui mans zijt toch rare wezens s hij zoo lang bij je geweest en heeft hij je niet eens van dat voorval met ettemie gesproken een antwoordde ol met eenige levendigheid is er met ettemie iets gebeurd
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, measured, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; average clarity, some disfluency, wide pitch range, audible breath; affect is mildly negative, neutral stance, neutral openness; reads as astonishment surprise, contemplation, sourness; style: narration, whispered; good recording, no background noise; genuineness 2.2/6; vocal-burst blend 1.4/10; 15.2s, DUTCH.
1724_10268_001179 · in -23.6 dBFS · gain +3.6 dB · mls-00086
BRGT — brightness of timbre ↑c-mls-VN1 · #19
This is a VoiceNet dimension, not an emotion: brightness of timbre (BRGT) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with brightness of timbre (BRGT) low — 0.11, lower than 89 % of clips in this corpus — and ends with it around average at 0.47, lower than 53 % of clips in this corpus. That is a total rise of 0.36.
It takes 4 clips to get there. Clip to clip the moves are +0.12, then +0.08, then +0.16 — a fairly even climb, though some clips carry more of the change than others.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
4 clips · 61 s · dutch · mls
k 4d_a 0.364d_b 0.364step_a 0.162step_b 0.162min_cos_consec —min_cos_anchor —dataset mlslang dutchspeaker 1724track 1724|wonderdokter_30_bosbototal 61.4slevel spread 0.6 dBmax seam 0.6 dB
Script — 4 chunks, 0 with a non-speech sound
Unchanged across all 4 clips: a child feminine voice · neutral-toned, average recording, slightly relaxed
(awe, anger, interest · normal-paced, normally alert, fairly steady, monologue)de groote vilthoed werd al vast afgenomen en wat brusk ter zijde geworpen het scheen schout gerrit wat warm en wat drukkend geworden in zijn ruim hoog vertrek om zich zelven echter over dit alles heen te zetten en bovenal om er niets van te laten doorschemeren voor het scherpe oog van zijne tegenpartij
full caption & clip details
A child feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, thin; slurred, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as awe, anger, interest; style: monologue, narration; average recording, quiet background; genuineness 1.0/6; vocal-burst blend 2.1/10; 17.1s, DUTCH.
1724_2757_003017 · in -24.0 dBFS · gain +4.0 dB · mls-00104
(disgust, bitterness, awe ·slow, very low-energy, variable, narration)dat verblijdt me dat gij uw jongenstijd nog herdenkt viel jacob jansz in met eenige levendigheid wel zeker en daarom zeg ik fij van die steile en stugge houding die gij tegen mij aanneemt wat zegt dat tusschen ons dat ik schout van delft ben geworden
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, slightly relaxed, variable; timbre is neutral-toned, dark, slightly rough, thin; clear, almost no disfluency, fairly narrow pitch, audible breath; affect is negative, neutral stance, vulnerable; reads as disgust, bitterness, awe; style: narration, storytelling; average recording, no background noise; genuineness 1.3/6; vocal-burst blend 2.6/10; 15.8s, DUTCH.
1724_2757_003091 · in -24.6 dBFS · gain +4.6 dB · mls-00104
(contemplation, concentration, interest·measured, very low-energy, fairly steady, whispered)gij maakt u vroolijk over eene onderstelling die een valschen grond heeft gerrit fransz viel graswinckel in met zekeren nadruk schoon ik mij niet inbeelde wereldsche wijsheid genoeg te bezitten om een goed magistraat te konnen zijn
full caption & clip details
A middle-aged somewhat feminine voice; delivery is very low-energy, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; clear, little disfluency, fairly narrow pitch, light breath; affect is mildly positive, neutral stance, neutral openness; reads as contemplation, concentration, interest; style: whispered, narration; average recording, no background noise; genuineness 1.3/6; vocal-burst blend 0.0/10; 13.9s, DUTCH.
1724_2757_003050 · in -24.4 dBFS · gain +4.5 dB · mls-00104
(contemplation, longing, disgust·normal-paced, normally alert, fairly steady, narration)weet ik toch dit van mij zelven dat ik eenmaal in zulk een ambt gesteld zijnde er de groote verantwoordelijkheid wel van zou kunnen beseffen en geen boosdoener zou sparen of beschermen ten koste van rustige burgers en eerlijke luiden
full caption & clip details
A young adult somewhat feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, thin; clear, almost no disfluency, fairly narrow pitch, light breath; affect is negative, neutral stance, slightly guarded; reads as contemplation, longing, disgust; style: narration, whispered; average recording, no background noise; genuineness 1.0/6; vocal-burst blend 1.5/10; 14.1s, DUTCH.
1724_2757_003397 · in -24.6 dBFS · gain +4.6 dB · mls-00104
RCQL — recording quality ↑c-mls-VN1 · #20
This is a VoiceNet dimension, not an emotion: recording quality (RCQL) describes the voice or the recording itself — how it sounds — rather than what the speaker feels. The rule asked it to sweep by at least 0.20.
The chain starts with recording quality (RCQL) low — 0.23, lower than 77 % of clips in this corpus — and ends with it around average at 0.56, higher than 56 % of clips in this corpus. That is a total rise of 0.33.
It takes 5 clips to get there. Clip to clip the moves are +0.12, then -0.15, then +0.15, then +0.22 — not a clean run: step 2 moves back the other way by 0.15 before the chain recovers.
No single step is larger than the 0.25 cap, which is exactly what stops this being a jump cut: the change has to be spread across the clips instead of landing all at once.
Same speaker? No similarity score is available here — the mls clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
5 clips · 83 s · portuguese · mls
k 5d_a 0.333d_b 0.333step_a 0.218step_b 0.218min_cos_consec —min_cos_anchor —dataset mlslang portuguesespeaker 2961track 2961|oalienista_1_machadodtotal 82.8slevel spread 3.2 dBmax seam 2.0 dB
Script — 5 chunks, 0 with a non-speech sound
Unchanged across all 5 clips: an elderly somewhat feminine voice · slightly cool, neutral-bright, average recording, quiet background, normally alert, moderately variable, some disfluency, wide pitch range
(sourness, jealousy and envy, contempt · measured, neutral tension, average clarity, cartoonish)foi então que um dos recantos desta lhe chamou especialmente a atenção o recanto psíquico o exame de patologia cerebral não havia na colônia e ainda no reino uma só autoridade em semelhante matéria mal explorada ou quase inexplorada
full caption & clip details
An elderly somewhat feminine voice; delivery is normally alert, measured, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, slightly rough, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as sourness, jealousy and envy, contempt; style: cartoonish, storytelling; average recording, quiet background; genuineness 3.3/6; vocal-burst blend 5.1/10; 18.9s, PORTUGUESE.
2961_3058_000241 · in -30.3 dBFS · gain +10.3 dB · mls-00116
(contempt, bitterness, malevolence malice·normal-paced, neutral tension, average clarity, cartoonish)simão bacamarte compreendeu que a ciência lusitana e particularmente a brasileira podia cobrir-se de louros imarcescíveis expressão usada por ele mesmo mas em um arroubo de intimidade doméstica exteriormente era modesto segundo convém aos sabedores
full caption & clip details
A middle-aged feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as contempt, bitterness, malevolence malice; style: cartoonish, storytelling; average recording, quiet background; genuineness 2.5/6; vocal-burst blend 6.1/10; 17.6s, PORTUGUESE.
2961_3058_000447 · in -30.5 dBFS · gain +10.5 dB · mls-00116
(malevolence malice, disgust, contempt ·measured, neutral tension, average clarity, dramatic)a saúde da alma bradou ele é a ocupação mais digna do médico do verdadeiro médico emendou crispim soares boticário da vila e um dos seus amigos e
full caption & clip details
An elderly somewhat feminine voice; delivery is normally alert, measured, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, slightly rough, thin; average clarity, some disfluency, wide pitch range, normal breath; affect is mildly positive, slightly dominant, slightly guarded; reads as malevolence malice, disgust, contempt; style: dramatic, monologue; average recording, quiet background; genuineness 3.2/6; vocal-burst blend 5.5/10; 13.0s, PORTUGUESE.
2961_3058_000386 · in -28.5 dBFS · gain +8.5 dB · mls-00116
(malevolence malice, contempt, sourness·normal-paced, slightly relaxed, clear, cartoonish)a vereança de itaguaí entre outros pecados de que é argüida pelos cronistas tinha o de não fazer caso dos dementes assim é que cada louco furioso era trancado em uma alcova na própria casa e não curado mas descurado até que a morte o vinha defraudar do benefício da vida
full caption & clip details
A middle-aged feminine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, balanced body; clear, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as malevolence malice, contempt, sourness; style: cartoonish, dramatic; average recording, quiet background; genuineness 2.2/6; vocal-burst blend 3.8/10; 19.2s, PORTUGUESE.
2961_3058_000274 · in -27.3 dBFS · gain +7.3 dB · mls-00116
(sourness, contempt, impatience and irritability·measured, neutral tension, average clarity, cartoonish)os mansos andavam à solta pela rua simão bacamarte entendeu desde logo reformar tão ruim costume pediu licença à câmara para agasalhar e tratar no edifício que ia construir
full caption & clip details
A child feminine voice; delivery is normally alert, measured, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, thin; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as sourness, contempt, impatience and irritability; style: cartoonish, storytelling; average recording, quiet background; genuineness 3.4/6; vocal-burst blend 5.7/10; 13.4s, PORTUGUESE.
2961_3058_000178 · in -27.8 dBFS · gain +7.8 dB · mls-00116