Sadness under rescue rule S3, k=2. the literal 'not sad -> clearly sad' pair. 90.0 % of clips score at or below zero on this emotion and the largest gap on its normalised axis is 0.449 (WIDER than the 0.25 step cap). This rule found 7,826 chains over 40,000 tracks; the strict rule found 0 at k=3.
These are not strict-rule trajectories. They come from a deliberately looser rule, built to recover examples on an emotion the strict rule cannot reach. What rule S3 changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'. What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained. Full explanation →
7,826chains this rule found
0the strict rule found (k=3)
0.449largest gap on the axis
90.0 %clips scoring ≤ 0
How to read a Script. Each chunk is one line:
a short tag of what the models heard in that clip, then the words spoken.
(underlined, plain ·
delivery, style) — the tag before the words. Emotions first, then how
it is delivered. Underlined descriptors are the ones that change across this chain
— anything identical on every clip is pulled out and stated once above, because a
value that never moves says nothing about a trajectory.
(ahem) — brackets inside the words
are a different thing: a real non-speech sound, printed where it happens. Most clips have
none; about a quarter do.
The full generated caption for any clip is still there, under
“full caption & clip details”. Its perceived-gender and background-noise
clauses were re-rendered from the numeric buckets, because the versions stored in the
corpus index had those two ladders running backwards.
The emotion clause has been re-derived, and it used to be wrong. The 40 emotion heads are not on a common scale — Interest has a median of 2.08 and is never zero, while Infatuation is zero on 87.7 % of clips — and the caption named an emotion whenever its raw score cleared an absolute 1.0. Interest therefore appeared in 94.8 % of captions and Sadness in almost none: the clause was reporting the scale of the head, not the emotion of the clip. An emotion is now named only when it lands in the top 10 % for that emotion, against a pooled tie-aware mid-rank ECDF over 132,833,726 utterances spanning every dataset and language. Interest now appears in 6.2 %, all 40 emotions occur, and a clip that is ordinary on all 40 says “no dominant emotion” rather than being forced to pick one (21.8 % of clips). This is the same scale the trajectory miner selects on, so the caption and the mining now refer to the same quantity: the mined target emotion is named in the final clip's caption on 73 % of chains, up from 46 %.
Sadness rising ↑sad-Sadness-S3-k2 · #1
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.09. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.09.
On the corpus-wide percentile scale those become 0.43, 0.98 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 10 s · emolia
hear it un-normalised (raw levels, max seam 3.0 dB)
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, measured, slightly relaxed
(normally alert, fairly steady, little disfluency, authoritative)Hello. Herzlich willkommen aus der Quantum Storm Star Wars Collection. Mein Name ist Deniz und ich möchte euch heute den
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; no dominant emotion; style: authoritative, didactic; good recording, no background noise; genuineness 1.6/6; vocal-burst blend 0.1/10; 6.7s, DE.
DE_-WhG-j-UudU_W000000 · in -19.3 dBFS · gain -0.7 dB · emolia-00238
(helplessness, sadness, longing·subdued, steady, no disfluency, formal)Gummibändern festgemacht, die wir von.
full caption & clip details
An adult masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, neutral openness; reads as helplessness, sadness, longing; style: formal, monologue; good recording, no background noise; genuineness 1.4/6; vocal-burst blend 2.0/10; 3.1s, DE.
DE_-WhG-j-UudU_W000024 · in -16.3 dBFS · gain -3.7 dB · emolia-00238
Sadness rising ↑sad-Sadness-S3-k2 · #2
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.47. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.47.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 33 s · emolia
hear it un-normalised (raw levels, max seam 5.8 dB)
Unchanged across all 2 clips: an adult masculine voice · neutral-bright, fairly smooth, balanced body, average recording, energised, light breath
(normal-paced, slightly relaxed, fairly steady, authoritative)Und ich rufe auf den Tagesordnungspunkt 11.
full caption & clip details
An adult masculine voice; delivery is energised, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; very clear, frequent disfluency, wide pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; no dominant emotion; style: authoritative, formal; average recording, no background noise; genuineness 1.4/6; vocal-burst blend 0.3/10; 3.2s, DE.
DE_-xg8lm69_K8_W000000 · in -18.2 dBFS · gain -1.8 dB · emolia-00224
(disappointment, distress, impatience and irritability·brisk, neutral tension, moderately variable, monologue)Und da würde ich sagen, liegt es nicht nur daran, dass das Wahlprozedere jetzt irgendwie kompliziert wäre. Ich meine, die Studierenden bekommen die Unterlagen geschickt. Es gibt tagelang Zeit, seine Stimme abzugeben. Ich denke, wir müssen auch überlegen, ob es was damit zu tun hat, (ahem) dass die Entscheidungskompetenzen in der Verfasst Studierendenschaft nicht gerade sehr ausgeprägt sind, um es vorsichtig zu sagen. Also, den ganzen Autonomieprozess, den es gab, sind ja Kompetenzen vom Ministerium an die Hochschulen (ahem) (ahem) verlagert worden, aber eben dort vor allem an die Präsidien und an die Hochschulräte. Und ich finde, man muss vielleicht
full caption & clip details
A young adult feminine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, fairly guarded; reads as disappointment, distress, impatience and irritability; style: monologue, dramatic; average recording, some background noise; genuineness 2.1/6; vocal-burst blend 2.7/10; 30.0s, DE.
DE_-xg8lm69_K8_W000052 · in -24.0 dBFS · gain +4.0 dB · emolia-00224
Sadness rising ↑sad-Sadness-S3-k2 · #3
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.24. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.24.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 14 s · emolia
hear it un-normalised (raw levels, max seam 1.3 dB)
Unchanged across all 2 clips: an elderly masculine voice · neutral-toned, dark, slightly thin, below-average recording, quiet background, slow, relaxed, steady
(emotional numbness, contemplation · lethargic, narrow pitch range, whispered, monologue)So, ich würd mal sagen, Willkommen zu einer neuen Aufnahme, außerhalb des Streams.
full caption & clip details
An elderly masculine voice; delivery is lethargic, slow, relaxed, steady; timbre is neutral-toned, dark, smooth, slightly thin; slurred, frequent disfluency, narrow pitch range, audible breath; affect is mildly negative, submissive, neutral openness; reads as emotional numbness, contemplation; style: whispered, monologue; below-average recording, quiet background; genuineness 4.3/6; vocal-burst blend 0.0/10; 9.0s, DE.
DE_02FqQjJyMty_W000000 · in -18.4 dBFS · gain -1.6 dB · emolia-00215
(sadness, helplessness, distress·very low-energy, fairly narrow pitch, whispered, casual)In (low mumble) den Videos habe ich ja nicht mal mit dabei. Würde eigentlich Sinn machen.
full caption & clip details
A young adult masculine voice; delivery is very low-energy, slow, relaxed, steady; timbre is neutral-toned, dark, fairly smooth, slightly thin; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly negative, submissive, neutral openness; reads as sadness, helplessness, distress; style: whispered, casual; below-average recording, quiet background; genuineness 5.7/6; vocal-burst blend 0.8/10; 4.6s, DE.
DE_02FqQjJyMty_W000002 · in -17.1 dBFS · gain -2.9 dB · emolia-00215
Sadness rising ↑sad-Sadness-S3-k2 · #4
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 1.02. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 1.02.
On the corpus-wide percentile scale those become 0.43, 0.98 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 10 s · emolia
hear it un-normalised (raw levels, max seam 6.2 dB)
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, normal-paced, normally alert, slightly relaxed
(fairly steady, some disfluency, conversational, authoritative)auch wenn's nicht stattfindet, ein bisschen Lagerstellen, man darf doch hier aufkommen. Lass uns doch jetzt gemeinsam kurz ein Zelt aufstellen. Los geht's!
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; no dominant emotion; style: conversational, authoritative; average recording, quiet background; genuineness 3.1/6; vocal-burst blend 0.0/10; 6.7s, DE.
DE_0FuG3tlx0yQ_W000000 · in -24.3 dBFS · gain +4.3 dB · emolia-00097
(sadness, triumph, helplessness·moderately variable, no disfluency, authoritative, dramatic)Doch selbst noch die besten Jahre sind voll Kummer und Schmerz.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, no disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as sadness, triumph, helplessness; style: authoritative, dramatic; good recording, quiet background; genuineness 2.3/6; vocal-burst blend 0.7/10; 3.3s, DE.
DE_0FuG3tlx0yQ_W000006 · in -18.1 dBFS · gain -1.9 dB · emolia-00097
Sadness rising ↑sad-Sadness-S3-k2 · #5
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 1.09. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 1.09.
On the corpus-wide percentile scale those become 0.43, 0.98 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 18 s · emolia
hear it un-normalised (raw levels, max seam 2.8 dB)
Unchanged across all 2 clips: a young adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, normal-paced, normally alert, slightly relaxed, some disfluency
(moderately variable, wide pitch range, conversational, playful)Hey Leute, willkommen zurück zu einem neuen Video hier auf Popel mit Zucker. Wie immer fangen wir an mit einem extra Schluck.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is mildly positive, slightly dominant, neutral openness; no dominant emotion; style: conversational, playful; good recording, quiet background; genuineness 3.2/6; vocal-burst blend 0.4/10; 5.6s, DE.
DE_0IWgiqWfPNA_W000000 · in -17.9 dBFS · gain -2.1 dB · emolia-00195
(fatigue exhaustion, pain, sadness·fairly steady, moderate pitch range, monologue, casual)zu leicht, glaube ich, meiner Meinung nach. Hätte man doppelt überlegen müssen, aber klar, man will die Rache, man will ihn endlich loswerden, nach all der Zeit, und dann laut man halt auch so eine, so eine Täuschung mal eher, aber das ist halt,
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as fatigue exhaustion, pain, sadness; style: monologue, casual; average recording, no background noise; genuineness 3.1/6; vocal-burst blend 1.1/10; 12.3s, DE.
DE_0IWgiqWfPNA_W000059 · in -20.7 dBFS · gain +0.7 dB · emolia-00195
Sadness rising ↑sad-Sadness-S3-k2 · #6
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.04, 1.00. In the first clip the scorer found no Sadness whatsoever (0.04); by the last it is at 1.00.
On the corpus-wide percentile scale those become 0.87, 0.98 — a total move of +0.10.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a middle-aged feminine voice · neutral-toned, fairly smooth, balanced body, average recording, quiet background, very low-energy, fairly steady
(doubt, confusion, pain · normal-paced, slightly relaxed, some disfluency, casual)Ist es überhaupt richtig, das Setting, was mir jetzt gerade angeboten wird? Und genau darum geht es in diesem Video. Und zwar um die Aspekte, die man berücksichtigen kann oder sollte, bevor man sich dafür oder dagegen entscheidet.
full caption & clip details
A middle-aged feminine voice; delivery is very low-energy, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, neutral openness; reads as doubt, confusion, pain; style: casual, monologue; average recording, quiet background; genuineness 3.0/6; vocal-burst blend 0.0/10; 16.3s, DE.
DE_0KZXkf63R18_W000001 · in -23.9 dBFS · gain +3.9 dB · emolia-00254
(sadness, helplessness, contemplation·slow, relaxed, frequent disfluency, monologue)Leider ist zurzeit (low mumble) auch unter Therapeuten und Trauma-Behandlenden die Situation folgend, dass (low mumble) bei vielen nicht ausreichend, ja, Wissen und Kenntnisse da sind.
full caption & clip details
A middle-aged somewhat feminine voice; delivery is very low-energy, slow, relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; somewhat unclear, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly positive, submissive, neutral openness; reads as sadness, helplessness, contemplation; style: monologue, ASMR; average recording, quiet background; genuineness 3.6/6; vocal-burst blend 0.0/10; 16.5s, DE.
DE_0KZXkf63R18_W000004 · in -20.6 dBFS · gain +0.6 dB · emolia-00254
Sadness rising ↑sad-Sadness-S3-k2 · #7
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.12. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.12.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, slightly relaxed, some disfluency, average clarity, moderate pitch range, light breath
(fatigue exhaustion · normal-paced, normally alert, moderately variable, casual)Karten und so weiter. Ich mach's gleich auf und zeig's euch richtig, aber ich zeig's einfach nur so kurz und dann.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, thin; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as fatigue exhaustion; style: casual, conversational; below-average recording, quiet background; genuineness 5.0/6; vocal-burst blend 2.8/10; 5.3s, DE.
DE_0PbfumbfjeQ_W000000 · in -19.7 dBFS · gain -0.3 dB · emolia-00161
(helplessness, sadness, distress·measured, subdued, fairly steady, monologue)Jetzt werde ich auf jeden Fall keine weiteren mehr bestellen. Das muss man jetzt erstmal alles verarbeiten, was ich hier.
full caption & clip details
An adult feminine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as helplessness, sadness, distress; style: monologue, ASMR; good recording, no background noise; genuineness 3.1/6; vocal-burst blend 0.7/10; 5.5s, DE.
DE_0PbfumbfjeQ_W000005 · in -18.4 dBFS · gain -1.6 dB · emolia-00161
Sadness rising ↑sad-Sadness-S3-k2 · #8
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.01. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.01.
On the corpus-wide percentile scale those become 0.43, 0.98 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(steady, newsreading, formal)Zusammen wollen wir uns verschiedene Begriffe aus dem queeren und feministischen Spektrum anschauen, Hintergründe zu den Themen beleuchten und Fragen klären.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: newsreading, formal; good recording, no background noise; genuineness 0.1/6; vocal-burst blend 0.2/10; 7.3s, DE.
DE_0S219bZ30GQ_W000000 · in -22.0 dBFS · gain +2.0 dB · emolia-00097
(contempt, sadness, distress·fairly steady, formal, newsreading)Die Begriffe Zwitter und Hermaphrodit sind für die meisten intersexuellen Menschen sehr beleidigend und verletzend. Du solltest sie also nicht benutzen.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as contempt, sadness, distress; style: formal, newsreading; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.2/10; 7.7s, DE.
DE_0S219bZ30GQ_W000012 · in -21.3 dBFS · gain +1.3 dB · emolia-00097
Sadness rising ↑sad-Sadness-S3-k2 · #9
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 1.30. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 1.30.
On the corpus-wide percentile scale those become 0.43, 0.99 — a total move of +0.56.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, normal-paced, normally alert, slightly relaxed
(concentration, hope enthusiasm optimism, contemplation · monologue, formal)Ihr Lieben, (low mumble) für uns Grüne war schon immer klar, die Zukunft, die wir wollen, die müssen wir auch selber erschaffen. Und auch wenn die Folgen der Corona-Krise uns noch lange beschäftigen werden, wenn nicht jetzt alle Probleme, alle Krisen und alle Herausforderungen gleichzeitig angeht, der verspielt die Zukunft der zukünftigen Generation.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as concentration, hope enthusiasm optimism, contemplation; style: monologue, formal; average recording, quiet background; genuineness 3.2/6; vocal-burst blend 0.0/10; 18.3s, DE.
DE_0fB4TCjiFIk_W000000 · in -15.9 dBFS · gain -4.1 dB · emolia-00077
(pride, helplessness, sadness·didactic, monologue)Und wir sind keine kleine Partei mehr. Wir sind wahnsinnig gewachsen. Als Nina und ich als Landesvorsitzende gewählt wurden.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as pride, helplessness, sadness; style: didactic, monologue; good recording, quiet background; genuineness 2.8/6; vocal-burst blend 0.0/10; 6.3s, DE.
DE_0fB4TCjiFIk_W000010 · in -14.9 dBFS · gain -5.1 dB · emolia-00077
Sadness rising ↑sad-Sadness-S3-k2 · #10
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.12. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.12.
On the corpus-wide percentile scale those become 0.43, 0.98 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an elderly somewhat feminine voice · slightly warm, balanced body, good recording, no background noise, slow, very low-energy, relaxed, steady
(sexual lust, fatigue exhaustion, infatuation · no disfluency, slurred, narrow pitch range, whispered)denn die Energien öffnen deinen Geist.
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, relaxed, steady; timbre is slightly warm, dark, smooth, balanced body; slurred, no disfluency, narrow pitch range, audible breath; affect is mildly positive, submissive, neutral openness; reads as sexual lust, fatigue exhaustion, infatuation; style: whispered, monologue; good recording, no background noise; genuineness 1.8/6; vocal-burst blend 5.2/10; 3.9s, DE.
DE_0ytNrKylOrc_W000002 · in -13.5 dBFS · gain -6.5 dB · emolia-00215
(distress, fear, sadness·frequent disfluency, somewhat unclear, fairly narrow pitch, ASMR)dass du als Mensch hier auf der Erde das vergessen hast.
full caption & clip details
A middle-aged somewhat feminine voice; delivery is very low-energy, slow, relaxed, steady; timbre is slightly warm, slightly dark, slightly rough, balanced body; somewhat unclear, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly positive, submissive, vulnerable; reads as distress, fear, sadness; style: ASMR, monologue; good recording, no background noise; genuineness 2.8/6; vocal-burst blend 4.2/10; 6.8s, DE.
DE_0ytNrKylOrc_W000003 · in -15.8 dBFS · gain -4.2 dB · emolia-00215
Sadness rising ↑sad-Sadness-S3-k2 · #11
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 1.67. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 1.67.
On the corpus-wide percentile scale those become 0.43, 1.00 — a total move of +0.57.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a middle-aged somewhat feminine voice · neutral-toned, slightly dark, fairly smooth, balanced body, average recording, quiet background, very low-energy, relaxed
(fatigue exhaustion, confusion, doubt · measured, average clarity, fairly narrow pitch, ASMR)Nachbarin auch mitnehmen und weiter, weiter viertn.
full caption & clip details
A middle-aged somewhat feminine voice; delivery is very low-energy, measured, relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; average clarity, frequent disfluency, fairly narrow pitch, light breath; affect is neutral, submissive, neutral openness; reads as fatigue exhaustion, confusion, doubt; style: ASMR, whispered; average recording, quiet background; genuineness 4.1/6; vocal-burst blend 0.2/10; 4.5s, DE.
DE_12MSkDIQFvk_W000006 · in -16.6 dBFS · gain -3.4 dB · emolia-00003
(sadness, pain, helplessness·slow, somewhat unclear, moderate pitch range, ASMR)Kleine Kinder alle angeblie, alleine geblieben waren. Der älteste war, war sechs Jahre.
full caption & clip details
A middle-aged somewhat feminine voice; delivery is very low-energy, slow, relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, audible breath; affect is neutral, submissive, neutral openness; reads as sadness, pain, helplessness; style: ASMR, whispered; average recording, quiet background; genuineness 3.6/6; vocal-burst blend 0.8/10; 6.3s, DE.
DE_12MSkDIQFvk_W000009 · in -17.0 dBFS · gain -3.0 dB · emolia-00003
Sadness rising ↑sad-Sadness-S3-k2 · #12
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.02. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.02.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(steady, formal, monologue)Startschuss für DSDS. Nicht mehr lange, und Florian Silbereisen-Fans kommen auch bei RTL auf ihre Kosten
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; no dominant emotion; style: formal, monologue; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.0/10; 7.9s, DE.
DE_163DYBR_hRg_W000000 · in -16.4 dBFS · gain -3.6 dB · emolia-00161
(confusion, helplessness, sadness·fairly steady, storytelling, formal)Ich sitze hier und weiß genauso wenig wie ihr, was heute passiert.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, slightly guarded; reads as confusion, helplessness, sadness; style: storytelling, formal; good recording, no background noise; genuineness 0.1/6; vocal-burst blend 0.5/10; 4.1s, DE.
DE_163DYBR_hRg_W000011 · in -16.4 dBFS · gain -3.6 dB · emolia-00161
Sadness rising ↑sad-Sadness-S3-k2 · #13
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.01, 1.21. In the first clip the scorer found no Sadness whatsoever (0.01); by the last it is at 1.21.
On the corpus-wide percentile scale those become 0.87, 0.99 — a total move of +0.12.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
A middle-aged feminine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as shame; style: storytelling, casual; average recording, quiet background; genuineness 2.7/6; vocal-burst blend 0.4/10; 4.1s, DE.
DE_17Ls9FmMIuk_W000001 · in -20.7 dBFS · gain +0.7 dB · emolia-00245
(longing, sadness, infatuation·slow, energised, relaxed, cartoonish)Drink, mein Söhnchen, von meiner Brust. Drink, dann wirst du ein starker Held. Ziehst mit den andern hinaus ins Feld.
full caption & clip details
A child feminine voice; delivery is energised, slow, relaxed, moderately variable; timbre is cool, neutral-bright, smooth, balanced body; slurred, almost no disfluency, very wide pitch range, no audible breath; affect is positive, slightly submissive, neutral openness; reads as longing, sadness, infatuation; style: cartoonish, storytelling; average recording, some background noise; genuineness 0.9/6; vocal-burst blend 2.7/10; 10.5s, DE.
DE_17Ls9FmMIuk_W000002 · in -18.2 dBFS · gain -1.8 dB · emolia-00245
Sadness rising ↑sad-Sadness-S3-k2 · #14
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 1.38. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 1.38.
On the corpus-wide percentile scale those become 0.43, 0.99 — a total move of +0.56.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an elderly somewhat feminine voice · slightly warm, neutral-bright, smooth, balanced body, no background noise, slow, very low-energy, relaxed
(contentment, affection, relief · no disfluency, ASMR, monologue)Es ist mir ein Bedürfnis, heute mit dir zusammen, mit euch allen zusammen.
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, relaxed, steady; timbre is slightly warm, neutral-bright, smooth, balanced body; clear, no disfluency, narrow pitch range, audible breath; affect is mildly positive, submissive, neutral openness; reads as contentment, affection, relief; style: ASMR, monologue; good recording, no background noise; genuineness 1.1/6; vocal-burst blend 2.4/10; 8.7s, DE.
DE_1A7Boronwik_W000000 · in -18.1 dBFS · gain -1.9 dB · emolia-00254
(contemplation, sadness, awe·frequent disfluency, whispered, ASMR)Ein Gebet, welches das transformiert, was so schwer auf unseren Schultern liegt. Ein Gebet, was sein Licht über die ganze Erde erstrahlen lässt und die Menschen dazu bringt, den Frieden wieder einzuladen.
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, relaxed, steady; timbre is slightly warm, neutral-bright, smooth, balanced body; clear, frequent disfluency, narrow pitch range, audible breath; affect is mildly positive, submissive, vulnerable; reads as contemplation, sadness, awe; style: whispered, ASMR; very good recording, no background noise; genuineness 1.3/6; vocal-burst blend 1.0/10; 25.1s, DE.
DE_1A7Boronwik_W000001 · in -18.8 dBFS · gain -1.2 dB · emolia-00254
Sadness rising ↑sad-Sadness-S3-k2 · #15
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 1.20. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 1.20.
On the corpus-wide percentile scale those become 0.43, 0.99 — a total move of +0.56.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a child feminine voice · neutral-toned, neutral-bright, fairly smooth, average recording, quiet background, moderately variable, wide pitch range
(impatience and irritability, anger · normal-paced, normally alert, slightly relaxed, conversational)Ach so. Auf der Suche nach dem a priori.
full caption & clip details
A child feminine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, thin; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly submissive, neutral openness; reads as impatience and irritability, anger; style: conversational, casual; average recording, quiet background; genuineness 4.1/6; vocal-burst blend 1.6/10; 3.1s, DE.
DE_1Pd71w8hAhI_W000001 · in -15.7 dBFS · gain -4.3 dB · emolia-00077
(helplessness, fatigue exhaustion, distress·measured, very low-energy, neutral tension, conversational)(low mumble) euhm, also, ich hab so das Gefühl, ich lass mir da auch Zeit, ja, das drängt mich niemand. (breathy giggle)
full caption & clip details
A young adult feminine voice; delivery is very low-energy, measured, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, frequent disfluency, wide pitch range, normal breath; affect is positive, slightly submissive, vulnerable; reads as helplessness, fatigue exhaustion, distress; style: conversational, casual; average recording, quiet background; genuineness 4.3/6; vocal-burst blend 1.9/10; 6.5s, DE.
DE_1Pd71w8hAhI_W000014 · in -17.7 dBFS · gain -2.3 dB · emolia-00077
Sadness rising ↑sad-Sadness-S3-k2 · #16
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.15. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.15.
On the corpus-wide percentile scale those become 0.43, 0.98 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-bright, fairly smooth, balanced body, average recording
(normal-paced, normally alert, slightly relaxed, authoritative)Antrag der Fraktion Bündnis 90, die GRÜNEN.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, fairly narrow pitch, no audible breath; affect is neutral, dominant, slightly guarded; no dominant emotion; style: authoritative, formal; average recording, quiet background; genuineness 1.5/6; vocal-burst blend 0.6/10; 3.1s, DE.
DE_1ZMIiMMqy9u_W000000 · in -20.9 dBFS · gain +0.9 dB · emolia-00123
(bitterness, sourness, contempt·brisk, highly aroused, tense, authoritative)Dieser Satz in unserem Grundgesetz ist sozusagen der moralische Imperativ, den die Väter und Mütter unseres Grundgesetzes in dieses Grundgesetz geschrieben haben, weil sie die Lehren aus Nationalsozialismus und der Entmenschlichung der Nazidiktatur gezogen haben.
full caption & clip details
An adult masculine voice; delivery is highly aroused, brisk, tense, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, balanced body; very clear, almost no disfluency, wide pitch range, normal breath; affect is elated, dominant, fairly guarded; reads as bitterness, sourness, contempt; style: authoritative, dramatic; average recording, some background noise; genuineness 0.9/6; vocal-burst blend 1.5/10; 16.5s, DE.
DE_1ZMIiMMqy9u_W000006 · in -20.6 dBFS · gain +0.6 dB · emolia-00123
Sadness rising ↑sad-Sadness-S3-k2 · #17
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 1.05. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 1.05.
On the corpus-wide percentile scale those become 0.43, 0.98 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, normal-paced, normally alert, light breath
(teasing, amusement, intoxication altered states of consciousness · neutral tension, moderately variable, some disfluency, conversational)Es, es, es sind auch, (ahem) äh, andere Sachen gemeint. Das war jetzt bloß ein Beispiel. Aber, (childlike giggle) äh, also ich glaube, sexy Klamotten, meinen Klamottenstil nicht nennen. Einfach nur, oh, ist bequem.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly submissive, neutral openness; reads as teasing, amusement, intoxication altered states of consciousness; style: conversational, playful; good recording, quiet background; genuineness 4.5/6; vocal-burst blend 0.9/10; 13.5s, DE.
DE_1c4KtO2dsKE_W000000 · in -12.4 dBFS · gain -7.5 dB · emolia-00167
(fatigue exhaustion, longing, sadness·relaxed, fairly steady, frequent disfluency, casual)Ja, natürlich, also, sie da behalten ist dann auch, ich mein, ja, das ist halt auch nicht meine Entscheidung, wer dann noch.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, light breath; affect is mildly positive, submissive, neutral openness; reads as fatigue exhaustion, longing, sadness; style: casual, monologue; average recording, quiet background; genuineness 5.0/6; vocal-burst blend 0.5/10; 8.8s, DE.
DE_1c4KtO2dsKE_W000013 · in -14.4 dBFS · gain -5.6 dB · emolia-00167
Sadness rising ↑sad-Sadness-S3-k2 · #18
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.85. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.85.
On the corpus-wide percentile scale those become 0.44, 1.00 — a total move of +0.56.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, normal-paced, normally alert, fairly steady
(affection · slightly relaxed, little disfluency, clear, authoritative)Tage der Ermutigung. Ich freue mich, dass du wieder eingeschaltet hast, dass du wieder dabei bist.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, wide pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as affection; style: authoritative, storytelling; good recording, no background noise; genuineness 1.4/6; vocal-burst blend 0.0/10; 4.9s, DE.
DE_1eb_VRL8MFI_W000000 · in -19.3 dBFS · gain -0.7 dB · emolia-00216
(pain, relief, helplessness·neutral tension, some disfluency, average clarity, conversational)Ich wurde so krank, dass ich nicht mehr laufen konnte, dass ich mich nicht mehr bewegen konnte und ich große Schmerzen hatte im ganzen Körper.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, neutral tension, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as pain, relief, helplessness; style: conversational, casual; good recording, quiet background; genuineness 3.1/6; vocal-burst blend 1.3/10; 6.9s, DE.
DE_1eb_VRL8MFI_W000004 · in -20.9 dBFS · gain +0.9 dB · emolia-00216
Sadness rising ↑sad-Sadness-S3-k2 · #19
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.15. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.15.
On the corpus-wide percentile scale those become 0.43, 0.98 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a middle-aged masculine voice · neutral-toned, neutral-bright, balanced body, average recording, quiet background, measured, slightly relaxed, frequent disfluency
(normally alert, fairly steady, didactic, monologue)Wie bereits in einem vergangenen Video erwähnt, ist die Frage nach gut oder schlecht schwer zu beantworten. Diese Frage ist aber sehr wichtig.
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: didactic, monologue; average recording, quiet background; genuineness 1.9/6; vocal-burst blend 0.0/10; 8.4s, DE.
DE_1gZLllrev9c_W000000 · in -19.9 dBFS · gain -0.1 dB · emolia-00259
(bitterness, contemplation, sadness·very low-energy, steady, monologue, narration)Menschen töten ist schlecht, da es die Menschheit näher an die Nicht-Existenz bringt. Die Besiedlung des weiteren Sonnensystems ist gut, da es die Existenzwahrscheinlichkeit der Menschheit nach Katastrophen von planetaren Ausmaßen erhöht. Selbstmord ist schlecht, da es die Menschheit näher an die Nicht-Existenz bringt.
full caption & clip details
A middle-aged masculine voice; delivery is very low-energy, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; average clarity, frequent disfluency, fairly narrow pitch, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as bitterness, contemplation, sadness; style: monologue, narration; average recording, quiet background; explicit content; genuineness 2.4/6; vocal-burst blend 0.0/10; 19.7s, DE.
DE_1gZLllrev9c_W000012 · in -21.6 dBFS · gain +1.6 dB · emolia-00259
Sadness rising ↑sad-Sadness-S3-k2 · #20
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 1.11. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 1.11.
On the corpus-wide percentile scale those become 0.43, 0.98 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a middle-aged masculine voice · neutral-toned, slightly rough, balanced body, average recording, quiet background, slow, relaxed, frequent disfluency
(anger, fatigue exhaustion, jealousy and envy · subdued, steady, whispered, monologue)Nach viel Nachdenken habe ich mich entschlossen, dieses Video vorzubereiten und es als eine Art Podcast vorzutragen. Das wird euch wehtun und es ist mir nicht egal. Nur Fakten zu meinem Kanal und zu mir.
full caption & clip details
A middle-aged masculine voice; delivery is subdued, slow, relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, frequent disfluency, fairly narrow pitch, normal breath; affect is neutral, neutral stance, fairly guarded; reads as anger, fatigue exhaustion, jealousy and envy; style: whispered, monologue; average recording, quiet background; genuineness 1.3/6; vocal-burst blend 0.7/10; 20.2s, DE.
DE_1m-geKYo8XA_W000000 · in -24.1 dBFS · gain +4.2 dB · emolia-00097
(disappointment, doubt, sadness·very low-energy, fairly steady, monologue, whispered)Wer nichts weiß, muss glauben. Aus dem Unwissen oder Wissen heraus handelt man entsprechend. Ich hoffte, dass ihr mit dem Wissen, es ist alles gesagt und gezeigt, handelt, und das war falsch. Mein Fehler war, dass ich glaubte, dieser
full caption & clip details
A middle-aged masculine voice; delivery is very low-energy, slow, relaxed, fairly steady; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; somewhat unclear, frequent disfluency, fairly narrow pitch, normal breath; affect is neutral, neutral stance, fairly guarded; reads as disappointment, doubt, sadness; style: monologue, whispered; average recording, quiet background; genuineness 2.1/6; vocal-burst blend 1.1/10; 24.7s, DE.
DE_1m-geKYo8XA_W000006 · in -23.3 dBFS · gain +3.3 dB · emolia-00097