Distress under rescue rule S3, k=2. generalisation check, gap 0.453. 90.5 % of clips score at or below zero on this emotion and the largest gap on its normalised axis is 0.453 (WIDER than the 0.25 step cap). This rule found 7,875 chains over 40,000 tracks; the strict rule found 0 at k=3.
These are not strict-rule trajectories. They come from a deliberately looser rule, built to recover examples on an emotion the strict rule cannot reach. What rule S3 changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'. What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained. Full explanation →
7,875chains this rule found
0the strict rule found (k=3)
0.453largest gap on the axis
90.5 %clips scoring ≤ 0
How to read a Script. Each chunk is one line:
a short tag of what the models heard in that clip, then the words spoken.
(underlined, plain ·
delivery, style) — the tag before the words. Emotions first, then how
it is delivered. Underlined descriptors are the ones that change across this chain
— anything identical on every clip is pulled out and stated once above, because a
value that never moves says nothing about a trajectory.
(ahem) — brackets inside the words
are a different thing: a real non-speech sound, printed where it happens. Most clips have
none; about a quarter do.
The full generated caption for any clip is still there, under
“full caption & clip details”. Its perceived-gender and background-noise
clauses were re-rendered from the numeric buckets, because the versions stored in the
corpus index had those two ladders running backwards.
The emotion clause has been re-derived, and it used to be wrong. The 40 emotion heads are not on a common scale — Interest has a median of 2.08 and is never zero, while Infatuation is zero on 87.7 % of clips — and the caption named an emotion whenever its raw score cleared an absolute 1.0. Interest therefore appeared in 94.8 % of captions and Sadness in almost none: the clause was reporting the scale of the head, not the emotion of the clip. An emotion is now named only when it lands in the top 10 % for that emotion, against a pooled tie-aware mid-rank ECDF over 132,833,726 utterances spanning every dataset and language. Interest now appears in 6.2 %, all 40 emotions occur, and a clip that is ordinary on all 40 says “no dominant emotion” rather than being forced to pick one (21.8 % of clips). This is the same scale the trajectory miner selects on, so the caption and the mining now refer to the same quantity: the mined target emotion is named in the final clip's caption on 73 % of chains, up from 46 %.
Distress rising ↑sad-Distress-S3-k2 · #1
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.19. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.19.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 25 s · emolia
hear it un-normalised (raw levels, max seam 0.7 dB)
Unchanged across all 2 clips: a young adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, normal-paced, normally alert, slightly relaxed
(hope enthusiasm optimism, elation, interest · casual, conversational)Hallo in die Runde. Ich freue mich sehr, dass ihr dabei seid, mit mir heute wieder über Skat sprechen wollt. Da könnt ihr es auch schon sehen, mir ist demnetzt tatsächlich mal wieder eine ganz interessante Partie über den Weg gelaufen, online, wie ihr seht.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as hope enthusiasm optimism, elation, interest; style: casual, conversational; good recording, quiet background; genuineness 3.3/6; vocal-burst blend 3.8/10; 13.1s, DE.
DE_-YzjWuJuZL8_W000000 · in -17.1 dBFS · gain -2.9 dB · emolia-00238
(doubt, distress, emotional numbness·monologue, casual)sowieso nicht mehr verteidigen könnt gegen keine Verteilung. Diese Karte ist de facto schon tot. Das haben wir eben im Spielverlauf auch gesehen. Kurz vor Schluss musste die Kreuz 10 sowieso angeboten werden.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as doubt, distress, emotional numbness; style: monologue, casual; average recording, quiet background; genuineness 3.7/6; vocal-burst blend 0.0/10; 12.1s, DE.
DE_-YzjWuJuZL8_W000019 · in -17.8 dBFS · gain -2.2 dB · emolia-00238
Distress rising ↑sad-Distress-S3-k2 · #2
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.02, 1.02. In the first clip the scorer found no Distress whatsoever (0.02); by the last it is at 1.02.
On the corpus-wide percentile scale those become 0.89, 0.98 — a total move of +0.09.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 18 s · emolia
hear it un-normalised (raw levels, max seam 2.6 dB)
Unchanged across all 2 clips: a young adult feminine voice · neutral-bright, fairly smooth, normal-paced, normally alert, neutral tension, moderately variable, light breath
(sourness, shame, thankfulness gratitude · some disfluency, average clarity, wide pitch range, conversational)Ich will euch heute zeigen, was ich wirklich innerhalb, wenn man Mutter ist, dann hat man wenig Zeit und was ich wirklich so
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is positive, neutral stance, neutral openness; reads as sourness, shame, thankfulness gratitude; style: conversational, casual; good recording, quiet background; genuineness 4.7/6; vocal-burst blend 2.4/10; 5.9s, DE.
DE_-f6GmCGhSBo_W000000 · in -19.7 dBFS · gain -0.3 dB · emolia-00224
(jealousy and envy, confusion, relief·frequent disfluency, somewhat unclear, fairly narrow pitch, casual)Sorry, love. Also, ihr habt bestimmt auch echt geile Produkte. Jetzt fragt ihr euch wahrscheinlich, wo hast du die denn her? Gibt's die jetzt nur in der Türkei? Nee, die hat mir eine liebe Kohle hingestellt. Irgendwie verklebt, das sieht alles so klumpig aus. Also,
full caption & clip details
A young adult somewhat masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, slightly thin; somewhat unclear, frequent disfluency, fairly narrow pitch, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as jealousy and envy, confusion, relief; style: casual, conversational; below-average recording, some background noise; genuineness 4.4/6; vocal-burst blend 7.6/10; 12.4s, DE.
DE_-f6GmCGhSBo_W000010 · in -22.3 dBFS · gain +2.3 dB · emolia-00224
Distress rising ↑sad-Distress-S3-k2 · #3
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.47. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.47.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 33 s · emolia
hear it un-normalised (raw levels, max seam 5.8 dB)
Unchanged across all 2 clips: an adult masculine voice · neutral-bright, fairly smooth, balanced body, average recording, energised, light breath
(normal-paced, slightly relaxed, fairly steady, authoritative)Und ich rufe auf den Tagesordnungspunkt 11.
full caption & clip details
An adult masculine voice; delivery is energised, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; very clear, frequent disfluency, wide pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; no dominant emotion; style: authoritative, formal; average recording, no background noise; genuineness 1.4/6; vocal-burst blend 0.3/10; 3.2s, DE.
DE_-xg8lm69_K8_W000000 · in -18.2 dBFS · gain -1.8 dB · emolia-00224
(disappointment, distress, impatience and irritability·brisk, neutral tension, moderately variable, monologue)Und da würde ich sagen, liegt es nicht nur daran, dass das Wahlprozedere jetzt irgendwie kompliziert wäre. Ich meine, die Studierenden bekommen die Unterlagen geschickt. Es gibt tagelang Zeit, seine Stimme abzugeben. Ich denke, wir müssen auch überlegen, ob es was damit zu tun hat, (ahem) dass die Entscheidungskompetenzen in der Verfasst Studierendenschaft nicht gerade sehr ausgeprägt sind, um es vorsichtig zu sagen. Also, den ganzen Autonomieprozess, den es gab, sind ja Kompetenzen vom Ministerium an die Hochschulen (ahem) (ahem) verlagert worden, aber eben dort vor allem an die Präsidien und an die Hochschulräte. Und ich finde, man muss vielleicht
full caption & clip details
A young adult feminine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, fairly guarded; reads as disappointment, distress, impatience and irritability; style: monologue, dramatic; average recording, some background noise; genuineness 2.1/6; vocal-burst blend 2.7/10; 30.0s, DE.
DE_-xg8lm69_K8_W000052 · in -24.0 dBFS · gain +4.0 dB · emolia-00224
Distress rising ↑sad-Distress-S3-k2 · #4
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.02. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.02.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 17 s · emolia
hear it un-normalised (raw levels, max seam 1.0 dB)
Unchanged across all 2 clips: an adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, measured, normally alert, slightly relaxed
(doubt · fairly steady, formal, didactic)Du kennst bestimmt so die Situation. Du hast schon lange irgendwas nicht mehr gegessen.
full caption & clip details
An adult feminine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; reads as doubt; style: formal, didactic; good recording, quiet background; genuineness 1.5/6; vocal-burst blend 0.0/10; 5.3s, DE.
DE_-xpBDPzvHII_W000000 · in -18.9 dBFS · gain -1.1 dB · emolia-00182
(jealousy and envy, distress, impatience and irritability·moderately variable, didactic, monologue)Und da kannst du dich doch nur entscheiden von deinem Gefühl her, wie im Supermarkt, wenn du dir ein Müsli kaufst, welche Packen gefällt mir am besten. Und das holst du dir dann und alles andere.
full caption & clip details
A middle-aged feminine voice; delivery is normally alert, measured, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as jealousy and envy, distress, impatience and irritability; style: didactic, monologue; average recording, quiet background; genuineness 1.7/6; vocal-burst blend 0.0/10; 12.0s, DE.
DE_-xpBDPzvHII_W000019 · in -17.9 dBFS · gain -2.1 dB · emolia-00182
Distress rising ↑sad-Distress-S3-k2 · #5
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.24. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.24.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 14 s · emolia
hear it un-normalised (raw levels, max seam 1.3 dB)
Unchanged across all 2 clips: an elderly masculine voice · neutral-toned, dark, slightly thin, below-average recording, quiet background, slow, relaxed, steady
(emotional numbness, contemplation · lethargic, narrow pitch range, whispered, monologue)So, ich würd mal sagen, Willkommen zu einer neuen Aufnahme, außerhalb des Streams.
full caption & clip details
An elderly masculine voice; delivery is lethargic, slow, relaxed, steady; timbre is neutral-toned, dark, smooth, slightly thin; slurred, frequent disfluency, narrow pitch range, audible breath; affect is mildly negative, submissive, neutral openness; reads as emotional numbness, contemplation; style: whispered, monologue; below-average recording, quiet background; genuineness 4.3/6; vocal-burst blend 0.0/10; 9.0s, DE.
DE_02FqQjJyMty_W000000 · in -18.4 dBFS · gain -1.6 dB · emolia-00215
(sadness, helplessness, distress·very low-energy, fairly narrow pitch, whispered, casual)In (low mumble) den Videos habe ich ja nicht mal mit dabei. Würde eigentlich Sinn machen.
full caption & clip details
A young adult masculine voice; delivery is very low-energy, slow, relaxed, steady; timbre is neutral-toned, dark, fairly smooth, slightly thin; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly negative, submissive, neutral openness; reads as sadness, helplessness, distress; style: whispered, casual; below-average recording, quiet background; genuineness 5.7/6; vocal-burst blend 0.8/10; 4.6s, DE.
DE_02FqQjJyMty_W000002 · in -17.1 dBFS · gain -2.9 dB · emolia-00215
Distress rising ↑sad-Distress-S3-k2 · #6
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.12. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.12.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, slightly relaxed, some disfluency, average clarity, moderate pitch range, light breath
(fatigue exhaustion · normal-paced, normally alert, moderately variable, casual)Karten und so weiter. Ich mach's gleich auf und zeig's euch richtig, aber ich zeig's einfach nur so kurz und dann.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, thin; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as fatigue exhaustion; style: casual, conversational; below-average recording, quiet background; genuineness 5.0/6; vocal-burst blend 2.8/10; 5.3s, DE.
DE_0PbfumbfjeQ_W000000 · in -19.7 dBFS · gain -0.3 dB · emolia-00161
(helplessness, sadness, distress·measured, subdued, fairly steady, monologue)Jetzt werde ich auf jeden Fall keine weiteren mehr bestellen. Das muss man jetzt erstmal alles verarbeiten, was ich hier.
full caption & clip details
An adult feminine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as helplessness, sadness, distress; style: monologue, ASMR; good recording, no background noise; genuineness 3.1/6; vocal-burst blend 0.7/10; 5.5s, DE.
DE_0PbfumbfjeQ_W000005 · in -18.4 dBFS · gain -1.6 dB · emolia-00161
Distress rising ↑sad-Distress-S3-k2 · #7
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.36. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.36.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(steady, no disfluency, newsreading, formal)Zusammen wollen wir uns verschiedene Begriffe aus dem queeren und feministischen Spektrum anschauen, Hintergründe zu den Themen beleuchten und Fragen klären.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: newsreading, formal; good recording, no background noise; genuineness 0.1/6; vocal-burst blend 0.2/10; 7.3s, DE.
DE_0S219bZ30GQ_W000000 · in -22.0 dBFS · gain +2.0 dB · emolia-00097
(fear, sadness, pain·fairly steady, little disfluency, narration, formal)Es gibt natürlich auch Fälle, in denen Kinder operiert werden müssen, weil sie sonst starke Schmerzen hätten oder sterben würden. Dagegen haben die Verbände von intersexuellen Menschen auch nichts.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as fear, sadness, pain; style: narration, formal; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.0/10; 9.9s, DE.
DE_0S219bZ30GQ_W000030 · in -22.7 dBFS · gain +2.7 dB · emolia-00097
Distress rising ↑sad-Distress-S3-k2 · #8
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.00. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.00.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, fairly smooth, balanced body, average recording, quiet background, frequent disfluency, average clarity
(contemplation, sourness, bitterness · normal-paced, normally alert, neutral tension, conversational)ist in dem Sinne ein besonderes Projekt gewesen, denn erstmal waren alle Mentorinnen sehr, sehr skeptisch, ob das überhaupt klappt. 嗯, (low mumble) denn, (low mumble) äh, keiner aus dem Team kannte sich so wirklich mit Mindstorms aus. Wir haben alle gesagt, so, ja, nee, äh, (ahem) macht's auch lieber so, nee, könntest auch so machen. Aber nein, sie sind stur geblieben, sind bei ihrer Idee geblieben, 嗯, (low mumble)
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, slightly guarded; reads as contemplation, sourness, bitterness; style: conversational, casual; average recording, quiet background; genuineness 5.4/6; vocal-burst blend 5.5/10; 19.1s, DE.
DE_0uIkaC-rBuI_W000000 · in -17.3 dBFS · gain -2.7 dB · emolia-00254
(helplessness, distress, confusion·slow, very low-energy, relaxed, whispered)Man kann auch alleine spielen oder mit ganz vielen anderen Leuten. Man spielt nach seinen eigenen Regeln und baut Welten, die
full caption & clip details
A middle-aged feminine voice; delivery is very low-energy, slow, relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; average clarity, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly positive, submissive, neutral openness; reads as helplessness, distress, confusion; style: whispered, monologue; average recording, quiet background; genuineness 3.3/6; vocal-burst blend 0.0/10; 10.7s, DE.
DE_0uIkaC-rBuI_W000005 · in -15.2 dBFS · gain -4.8 dB · emolia-00254
Distress rising ↑sad-Distress-S3-k2 · #9
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.08. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.08.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an elderly somewhat feminine voice · slightly warm, dark, smooth, slow, very low-energy, relaxed, steady, frequent disfluency
(sexual lust, contemplation, fatigue exhaustion · somewhat unclear, minimal breath, whispered, ASMR)Und so lasse dich tragen von den Energiewellen, welche sich bereits zu dir bewegen. Es sind die Energiewellen, welche über meinem Kanal zu dir fließen.
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, relaxed, steady; timbre is slightly warm, dark, smooth, balanced body; somewhat unclear, frequent disfluency, narrow pitch range, minimal breath; affect is mildly positive, submissive, neutral openness; reads as sexual lust, contemplation, fatigue exhaustion; style: whispered, ASMR; average recording, quiet background; genuineness 1.7/6; vocal-burst blend 3.3/10; 15.1s, DE.
DE_0ytNrKylOrc_W000000 · in -16.7 dBFS · gain -3.3 dB · emolia-00215
An elderly somewhat feminine voice; delivery is very low-energy, slow, relaxed, steady; timbre is slightly warm, dark, smooth, very full; slurred, frequent disfluency, narrow pitch range, audible breath; affect is mildly positive, submissive, vulnerable; reads as contemplation, longing, relief; style: whispered, ASMR; very good recording, no background noise; genuineness 1.5/6; vocal-burst blend 6.0/10; 4.8s, DE.
DE_0ytNrKylOrc_W000001 · in -16.1 dBFS · gain -3.9 dB · emolia-00215
Distress rising ↑sad-Distress-S3-k2 · #10
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.11. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.11.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an elderly somewhat feminine voice · neutral-toned, slightly dark, slightly rough, very low-energy, relaxed, fairly steady, frequent disfluency, somewhat unclear
(fatigue exhaustion, contemplation, longing · slow, monologue, ASMR)Und uns gewarnt, ja, nicht hinausschauen und ruhig bleiben und, (ahem) äh, und wir Kinder haben das nicht verstanden und,
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, relaxed, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly positive, submissive, neutral openness; reads as fatigue exhaustion, contemplation, longing; style: monologue, ASMR; below-average recording, some background noise; genuineness 3.0/6; vocal-burst blend 0.9/10; 9.5s, DE.
DE_12MSkDIQFvk_W000002 · in -16.8 dBFS · gain -3.2 dB · emolia-00003
(helplessness, fatigue exhaustion, distress·measured, ASMR, whispered)sind wir nicht ruhig geblieben, immer wieder bei die Fenster auße geschaut und beobachtet, was los ist. Bis wir sahen, dass sie uns relieferten,
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, measured, relaxed, fairly steady; timbre is neutral-toned, slightly dark, slightly rough, thin; somewhat unclear, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly negative, submissive, neutral openness; reads as helplessness, fatigue exhaustion, distress; style: ASMR, whispered; average recording, quiet background; genuineness 3.5/6; vocal-burst blend 0.7/10; 11.5s, DE.
DE_12MSkDIQFvk_W000003 · in -19.6 dBFS · gain -0.4 dB · emolia-00003
Distress rising ↑sad-Distress-S3-k2 · #11
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.02. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.02.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(steady, formal, monologue)Startschuss für DSDS. Nicht mehr lange, und Florian Silbereisen-Fans kommen auch bei RTL auf ihre Kosten
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; no dominant emotion; style: formal, monologue; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.0/10; 7.9s, DE.
DE_163DYBR_hRg_W000000 · in -16.4 dBFS · gain -3.6 dB · emolia-00161
(confusion, helplessness, sadness·fairly steady, storytelling, formal)Ich sitze hier und weiß genauso wenig wie ihr, was heute passiert.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is mildly positive, slightly submissive, slightly guarded; reads as confusion, helplessness, sadness; style: storytelling, formal; good recording, no background noise; genuineness 0.1/6; vocal-burst blend 0.5/10; 4.1s, DE.
DE_163DYBR_hRg_W000011 · in -16.4 dBFS · gain -3.6 dB · emolia-00161
Distress rising ↑sad-Distress-S3-k2 · #12
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.15. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.15.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an elderly somewhat feminine voice · slightly warm, smooth, balanced body, no background noise, slow, very low-energy, relaxed, steady
(contentment, affection, relief · no disfluency, clear, narrow pitch range, ASMR)Es ist mir ein Bedürfnis, heute mit dir zusammen, mit euch allen zusammen.
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, relaxed, steady; timbre is slightly warm, neutral-bright, smooth, balanced body; clear, no disfluency, narrow pitch range, audible breath; affect is mildly positive, submissive, neutral openness; reads as contentment, affection, relief; style: ASMR, monologue; good recording, no background noise; genuineness 1.1/6; vocal-burst blend 2.4/10; 8.7s, DE.
DE_1A7Boronwik_W000000 · in -18.1 dBFS · gain -1.9 dB · emolia-00254
(shame, sadness, relief ·frequent disfluency, slurred, fairly narrow pitch, whispered)Und ich bin mir meiner Verantwortung bewusst.
full caption & clip details
An elderly somewhat feminine voice; delivery is very low-energy, slow, relaxed, steady; timbre is slightly warm, dark, smooth, balanced body; slurred, frequent disfluency, fairly narrow pitch, audible breath; affect is mildly positive, submissive, vulnerable; reads as shame, sadness, relief; style: whispered, monologue; very good recording, no background noise; genuineness 1.6/6; vocal-burst blend 3.8/10; 4.6s, DE.
DE_1A7Boronwik_W000006 · in -19.8 dBFS · gain -0.2 dB · emolia-00254
Distress rising ↑sad-Distress-S3-k2 · #13
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.24. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.24.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, quiet background, neutral tension, moderately variable, somewhat unclear, normal breath
(embarrassment, fatigue exhaustion, infatuation · normal-paced, normally alert, some disfluency, casual)Ich drück mich immer falsch auf, aus, in dieser, also, filterraum, ich muss auch mal die richtige Sprache lernen.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, slightly thin; somewhat unclear, some disfluency, moderate pitch range, normal breath; affect is positive, slightly submissive, slightly vulnerable; reads as embarrassment, fatigue exhaustion, infatuation; style: casual, conversational; below-average recording, quiet background; genuineness 5.3/6; vocal-burst blend 4.9/10; 6.3s, DE.
DE_1Pd71w8hAhI_W000000 · in -21.4 dBFS · gain +1.4 dB · emolia-00077
(helplessness, fatigue exhaustion, distress·measured, very low-energy, frequent disfluency, conversational)(low mumble) euhm, also, ich hab so das Gefühl, ich lass mir da auch Zeit, ja, das drängt mich niemand. (breathy giggle)
full caption & clip details
A young adult feminine voice; delivery is very low-energy, measured, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, frequent disfluency, wide pitch range, normal breath; affect is positive, slightly submissive, vulnerable; reads as helplessness, fatigue exhaustion, distress; style: conversational, casual; average recording, quiet background; genuineness 4.3/6; vocal-burst blend 1.9/10; 6.5s, DE.
DE_1Pd71w8hAhI_W000014 · in -17.7 dBFS · gain -2.3 dB · emolia-00077
Distress rising ↑sad-Distress-S3-k2 · #14
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.02. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.02.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, average recording, quiet background, normal-paced, normally alert
(steady, no disfluency, fairly narrow pitch, authoritative)Antrag der Fraktion Bündnis 90, die GRÜNEN.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, fairly narrow pitch, no audible breath; affect is neutral, dominant, slightly guarded; no dominant emotion; style: authoritative, formal; average recording, quiet background; genuineness 1.5/6; vocal-burst blend 0.6/10; 3.1s, DE.
DE_1ZMIiMMqy9u_W000000 · in -20.9 dBFS · gain +0.9 dB · emolia-00123
(disappointment, bitterness, sadness·fairly steady, almost no disfluency, moderate pitch range, newsreading)Wo rechtsextreme Gewalttaten die Gesundheit und das Leben von Migrantinnen und Migranten, von Antifaschistinnen und Antifaschisten bedrohen. 600 gemeldete Straftaten, die in den Bereich politisch motivierte Kriminalität rechts zuzuordnen sind, gab es allein im Jahr 2017, darunter 17 Gewalttaten.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, fairly guarded; reads as disappointment, bitterness, sadness; style: newsreading, formal; average recording, quiet background; genuineness 0.5/6; vocal-burst blend 0.7/10; 21.6s, DE.
DE_1ZMIiMMqy9u_W000072 · in -21.8 dBFS · gain +1.8 dB · emolia-00123
Distress rising ↑sad-Distress-S3-k2 · #15
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.33. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.33.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult masculine voice · neutral-bright, fairly smooth, balanced body, average recording, quiet background, normal-paced, energised, neutral tension
(disgust, interest, teasing · frequent disfluency, very clear, very wide pitch range, storytelling)3G Symmension 3. Ihr wisst Bescheid. Ganz genau. Und heute befinden wir uns, wenn ich das hier alles richtig sehe, und mich gut informiert hab, in Pyra Media. Ja, (ahem) ganz genau.
full caption & clip details
A young adult masculine voice; delivery is energised, normal-paced, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, balanced body; very clear, frequent disfluency, very wide pitch range, normal breath; affect is positive, slightly dominant, neutral openness; reads as disgust, interest, teasing; style: storytelling, playful; average recording, quiet background; genuineness 3.7/6; vocal-burst blend 1.9/10; 14.1s, DE.
DE_1aPh9BCqtRu_W000000 · in -15.4 dBFS · gain -4.6 dB · emolia-00022
(confusion, helplessness, doubt·some disfluency, average clarity, wide pitch range, casual)Was ist das denn? Auf einmal spot eine Kugel und ich kann nicht weiter Staubsaugermeister machen hier. Schön, dass es dafür ein Kristall gibt. Ich weiß nicht, wo ich laufen darf.
full caption & clip details
A young adult masculine voice; delivery is energised, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, normal breath; affect is elated, slightly dominant, neutral openness; reads as confusion, helplessness, doubt; style: casual, dramatic; average recording, quiet background; mildly explicit content; genuineness 3.8/6; vocal-burst blend 0.0/10; 12.3s, DE.
DE_1aPh9BCqtRu_W000028 · in -16.8 dBFS · gain -3.2 dB · emolia-00022
Distress rising ↑sad-Distress-S3-k2 · #16
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.06. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.06.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, quiet background, normal-paced, normally alert
(teasing, amusement, intoxication altered states of consciousness · neutral tension, moderately variable, wide pitch range, conversational)Es, es, es sind auch, (ahem) äh, andere Sachen gemeint. Das war jetzt bloß ein Beispiel. Aber, (childlike giggle) äh, also ich glaube, sexy Klamotten, meinen Klamottenstil nicht nennen. Einfach nur, oh, ist bequem.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly submissive, neutral openness; reads as teasing, amusement, intoxication altered states of consciousness; style: conversational, playful; good recording, quiet background; genuineness 4.5/6; vocal-burst blend 0.9/10; 13.5s, DE.
DE_1c4KtO2dsKE_W000000 · in -12.4 dBFS · gain -7.5 dB · emolia-00167
(shame, sadness, helplessness·slightly relaxed, fairly steady, moderate pitch range, conversational)Ich hab schon mal wissentlich geflirtet, aber mir ist es auch schon mal passiert, dass irgendwie ein Verhalten von mir als flirten gedeutet wurde und von mir da nicht so gemeint war. Wir haben hier noch eine letzte.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as shame, sadness, helplessness; style: conversational, casual; good recording, quiet background; genuineness 2.3/6; vocal-burst blend 0.0/10; 9.9s, DE.
DE_1c4KtO2dsKE_W000024 · in -13.0 dBFS · gain -7.0 dB · emolia-00167
Distress rising ↑sad-Distress-S3-k2 · #17
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.85. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.85.
On the corpus-wide percentile scale those become 0.44, 1.00 — a total move of +0.56.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, normal-paced, normally alert, fairly steady
(affection · slightly relaxed, little disfluency, clear, authoritative)Tage der Ermutigung. Ich freue mich, dass du wieder eingeschaltet hast, dass du wieder dabei bist.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, wide pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as affection; style: authoritative, storytelling; good recording, no background noise; genuineness 1.4/6; vocal-burst blend 0.0/10; 4.9s, DE.
DE_1eb_VRL8MFI_W000000 · in -19.3 dBFS · gain -0.7 dB · emolia-00216
(pain, relief, helplessness·neutral tension, some disfluency, average clarity, conversational)Ich wurde so krank, dass ich nicht mehr laufen konnte, dass ich mich nicht mehr bewegen konnte und ich große Schmerzen hatte im ganzen Körper.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, neutral tension, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as pain, relief, helplessness; style: conversational, casual; good recording, quiet background; genuineness 3.1/6; vocal-burst blend 1.3/10; 6.9s, DE.
DE_1eb_VRL8MFI_W000004 · in -20.9 dBFS · gain +0.9 dB · emolia-00216
Distress rising ↑sad-Distress-S3-k2 · #18
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.36. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.36.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult masculine voice · thin, below-average recording, quiet background, fast, some disfluency
(affection · normally alert, fully relaxed, fairly steady, casual)Rock, natürlich, gib ihm die Frisur. Das war kleiner Jojo.
full caption & clip details
A young adult masculine voice; delivery is normally alert, fast, fully relaxed, fairly steady; timbre is neutral-toned, dark, slightly rough, thin; slurred, some disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, neutral openness; reads as affection; style: casual, conversational; below-average recording, quiet background; genuineness 5.9/6; vocal-burst blend 1.9/10; 3.5s, DE.
DE_1g2ByepregY_W000000 · in -15.3 dBFS · gain -4.7 dB · emolia-00161
(fear, distress, astonishment surprise·highly aroused, tense, moderately variable, casual)Ja. What the fuck? Das Ding ist halt so, du spritzt das Eimer ab und es ist kaputt. So ein Zweig gab es in irgendeinem Modus.
full caption & clip details
A young adult somewhat masculine voice; delivery is highly aroused, fast, tense, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, thin; average clarity, some disfluency, very wide pitch range, normal breath; affect is elated, very dominant, neutral openness; reads as fear, distress, astonishment surprise; style: casual, dramatic; below-average recording, quiet background; genuineness 4.3/6; vocal-burst blend 1.7/10; 7.1s, DE.
DE_1g2ByepregY_W000019 · in -18.5 dBFS · gain -1.5 dB · emolia-00161
Distress rising ↑sad-Distress-S3-k2 · #19
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.50. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.50.
On the corpus-wide percentile scale those become 0.44, 0.99 — a total move of +0.55.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a middle-aged masculine voice · neutral-toned, neutral-bright, balanced body, average recording, quiet background, measured, slightly relaxed, average clarity
(normally alert, fairly steady, frequent disfluency, didactic)Wie bereits in einem vergangenen Video erwähnt, ist die Frage nach gut oder schlecht schwer zu beantworten. Diese Frage ist aber sehr wichtig.
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: didactic, monologue; average recording, quiet background; genuineness 1.9/6; vocal-burst blend 0.0/10; 8.4s, DE.
DE_1gZLllrev9c_W000000 · in -19.9 dBFS · gain -0.1 dB · emolia-00259
(bitterness, helplessness, sourness·subdued, steady, some disfluency, monologue)Dies lässt sich als evolutionärer Prozess betrachten, in dem die Moral überlebt, welche ihren Trägern die besten Voraussetzungen zum Überleben gibt und jede Moral durch beispielsweise Missverständnisse langsam mutiert. Da die Existenz ihrer Träger das Axiom meiner Moral ist, ist sie zwingend immer am besten dafür geeignet und wird sich, ob nun durch mich oder nicht, irgendwann durchsetzen.
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; average clarity, some disfluency, fairly narrow pitch, light breath; affect is neutral, neutral stance, fairly guarded; reads as bitterness, helplessness, sourness; style: monologue, narration; average recording, quiet background; genuineness 1.3/6; vocal-burst blend 0.1/10; 27.7s, DE.
DE_1gZLllrev9c_W000017 · in -19.2 dBFS · gain -0.8 dB · emolia-00259
Distress rising ↑sad-Distress-S3-k2 · #20
This is not a strict-rule trajectory. It comes from rescue rule S3, which exists because the strict rule returns nothing at all for Distress.
The raw scorer output across the chain is 0.00, 1.09. In the first clip the scorer found no Distress whatsoever (0.00); by the last it is at 1.09.
On the corpus-wide percentile scale those become 0.44, 0.98 — a total move of +0.54.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Distress, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Distress sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Absent to present. The chain must start at or below 0.05 (the emotion is absent) and end at or above 1.0 (it is clearly present). The most literal reading of 'from not sad to sad'.
What it costs: Says nothing about the SHAPE of the path -- only that it begins absent and ends present. Any intermediate clip is unconstrained.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult feminine voice · neutral-toned, fairly smooth, balanced body, good recording, normally alert, slightly relaxed, some disfluency, average clarity
(elation, pleasure ecstasy, contentment · brisk, moderately variable, wide pitch range, playful)Hi, hallo, ich bin's Alena. Ich freue mich, dass ich dir heute meine Morgenroutine zeigen darf. Diese Morgenroutine ist jeden Montag bis jeden Freitag, denn ich gehe arbeiten. Mein Tag beginnt um 5.15 Uhr oder auch 5.30 Uhr, je nachdem, ob ich mir noch wie heute die Haare machen muss.
full caption & clip details
An adult feminine voice; delivery is normally alert, brisk, slightly relaxed, moderately variable; timbre is neutral-toned, slightly bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is positive, slightly dominant, neutral openness; reads as elation, pleasure ecstasy, contentment; style: playful, conversational; good recording, quiet background; genuineness 1.8/6; vocal-burst blend 1.4/10; 16.1s, DE.
DE_2R_mAsRw2RI_W000002 · in -13.8 dBFS · gain -6.2 dB · emolia-00258
(fatigue exhaustion, longing, contentment ·normal-paced, fairly steady, moderate pitch range, casual)Dehne ich mich am Morgen immer ganz gerne und entspanne nochmal so ein bisschen. Dehne auch meinen Schultern und Nackenbereich dadurch, dass ich ja einen ganzen Tag am Schreibtisch am Computer sitze auf der Arbeit.
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as fatigue exhaustion, longing, contentment; style: casual, monologue; good recording, no background noise; genuineness 3.5/6; vocal-burst blend 0.7/10; 10.7s, DE.
DE_2R_mAsRw2RI_W000033 · in -13.6 dBFS · gain -6.4 dB · emolia-00258