Sadness under rescue rule S4, k=2. one hop straight across the gap, step cap lifted. 90.0 % of clips score at or below zero on this emotion and the largest gap on its normalised axis is 0.449 (WIDER than the 0.25 step cap). This rule found 19,060 chains over 40,000 tracks; the strict rule found 0 at k=3.
These are not strict-rule trajectories. They come from a deliberately looser rule, built to recover examples on an emotion the strict rule cannot reach. What rule S4 changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop. What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression. Full explanation →
19,060chains this rule found
0the strict rule found (k=3)
0.449largest gap on the axis
90.0 %clips scoring ≤ 0
How to read a Script. Each chunk is one line:
a short tag of what the models heard in that clip, then the words spoken.
(underlined, plain ·
delivery, style) — the tag before the words. Emotions first, then how
it is delivered. Underlined descriptors are the ones that change across this chain
— anything identical on every clip is pulled out and stated once above, because a
value that never moves says nothing about a trajectory.
(ahem) — brackets inside the words
are a different thing: a real non-speech sound, printed where it happens. Most clips have
none; about a quarter do.
The full generated caption for any clip is still there, under
“full caption & clip details”. Its perceived-gender and background-noise
clauses were re-rendered from the numeric buckets, because the versions stored in the
corpus index had those two ladders running backwards.
The emotion clause has been re-derived, and it used to be wrong. The 40 emotion heads are not on a common scale — Interest has a median of 2.08 and is never zero, while Infatuation is zero on 87.7 % of clips — and the caption named an emotion whenever its raw score cleared an absolute 1.0. Interest therefore appeared in 94.8 % of captions and Sadness in almost none: the clause was reporting the scale of the head, not the emotion of the clip. An emotion is now named only when it lands in the top 10 % for that emotion, against a pooled tie-aware mid-rank ECDF over 132,833,726 utterances spanning every dataset and language. Interest now appears in 6.2 %, all 40 emotions occur, and a clip that is ordinary on all 40 says “no dominant emotion” rather than being forced to pick one (21.8 % of clips). This is the same scale the trajectory miner selects on, so the caption and the mining now refer to the same quantity: the mined target emotion is named in the final clip's caption on 73 % of chains, up from 46 %.
Sadness rising ↑sad-Sadness-S4-k2 · #1
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.44. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.44.
On the corpus-wide percentile scale those become 0.43, 0.92 — a total move of +0.49.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 23 s · emolia
hear it un-normalised (raw levels, max seam 0.2 dB)
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, no background noise, normal-paced, normally alert, slightly relaxed
(some disfluency, average clarity, monologue, didactic)Und zwar stellt sich in unserem Fall die Frage, ob denn der Familienrichter, Dr. Markus Bühler,
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: monologue, didactic; good recording, no background noise; genuineness 1.7/6; vocal-burst blend 0.0/10; 6.2s, DE.
DE_--z5fTsHDac_W000000 · in -18.5 dBFS · gain -1.5 dB · emolia-00037
(sourness, malevolence malice, sadness·little disfluency, clear, didactic, monologue)dachte, der Vater ist der Böse und die Mutter ist die Gute, die mir auch das Essen, wahrscheinlich die Spätzle mit Soße auf den Tisch stellt. Und er hat dann ein Hass gegen seinen Vater entwickelt und eben nicht gegen die Mutter, wobei wahrscheinlich die Mutter den Vater benutzt hat.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as sourness, malevolence malice, sadness; style: didactic, monologue; average recording, no background noise; genuineness 0.6/6; vocal-burst blend 0.3/10; 16.5s, DE.
DE_--z5fTsHDac_W000015 · in -18.3 dBFS · gain -1.7 dB · emolia-00037
Sadness rising ↑sad-Sadness-S4-k2 · #2
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.79. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.79.
On the corpus-wide percentile scale those become 0.43, 0.96 — a total move of +0.53.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 16 s · emolia
hear it un-normalised (raw levels, max seam 3.9 dB)
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, normally alert, slightly relaxed, fairly steady, some disfluency
(anger · measured, somewhat unclear, monologue)Ja, meine Damen, meine Herren, die FDP-Bundestagsfraktion hat natürlich den Fall Nawalny im Zusammenhang auch mit den deutsch-russischen Beziehungen heute (ahem) diskutiert. (low mumble)
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as anger; style: monologue; average recording, no background noise; genuineness 3.3/6; vocal-burst blend 0.0/10; 11.1s, DE.
DE_-05ZmfVq_H8_W000000 · in -16.3 dBFS · gain -3.7 dB · emolia-00037
(sadness, disappointment, anger ·normal-paced, average clarity, conversational, casual)Etwas, das man durchaus spürt im russischen Staatsapparat. Also, ich würde das jetzt nicht einfach so.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as sadness, disappointment, anger; style: conversational, casual; good recording, quiet background; genuineness 3.8/6; vocal-burst blend 0.0/10; 5.2s, DE.
DE_-05ZmfVq_H8_W000009 · in -12.4 dBFS · gain -7.6 dB · emolia-00037
Sadness rising ↑sad-Sadness-S4-k2 · #3
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.41. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.41.
On the corpus-wide percentile scale those become 0.43, 0.91 — a total move of +0.48.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 24 s · emolia
hear it un-normalised (raw levels, max seam 1.0 dB)
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, normal-paced, normally alert, slightly relaxed, fairly steady
(no disfluency, clear, didactic, formal)Tokens sollten ja schon am 17. bzw. 18. Dezember
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: didactic, formal; good recording, no background noise; genuineness 1.5/6; vocal-burst blend 0.0/10; 4.4s, DE.
DE_-6F1Ss25DJQ_W000000 · in -19.4 dBFS · gain -0.7 dB · emolia-00037
(emotional numbness, fear, interest·some disfluency, average clarity, monologue, casual)(low mumble) eh, dadurch soll also entsprechend die Inflation ein bisschen gestoppt werden, ist aber im Grunde alles Quatsch, denn im, gerade am Anfang ist die Inflation sicherlich höher als das, was hier reinfließen wird, geht ja auch nicht anders, ist ja jetzt erst am Anfang, (low mumble) es können ja im Prinzip nicht mehr Tokens, (low mumble) gelockt und gestaked werden als,
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as emotional numbness, fear, interest; style: monologue, casual; average recording, quiet background; genuineness 4.7/6; vocal-burst blend 1.8/10; 19.4s, DE.
DE_-6F1Ss25DJQ_W000122 · in -20.4 dBFS · gain +0.4 dB · emolia-00037
Sadness rising ↑sad-Sadness-S4-k2 · #4
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.74. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.74.
On the corpus-wide percentile scale those become 0.43, 0.95 — a total move of +0.52.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 40 s · emolia
hear it un-normalised (raw levels, max seam 1.7 dB)
Unchanged across all 2 clips: a middle-aged feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, normal-paced, normally alert, slightly relaxed
(fear, anger, relief · clear, monologue, casual)Ich würde sagen, da ist auf jeden Fall Bedarf der Verbesserung, dass man sich nicht schämen muss, in diesem Beruf zu arbeiten. Es lohnt sich auf jeden Fall, weil der Schmutz stirbt nicht aus.
full caption & clip details
A middle-aged feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as fear, anger, relief; style: monologue, casual; good recording, quiet background; genuineness 2.0/6; vocal-burst blend 0.0/10; 10.1s, DE.
DE_-8Xv740j-8Q_W000000 · in -17.5 dBFS · gain -2.5 dB · emolia-00037
(bitterness, jealousy and envy, pain·average clarity, monologue, authoritative)Die erste Kundin ist bereits in Beratung. Die ersten Gespräche sind angebracht und damit wird auch der Webauftritt noch weiter dann optimiert. Und wir wünschen jetzt dir auch viel Erfolg dabei, wenn du dich in der Reinigungsbranche selbstständig machen möchtest oder jemand kennst. Du kannst gerne den Link unterhalb des Videos nutzen für ein kostenfreies Erstgespräch. Wir würden dann gemeinsam als Team zunächst sicherstellen, dass der AVGS dafür beantragt wird. Was ist ein AVGS?
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as bitterness, jealousy and envy, pain; style: monologue, authoritative; average recording, quiet background; genuineness 1.8/6; vocal-burst blend 1.3/10; 30.0s, DE.
DE_-8Xv740j-8Q_W000047 · in -19.1 dBFS · gain -0.9 dB · emolia-00037
Sadness rising ↑sad-Sadness-S4-k2 · #5
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 0.49. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 0.49.
On the corpus-wide percentile scale those become 0.43, 0.92 — a total move of +0.49.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
2 clips · 18 s · emolia
hear it un-normalised (raw levels, max seam 2.1 dB)
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, normal-paced, normally alert, slightly relaxed
(average clarity, conversational, authoritative)Und jetzt der Wochendurchblick mit Florian Streibl. Liebe Zuschauerinnen und Zuschauer, willkommen zum Wochendurchblick.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: conversational, authoritative; good recording, quiet background; genuineness 1.6/6; vocal-burst blend 0.3/10; 6.5s, DE.
DE_-8c1BTgFP3k_W000000 · in -18.5 dBFS · gain -1.5 dB · emolia-00247
(jealousy and envy, disappointment, bitterness·clear, monologue, didactic)Doch angesichts der Aggressivität mit der Russlandspräsident Putin den Westen geradezu herausfordert, hätte ich das so klar und energisch eher vom Bundeskanzler Scholz erwartet.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as jealousy and envy, disappointment, bitterness; style: monologue, didactic; good recording, no background noise; genuineness 1.6/6; vocal-burst blend 0.0/10; 11.4s, DE.
DE_-8c1BTgFP3k_W000013 · in -20.5 dBFS · gain +0.5 dB · emolia-00247
Sadness rising ↑sad-Sadness-S4-k2 · #6
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 0.00. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 0.00.
On the corpus-wide percentile scale those become 0.43, 0.86 — a total move of +0.43.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(contempt, disgust, sourness · almost no disfluency, formal, newsreading)Sehr geehrte Frau Präsidentin, als Verfasser der Stellungnahme des Auswärtigen Ausschusses will ich in erster Linie über die externen Aspekte unseres Berichtes sprechen.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as contempt, disgust, sourness; style: formal, newsreading; good recording, no background noise; genuineness 1.5/6; vocal-burst blend 0.0/10; 8.3s, DE.
DE_-ArmcowDha4_W000000 · in -17.8 dBFS · gain -2.2 dB · emolia-00224
(pride·no disfluency, formal, authoritative)Dazu gehört auch der weltweite Kampf für den Klimaschutz.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as pride; style: formal, authoritative; good recording, no background noise; genuineness 1.2/6; vocal-burst blend 0.0/10; 3.3s, DE.
DE_-ArmcowDha4_W000003 · in -17.0 dBFS · gain -3.0 dB · emolia-00224
Sadness rising ↑sad-Sadness-S4-k2 · #7
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.38. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.38.
On the corpus-wide percentile scale those become 0.43, 0.91 — a total move of +0.48.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a middle-aged masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, normal-paced, normally alert, slightly relaxed
(emotional numbness, helplessness, fear · some disfluency, formal, storytelling)Egal was Sie machen, Ihre VR-Brille möchte einfach nicht über Ihre Seehilfe passen.
full caption & clip details
A middle-aged masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as emotional numbness, helplessness, fear; style: formal, storytelling; good recording, quiet background; genuineness 2.2/6; vocal-burst blend 0.0/10; 4.7s, DE.
DE_-Jk4ZBOnugM_W000000 · in -16.3 dBFS · gain -3.7 dB · emolia-00147
(contemplation, concentration, helplessness ·frequent disfluency, monologue, casual)Wenn man abends nochmal kurz vorm Bett die Brille dann anzieht, ja, hat man sonst eventuell Probleme einzuschlafen und das soll dann auch noch minimiert werden dadurch, durch dieses Blutprotekt. Von daher, wie gesagt, muss man selbst entscheiden, man braucht es nicht.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, frequent disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as contemplation, concentration, helplessness; style: monologue, casual; average recording, quiet background; genuineness 5.4/6; vocal-burst blend 0.0/10; 16.4s, DE.
DE_-Jk4ZBOnugM_W000074 · in -18.1 dBFS · gain -1.9 dB · emolia-00147
Sadness rising ↑sad-Sadness-S4-k2 · #8
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.66. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.66.
On the corpus-wide percentile scale those become 0.43, 0.94 — a total move of +0.51.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normally alert, slightly relaxed
(disappointment · brisk, dramatic, storytelling)Diese Woche gibt es unseren Beitrag zum Thema verlassene Orte etwas später wie gewohnt, aber immer noch pünktlich zum Mittwoch und daran wird sich auch in der Zukunft nichts ändern.
full caption & clip details
A young adult masculine voice; delivery is normally alert, brisk, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as disappointment; style: dramatic, storytelling; good recording, no background noise; genuineness 1.7/6; vocal-burst blend 1.2/10; 8.9s, DE.
DE_-NcjdeCDcbY_W000000 · in -20.7 dBFS · gain +0.7 dB · emolia-00182
(disgust, fear, sourness·normal-paced, formal, monologue)im ehemaligen Eingangsbereich zur Straßenseite hatte es vor einigen Jahren stark gebrannt weswegen es aufgrund der Witterung zu einem Dach- und Deckendurchbruch kam somit konnte man gerade den vorderen Teil gar nicht mehr betreten
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as disgust, fear, sourness; style: formal, monologue; good recording, no background noise; genuineness 0.3/6; vocal-burst blend 0.1/10; 12.2s, DE.
DE_-NcjdeCDcbY_W000007 · in -21.3 dBFS · gain +1.3 dB · emolia-00182
Sadness rising ↑sad-Sadness-S4-k2 · #9
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.23. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.23.
On the corpus-wide percentile scale those become 0.43, 0.89 — a total move of +0.46.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(emotional numbness · formal, didactic)Anfang September. Der Rübenlagerplatz der Kautfabrik ist noch leer. Die Absetzanlage außer Betrieb.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as emotional numbness; style: formal, didactic; good recording, no background noise; genuineness 0.2/6; vocal-burst blend 0.2/10; 5.8s, DE.
DE_-QLrfFrKsMI_W000000 · in -19.8 dBFS · gain -0.2 dB · emolia-00147
(formal, newsreading)Jedem Rübenbauern stehen vier Prozent seiner gesamten Liefermenge an zerkleinerten Rückständen zu. Sie finden besonders als hochwertiges Viehfutterverwendung.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: formal, newsreading; good recording, no background noise; genuineness 0.0/6; vocal-burst blend 0.0/10; 8.8s, DE.
DE_-QLrfFrKsMI_W000087 · in -16.9 dBFS · gain -3.1 dB · emolia-00147
Sadness rising ↑sad-Sadness-S4-k2 · #10
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.09. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.09.
On the corpus-wide percentile scale those become 0.43, 0.88 — a total move of +0.45.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult somewhat feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, average recording, quiet background, slightly relaxed, fairly steady
(malevolence malice · measured, subdued, frequent disfluency, didactic)Heute möchte ich ein Seifenblasenpapier machen. Das heißt, ich puste Seifenblasen, bunte Seifenblasen auf ein Papier, die darauf platzen und dort ihre farbige Spur hinterlassen.
full caption & clip details
A young adult somewhat feminine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, frequent disfluency, fairly narrow pitch, normal breath; affect is mildly positive, neutral stance, slightly guarded; reads as malevolence malice; style: didactic, whispered; average recording, quiet background; genuineness 1.9/6; vocal-burst blend 0.5/10; 15.6s, DE.
DE_-U2hZsyyyRu_W000000 · in -22.5 dBFS · gain +2.5 dB · emolia-00147
(emotional numbness, fatigue exhaustion, disappointment·normal-paced, normally alert, some disfluency, casual)Ich habe da eine Batterie eingelegt, aber das Ganze ist eine Fehlkonstruktion. Hier sind so Stege in der Abdeckung und es lässt sich dann jetzt hier nicht mehr drüber schieben und befestigen. Also als Kinderspielzeug ungeeignet, aber
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, neutral openness; reads as emotional numbness, fatigue exhaustion, disappointment; style: casual, monologue; average recording, quiet background; genuineness 3.7/6; vocal-burst blend 0.0/10; 13.8s, DE.
DE_-U2hZsyyyRu_W000023 · in -17.9 dBFS · gain -2.1 dB · emolia-00147
Sadness rising ↑sad-Sadness-S4-k2 · #11
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 0.05. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 0.05.
On the corpus-wide percentile scale those become 0.43, 0.87 — a total move of +0.44.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, balanced body, good recording, measured, slightly relaxed, fairly steady, moderate pitch range
(normally alert, little disfluency, average clarity, authoritative)Hello. Herzlich willkommen aus der Quantum Storm Star Wars Collection. Mein Name ist Deniz und ich möchte euch heute den
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, little disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; no dominant emotion; style: authoritative, didactic; good recording, no background noise; genuineness 1.6/6; vocal-burst blend 0.1/10; 6.7s, DE.
DE_-WhG-j-UudU_W000000 · in -19.3 dBFS · gain -0.7 dB · emolia-00238
(fear, malevolence malice, sourness·subdued, some disfluency, somewhat unclear, monologue)Dann können die Storytrooper sich auch hinsetzen. Naja, auf so einer langen Fahrt vielleicht auch gar nicht schlecht.
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; somewhat unclear, some disfluency, moderate pitch range, light breath; affect is mildly negative, neutral stance, slightly guarded; reads as fear, malevolence malice, sourness; style: monologue, conversational; good recording, quiet background; genuineness 2.7/6; vocal-burst blend 0.7/10; 5.1s, DE.
DE_-WhG-j-UudU_W000056 · in -20.5 dBFS · gain +0.5 dB · emolia-00238
Sadness rising ↑sad-Sadness-S4-k2 · #12
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 0.26. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 0.26.
On the corpus-wide percentile scale those become 0.43, 0.90 — a total move of +0.47.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, normal-paced, normally alert, fairly steady
(hope enthusiasm optimism, elation, interest · slightly relaxed, casual, conversational)Hallo in die Runde. Ich freue mich sehr, dass ihr dabei seid, mit mir heute wieder über Skat sprechen wollt. Da könnt ihr es auch schon sehen, mir ist demnetzt tatsächlich mal wieder eine ganz interessante Partie über den Weg gelaufen, online, wie ihr seht.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is mildly positive, slightly dominant, slightly guarded; reads as hope enthusiasm optimism, elation, interest; style: casual, conversational; good recording, quiet background; genuineness 3.3/6; vocal-burst blend 3.8/10; 13.1s, DE.
DE_-YzjWuJuZL8_W000000 · in -17.1 dBFS · gain -2.9 dB · emolia-00238
(relief, affection, hope enthusiasm optimism ·neutral tension, casual, monologue)Nützen tut's auch nicht. Spaß am Skat hab ich trotzdem nicht verloren. War aber, das muss ich jetzt sagen, tatsächlich richtig eine Therapiesitzung hier heute für mich. Damit bin ich jetzt auch durch. Hoffe, ihr hattet euren Spaß. Hoffe, wir sehen uns bald mit weiteren interessanten Verteilungen. Wünsche euch noch einen hervorragenden Rest Wochenende und eine schöne Woche. Bis bald. Gut Blatt.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, neutral tension, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as relief, affection, hope enthusiasm optimism; style: casual, monologue; average recording, quiet background; genuineness 4.0/6; vocal-burst blend 3.7/10; 20.0s, DE.
DE_-YzjWuJuZL8_W000039 · in -18.2 dBFS · gain -1.8 dB · emolia-00238
Sadness rising ↑sad-Sadness-S4-k2 · #13
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.07. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.07.
On the corpus-wide percentile scale those become 0.43, 0.88 — a total move of +0.45.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, balanced body, quiet background, slightly relaxed, fairly steady, some disfluency, average clarity
(normal-paced, normally alert, storytelling, monologue)Wir machen eine Geisbergrunde heute miteinander. Man kann um den ganzen Geisberg rum spazieren. Und das Video, das werden wir heute rund um den Geisberg machen.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: storytelling, monologue; good recording, quiet background; genuineness 1.9/6; vocal-burst blend 1.0/10; 9.6s, DE.
DE_-bnJC1eLO5I_W000000 · in -18.2 dBFS · gain -1.8 dB · emolia-00113
(relief, contemplation, pride·measured, subdued, monologue, casual)Und wenn nachher schaust, die Dinge regeln sich oft ganz für selber, oder du warst dem Moment schon was zu tun ist, oder du warst es zu einem späteren Zeitpunkt. Und im Nachhinein sagst du immer, wa, eigentlich ist ganz gut, dass das passiert ist, weil sonst hätte das andere, viel Größere, nicht passieren können.
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, slightly rough, balanced body; average clarity, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as relief, contemplation, pride; style: monologue, casual; average recording, quiet background; genuineness 3.4/6; vocal-burst blend 3.5/10; 18.7s, DE.
DE_-bnJC1eLO5I_W000029 · in -18.5 dBFS · gain -1.5 dB · emolia-00113
Sadness rising ↑sad-Sadness-S4-k2 · #14
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 1.33. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 1.33.
On the corpus-wide percentile scale those become 0.39, 0.99 — a total move of +0.59.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, normally alert, slightly relaxed, fairly steady
(disgust, concentration, infatuation · brisk, clear, newsreading, formal)Am Freitag startet der CDU-Bundesparteitag in Leipzig. Der CDU-Bundesparteitag ist das höchste beschlussfassende Gremium der CDU Deutschlands. Was sich kompliziert anhört, ist eigentlich ganz einfach. Dort werden die wichtigsten inhaltlichen und personellen Entscheidungen gekriegt.
full caption & clip details
An adult masculine voice; delivery is normally alert, brisk, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as disgust, concentration, infatuation; style: newsreading, formal; good recording, no background noise; genuineness 0.4/6; vocal-burst blend 0.1/10; 15.6s, DE.
DE_-e7x_OIHd0g_W000000 · in -19.9 dBFS · gain -0.1 dB · emolia-00247
(bitterness, disappointment, triumph·normal-paced, average clarity, monologue, authoritative)Aus ganz Deutschland kommen rund 1000 Delegierte nach Leipzig.
full caption & clip details
A young adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, almost no disfluency, moderate pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; reads as bitterness, disappointment, triumph; style: monologue, authoritative; good recording, quiet background; genuineness 0.0/6; vocal-burst blend 1.9/10; 30.0s, DE.
DE_-e7x_OIHd0g_W000001 · in -18.8 dBFS · gain -1.2 dB · emolia-00247
Sadness rising ↑sad-Sadness-S4-k2 · #15
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.70. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.70.
On the corpus-wide percentile scale those become 0.43, 0.95 — a total move of +0.52.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, quiet background, normal-paced, normally alert, neutral tension
(sourness, shame, thankfulness gratitude · some disfluency, average clarity, wide pitch range, conversational)Ich will euch heute zeigen, was ich wirklich innerhalb, wenn man Mutter ist, dann hat man wenig Zeit und was ich wirklich so
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; average clarity, some disfluency, wide pitch range, light breath; affect is positive, neutral stance, neutral openness; reads as sourness, shame, thankfulness gratitude; style: conversational, casual; good recording, quiet background; genuineness 4.7/6; vocal-burst blend 2.4/10; 5.9s, DE.
DE_-f6GmCGhSBo_W000000 · in -19.7 dBFS · gain -0.3 dB · emolia-00224
(confusion, longing, sadness·frequent disfluency, somewhat unclear, moderate pitch range, conversational)(exhausted groan) euh, ich hab dann noch, in der Zeit hab ich auch noch Sami (low mumble) angemacht, von wegen, er hätte mir ja nicht geholfen, oder, (low mumble) euh,
full caption & clip details
A young adult feminine voice; delivery is normally alert, normal-paced, neutral tension, moderately variable; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, normal breath; affect is neutral, neutral stance, neutral openness; reads as confusion, longing, sadness; style: conversational, casual; average recording, quiet background; genuineness 4.5/6; vocal-burst blend 2.9/10; 6.9s, DE.
DE_-f6GmCGhSBo_W000013 · in -26.5 dBFS · gain +6.5 dB · emolia-00224
Sadness rising ↑sad-Sadness-S4-k2 · #16
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.36. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.36.
On the corpus-wide percentile scale those become 0.43, 0.91 — a total move of +0.48.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, no background noise, normal-paced, normally alert
(concentration, pride · almost no disfluency, formal, newsreading)Mehr Tempo auf dem Weg zur Klimaneutralität, das fordert die Organisation für wirtschaftliche Entwicklung und Zusammenarbeit in ihrem neuesten Umweltprüfbericht von der Bundesregierung. Ein Appell der OECD lautet deshalb mehr E-Mobilität und mehr Verkehr auf die Schiene.
full caption & clip details
An adult feminine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, almost no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as concentration, pride; style: formal, newsreading; good recording, no background noise; genuineness 0.6/6; vocal-burst blend 0.0/10; 16.4s, DE.
DE_-ffmw_U0FWY_W000000 · in -16.3 dBFS · gain -3.7 dB · emolia-00147
(emotional numbness, fear, sadness·no disfluency, formal, newsreading)denn Sturzfluten wie im Ahrtal werden häufiger. Sie haben zwischen den Jahren 2000 und 2021 insgesamt einen Schaden von 71 Milliarden Euro verursacht.
full caption & clip details
An adult masculine voice; delivery is normally alert, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, no disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, slightly guarded; reads as emotional numbness, fear, sadness; style: formal, newsreading; good recording, no background noise; genuineness 0.1/6; vocal-burst blend 0.2/10; 9.7s, DE.
DE_-ffmw_U0FWY_W000007 · in -15.5 dBFS · gain -4.5 dB · emolia-00147
Sadness rising ↑sad-Sadness-S4-k2 · #17
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.00. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.00.
On the corpus-wide percentile scale those become 0.43, 0.86 — a total move of +0.43.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: a young adult masculine voice · neutral-toned, fairly smooth, balanced body, average recording, quiet background, measured, fairly steady, somewhat unclear
(normally alert, relaxed, frequent disfluency, casual)Der war sogar weg, also so hoch stand da das Wasser, (low mumble) also knapp zwei Meter hoch. (low mumble) Das war nicht so eine gute Sache.
full caption & clip details
A young adult masculine voice; delivery is normally alert, measured, relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; somewhat unclear, frequent disfluency, moderate pitch range, normal breath; affect is neutral, submissive, neutral openness; no dominant emotion; style: casual, conversational; average recording, quiet background; genuineness 4.4/6; vocal-burst blend 0.0/10; 8.3s, DE.
DE_-iFXnOKFbJk_W000003 · in -24.9 dBFS · gain +4.9 dB · emolia-00095
(disappointment, fatigue exhaustion, helplessness·subdued, slightly relaxed, some disfluency, monologue)Schön aktiv an diesen ganzen Themen und nutzen in der Regel LoRaWan dafür, um die Daten da zu übertragen. Genau. Wenn es keine Fragen gibt, wäre ich fertig. Vielen Dank.
full caption & clip details
An adult masculine voice; delivery is subdued, measured, slightly relaxed, fairly steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; somewhat unclear, some disfluency, fairly narrow pitch, normal breath; affect is neutral, neutral stance, slightly guarded; reads as disappointment, fatigue exhaustion, helplessness; style: monologue, whispered; average recording, quiet background; genuineness 2.7/6; vocal-burst blend 0.0/10; 10.6s, DE.
DE_-iFXnOKFbJk_W000081 · in -22.6 dBFS · gain +2.6 dB · emolia-00095
Sadness rising ↑sad-Sadness-S4-k2 · #18
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 0.73. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 0.73.
On the corpus-wide percentile scale those become 0.43, 0.95 — a total move of +0.52.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-bright, fairly smooth, average recording, energised, wide pitch range
(normal-paced, slightly relaxed, fairly steady, authoritative)Und ich rufe auf den Tagesordnungspunkt 11.
full caption & clip details
An adult masculine voice; delivery is energised, normal-paced, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; very clear, frequent disfluency, wide pitch range, light breath; affect is neutral, slightly dominant, slightly guarded; no dominant emotion; style: authoritative, formal; average recording, no background noise; genuineness 1.4/6; vocal-burst blend 0.3/10; 3.2s, DE.
DE_-xg8lm69_K8_W000000 · in -18.2 dBFS · gain -1.8 dB · emolia-00224
(sourness, interest, disappointment·brisk, neutral tension, moderately variable, dramatic)Warum hat man dieses Gefühl? Wann hat man dieses Gefühl, wenn man das Gefühl hat, dass es einen Unterschied macht, wem man sozusagen mit seiner Stimme beauftragt, für die eigenen Anliegen einzustehen? Und an dieser Stelle, glaube ich, müssen wir auch darüber unterhalten, wie können wir dieses Gefühl weiter stärken. Das ist (ahem) eine insgesamt Verantwortung, der wir nachgehen wollen, insgesamt mit der hrg-Novelle. Und insofern freue ich mich auf die Diskussion dort im Sinne von einer größeren Demokratisierung aufgrund (ahem) im Sinne einer weiteren
full caption & clip details
A young adult feminine voice; delivery is energised, brisk, neutral tension, moderately variable; timbre is slightly cool, neutral-bright, fairly smooth, thin; average clarity, some disfluency, wide pitch range, normal breath; affect is neutral, slightly dominant, fairly guarded; reads as sourness, interest, disappointment; style: dramatic, ranting; average recording, some background noise; genuineness 2.2/6; vocal-burst blend 2.1/10; 28.1s, DE.
DE_-xg8lm69_K8_W000058 · in -21.1 dBFS · gain +1.1 dB · emolia-00224
Sadness rising ↑sad-Sadness-S4-k2 · #19
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is -0.00, 0.00. In the first clip the scorer found no Sadness whatsoever (-0.00); by the last it is at 0.00.
On the corpus-wide percentile scale those become 0.43, 0.86 — a total move of +0.43.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult feminine voice · neutral-toned, neutral-bright, fairly smooth, balanced body, good recording, measured, normally alert, slightly relaxed
(doubt · some disfluency, formal, didactic)Du kennst bestimmt so die Situation. Du hast schon lange irgendwas nicht mehr gegessen.
full caption & clip details
An adult feminine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, some disfluency, moderate pitch range, light breath; affect is neutral, neutral stance, neutral openness; reads as doubt; style: formal, didactic; good recording, quiet background; genuineness 1.5/6; vocal-burst blend 0.0/10; 5.3s, DE.
DE_-xpBDPzvHII_W000000 · in -18.9 dBFS · gain -1.1 dB · emolia-00182
(confusion, disappointment, distress·little disfluency, monologue, narration)Und dann kaufst du's, aber wie das hergestellt wurde, das interessiert uns doch nicht. Wir haben uns dafür entschieden, schmeckt oder schmeckt nicht, ja.
full caption & clip details
A middle-aged somewhat feminine voice; delivery is normally alert, measured, slightly relaxed, fairly steady; timbre is neutral-toned, neutral-bright, fairly smooth, balanced body; clear, little disfluency, moderate pitch range, light breath; affect is mildly positive, neutral stance, slightly guarded; reads as confusion, disappointment, distress; style: monologue, narration; good recording, no background noise; genuineness 1.5/6; vocal-burst blend 0.6/10; 9.0s, DE.
DE_-xpBDPzvHII_W000021 · in -19.0 dBFS · gain -1.0 dB · emolia-00182
Sadness rising ↑sad-Sadness-S4-k2 · #20
This is not a strict-rule trajectory. It comes from rescue rule S4, which exists because the strict rule returns nothing at all for Sadness.
The raw scorer output across the chain is 0.00, 0.17. In the first clip the scorer found no Sadness whatsoever (0.00); by the last it is at 0.17.
On the corpus-wide percentile scale those become 0.43, 0.89 — a total move of +0.46.
That is why the strict rule cannot build this chain. Roughly 90 % of the corpus scores exactly zero on Sadness, and tied values all collapse onto one point (about 0.45). The clips that genuinely carry Sadness sit above 0.85. There is simply nothing in between to step onto, so the only move available is one jump far wider than the 0.25 per-step cap.
What this rule changes: Same total move, bigger allowed step. Keep the 0.25 end-to-end requirement on the normalised scale but lift the per-step cap so a chain may cross the gap in one hop.
What it costs: The chain is no longer a smooth ramp. It is allowed to jump the whole distance in a single step, which is exactly what the step cap existed to forbid. Use it for k=2 pairs, where there is only one step anyway, and be careful reading it as a gradual progression.
Same speaker? No similarity score is available here — the emolia clips in this chain are not covered by either speaker-embedding store. The chain therefore rests on the corpus's own speaker/track labelling, which is not the same as a measured check.
Voice consistency: these clips are separate recordings joined together, and no voice-similarity check could be run for this sample, so there is no measurement of how closely the voices match. You may hear the voice shift between segments. Voice conversion has not been applied yet in this build. A planned pass will re-render every segment onto the first segment's voice, which removes this effect entirely.
Unchanged across all 2 clips: an adult masculine voice · neutral-toned, slightly dark, balanced body, average recording, quiet background, measured, slightly relaxed, steady
(normally alert, didactic, monologue)Die zehnte Etappe der Tour de France 2022 startete in Morsin, in Haute-Savoyen.
full caption & clip details
An adult masculine voice; delivery is normally alert, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, fairly smooth, balanced body; somewhat unclear, some disfluency, fairly narrow pitch, normal breath; affect is neutral, neutral stance, slightly guarded; no dominant emotion; style: didactic, monologue; average recording, quiet background; genuineness 1.8/6; vocal-burst blend 0.1/10; 6.9s, DE.
DE_0068rhxjjg8_W000000 · in -20.8 dBFS · gain +0.8 dB · emolia-00215
(fatigue exhaustion, disappointment, bitterness·subdued, monologue, casual)Aber nach einigen Kilometern stand ein Polizist, stoppte uns, weil die Straße noch nicht offiziell freigegeben gewesen ist. So hatten wir nochmals circa eine halbe Stunde zu warten.
full caption & clip details
A middle-aged masculine voice; delivery is subdued, measured, slightly relaxed, steady; timbre is neutral-toned, slightly dark, slightly rough, balanced body; somewhat unclear, some disfluency, fairly narrow pitch, normal breath; affect is neutral, neutral stance, slightly guarded; reads as fatigue exhaustion, disappointment, bitterness; style: monologue, casual; average recording, quiet background; genuineness 2.8/6; vocal-burst blend 0.1/10; 12.5s, DE.
DE_0068rhxjjg8_W000012 · in -23.7 dBFS · gain +3.7 dB · emolia-00215