The Tragedy Academy

Source-voice

← The Corpus·The Source Voice

Tragedy Academy Studios

A synthetic voice cannot exceed its source.

There is a voice model of Jay Hicks in the Respeecher marketplace. These seven films are the instrument it was made from — seven genres, six characters, one performer, and the full breadth of what is actually in there.

 

Hear the synthetic model →

 


Why this page exists

The ceiling of a voice model is the range of the person it came from.

A model inherits whatever behavioral range exists in its source. How far a character can be pushed, whether a register holds under strain, what happens at the extremes — all of it is set before a single line is ever synthesised.

This is that source at full extension. Six characters, six genres, one performer. Start at the top — the first one is the furthest thing from a podcast voice on this page.

Each film is a single performance, not a compilation. No line-by-line assembly, no comping between takes.

 


01 — Supper Time

Elias Rook · 0:53 · Survival horror · Photoreal

A cattle farmer holds a family supper that has gone very wrong. Constricted throat placement with heavy vocal fry, authentic southwest regional vowel shifts, harsh glottal attacks, ragged unfiltered inhalations driving the tempo.

He is at peak volume for forty-seven straight seconds. Then the last word drops twenty-five decibels to a whisper. The furthest point in the reel from the commercial voice.

02 — Fish & Chip

Chip · 1:06 · Deadpan comedy · 2D animation

A french fry tows a goldfish around a county fair planning a heist. High nasal mix with tight laryngeal constriction, a twelve-decibel dynamic window, breath completely suppressed, diction overly formal on purpose.

Never raises his voice once. The hardest register to sustain, because nothing is allowed to help you.

03 — Unpunished!

“Broken Leg” · 1:17 · Theatrical tragedy · Dark fantasy / narrative

A ruined performer in a collapsed opera house, named for “break a leg.” Roughly thirty-five decibels between murmur and scream. Deep wounded chest pushing into high pharyngeal strain, theatrical gasps as punctuation, articulation holding crisp under distortion.

At 0:54 the theatrical grandeur drops away and the hatred becomes personal. He tears the mask off at the end and bows bare-faced to an empty house.

04 — Three Minutes

Dr. Anselm Vosk · 0:39 · Manic · Steampunk / fantasy

An inventor in a lantern-lit workshop, three minutes from something. Extended upper pitch ceiling with head-voice glides, octave jumps on emotional peaks, explosive laughter, articulation staying hyper-precise through extreme tempo shifts.

At 0:22 the explanation breaks into uncontrolled laughter mid-thought. That transition is the whole performance.

05 — Divided & Dying

0:25 · Epic narration · Fantasy / trailer

An unseen narrator over a fallen empire in the rain. Low chest and sub-glottal resonance, compressed dynamics, elongated vowels, deep controlled inhalations carrying the weight.

No character on screen at all. The voice has to build the world by itself.

06 — Onboarding

Professor Gantt Flotsam · 0:28 · Warm institutional · Animated family / education

A professor welcomes a corridor of schoolchildren using the language of a corporate onboarding deck. Mid-high placement with a brightened smile vector, narrow dynamic range, hidden breath, pitch lifts at phrase ends.

The comedy is entirely in the mismatch — growth mindset and agility, delivered to eight-year-olds with total sincerity.

 


The zero point

And here is the same voice being ordinary.

Range only means something against a reference. This is the professional instrument doing paid, unremarkable, competent work — five brands, five stances, no character. Deep chest for premium automotive, warm mid-chest for financial advisory, light nasal conversational for consumer, gritty chest-throat for apparel, softened attack for healthcare.

Commercial reel · 1:18 · Everything above is measured as distance travelled from this.

 


What holds it together

One instrument, six configurations.

Independent acoustic review of the full set identified the through-line under every character: over-articulated consonant termination — crisp stops, hard T and K — and rhythmically structured breath intakes placed before emotional pivots rather than after them.

That signature survives vocal fry, head voice, nasal constriction and full theatrical strain. It is what makes the range legible as range rather than as six different actors.

The same voice has been licensed as a synthetic model built with Respeecher's Emmy-awarded team, trained on a curated corpus of unscripted long-form conversation.

The corpus contained no animation scripts, no healthcare copy and no game dialogue. Prompted into those verticals cold, the model produced them in a single pass — including organic self-correction and hesitation in positions where a person would actually stall.

 


The range is the asset.

Every character on this page was written, performed, directed and finished in-house — six worlds built to find the edges of one instrument. What the model can reach was decided here.

The synthetic voice is licensed through Respeecher. The corpus it was trained on is licensed directly.

Authenticity Amplified — the corpus →

Hear the synthetic model →

 

Tragedy Academy Studios · The Source Voice

← The Corpus·The Source Voice