Voices with a chuckle.

Orpheus Q8 Tara is included below with six new recordings using the same emotional and speech passages, including native laughter, chuckle, and sigh cues. Its listening copies use the same −26 LUFS target as Azelma/Eponine.

Every saved voice/reference for Chatterbox Nano and Chatterbox Turbo in this archive: Azelma, Eponine, Alba, and Tara-reference. Each has a conversation clip and a vocal-events prompt containing [chuckle] and [sigh]. Listen to judge the actual laughter, voice quality, and consistency; no quality ranking is implied.

The Azelma and Eponine references lead because publisher metadata identifies them as female. Tara-reference is conditioned on a short synthetic Orpheus Tara clip; it is not Orpheus running inside Chatterbox. These are all saved voices here, not an exhaustive list of voices either model can condition on.

The Chatterbox timings below were measured on orbital, MacBook Air M1, 8 GB, CPU-only. They are full waveform generation times, excluding model download/load and voice conditioning. They do not measure streaming or network latency. RTF = generation seconds / audio seconds; below 1 means faster than real time. Batch/offline listening is supported regardless of RTF.

Listening copies retain their original gain-only treatment: Azelma/Eponine approximately −26 LUFS; Alba/Tara-reference approximately −23 LUFS (some peak-constrained clips are quieter). These two audition sets are not level-matched against each other. Original audio is linked for every clip.

Orpheus Q8 · Tara · emotion and laughter

Six new recordings using the same passages as the other demos: expression, laughter, chuckle and sigh, ordinary speech, conversation, and technical pronunciation. Orpheus uses native <laugh>, <chuckle>, and <sigh> cues; the remaining text is unchanged.

Listening copies use static gain toward −26 LUFS, matching the Azelma/Eponine samples. Older Alba/Tara-reference Chatterbox clips use −23 LUFS. No pitch or speed changes; original WAVs are linked below.

Expression · surprise and delight

12.03 s audio · 23.31 s full request

Passage and original WAV

Wait, you actually got it working? Oh, that's brilliant! I was starting to think this little machine had given up on us. All right. Take a breath. We have time to get this right.

Original WAV

Laughter · laugh and chuckle

7.08 s audio · 14.46 s full request

Passage and original WAV

Hey Aku. <laugh> The laptop found its voice. <chuckle> That little victory made my day.

Original WAV

Vocal events · chuckle and sigh

10.75 s audio · 21.11 s full request

Passage and original WAV

Well, that was subtle. <chuckle> The desktop survived, and nobody had to buy a new graphics card. <sigh> I could get used to this.

Original WAV

Ordinary speech

5.97 s audio · 11.88 s full request

Passage and original WAV

Hey Aku. The laptop found its voice. That little victory made my day.

Original WAV

Conversation

14.85 s audio · 28.36 s full request

Passage and original WAV

Hey Aku. I found a few voices worth listening to. Take your time; I care more about whether this sounds natural than whether it wins a benchmark. Does the rhythm feel right, or does it sound like someone reading a script?

Original WAV

Technical pronunciation

14.59 s audio · 28.62 s full request

Passage and original WAV

The backup finished at 9:42 p.m. We copied 3.7 gigabytes, verified the SHA-256 checksum, and left the GPU alone. The address is 192.168.1.3. Please don't reboot the machine yet.

Original WAV

Recording and timing details

Generated September 6 on eimu through the existing Orpheus-3b-FT-Q8_0 Tara service. Full request time includes queueing, model generation, SNAC decoding, server WAV writing and HTTP transfer. It is not isolated generation time or measured first-audio latency. The service's existing settings and GPU pacing were retained.

Prompts, timing records, levels and audio checksums

Earlier Tabi Q8/Q4 benchmark recordings

Orpheus · Tara

Original Orpheus-3b-FT speech from Tabi, with Q8 and Q4 recordings. These are saved benchmark clips from September 5 PDT / September 6 UTC, using an older passage without laughter tags. They are separate from the Chatterbox auditions.

The original WAVs have no level matching. Their request timings include streaming SNAC decoding and exclude initialization. The Q4 runs use an improved runtime; these recordings do not isolate quantization or provide a matched timing comparison with the other models.

Spoken passage and recording details

Hey Aku, the MacBook is handling my voice now. Your gaming GPU can take the night off.

This is the saved benchmark prompt; the recordings were generated on Tabi. Orpheus used Intel Vulkan for the model and CPU SNAC for audio decoding. The MacBook sentence describes the old fixture, not current voice routing.

Audio provenance and checksums

Orpheus Q8 · Tara

Original runtime · CPU SNAC, one thread.

Run 1

27.77 s request · 6.83 s audio · RTF 4.068 · 1.71 s first audio

Original WAV

Run 2

42.54 s request · 6.83 s audio · RTF 6.232 · 1.98 s first audio

Original WAV · Saved measurements

Chatterbox Nano

Chatterbox Nano · Azelma

Conversation

9.12 s generation · 12.44 s audio · RTF 0.733

Prompt and files

Hey Aku. I found a few voices worth listening to. Take your time; I care more about whether this sounds natural than whether it wins a benchmark. Does the rhythm feel right, or does it sound like someone reading a script?

Original WAV · Listening WAV

Chuckle and sigh

6.43 s generation · 8.60 s audio · RTF 0.748

Prompt and files

Well, that was subtle. [chuckle] The desktop survived, and nobody had to buy a new graphics card. [sigh] I could get used to this.

Original WAV · Listening WAV
More saved clips

Expression · Original WAV

Technical · Original WAV

Chatterbox Nano · Eponine

Conversation

10.60 s generation · 14.12 s audio · RTF 0.751

Prompt and files

Hey Aku. I found a few voices worth listening to. Take your time; I care more about whether this sounds natural than whether it wins a benchmark. Does the rhythm feel right, or does it sound like someone reading a script?

Original WAV · Listening WAV

Chuckle and sigh

6.50 s generation · 8.52 s audio · RTF 0.762

Prompt and files

Well, that was subtle. [chuckle] The desktop survived, and nobody had to buy a new graphics card. [sigh] I could get used to this.

Original WAV · Listening WAV
More saved clips

Expression · Original WAV

Technical · Original WAV

Chatterbox Nano · Alba

Conversation

9.61 s generation · 13.18 s audio · RTF 0.729

Prompt and files

Hey Aku. I found a few voices worth listening to. Take your time; I care more about whether this sounds natural than whether it wins a benchmark. Does the rhythm feel right, or does it sound like someone reading a script?

Original WAV · Listening WAV

Chuckle and sigh

6.83 s generation · 8.46 s audio · RTF 0.807

Prompt and files

Well, that was subtle. [chuckle] The desktop survived, and nobody had to buy a new graphics card. [sigh] I could get used to this.

Original WAV · Listening WAV
More saved clips

Expression · Original WAV

Technical · Original WAV

Chatterbox Nano · Tara-reference

Conversation

8.71 s generation · 11.74 s audio · RTF 0.742

Prompt and files

Hey Aku. I found a few voices worth listening to. Take your time; I care more about whether this sounds natural than whether it wins a benchmark. Does the rhythm feel right, or does it sound like someone reading a script?

Original WAV · Listening WAV

Chuckle and sigh

6.11 s generation · 8.74 s audio · RTF 0.699

Prompt and files

Well, that was subtle. [chuckle] The desktop survived, and nobody had to buy a new graphics card. [sigh] I could get used to this.

Original WAV · Listening WAV
More saved clips

Expression · Original WAV

Technical · Original WAV

Chatterbox Turbo

Chatterbox Turbo · Azelma

Conversation

23.05 s generation · 13.80 s audio · RTF 1.670

Prompt and files

Hey Aku. I found a few voices worth listening to. Take your time; I care more about whether this sounds natural than whether it wins a benchmark. Does the rhythm feel right, or does it sound like someone reading a script?

Original WAV · Listening WAV

Chuckle and sigh

15.98 s generation · 9.44 s audio · RTF 1.692

Prompt and files

Well, that was subtle. [chuckle] The desktop survived, and nobody had to buy a new graphics card. [sigh] I could get used to this.

Original WAV · Listening WAV
More saved clips

Expression · Original WAV

Technical · Original WAV

Chatterbox Turbo · Eponine

Conversation

26.20 s generation · 14.00 s audio · RTF 1.872

Prompt and files

Hey Aku. I found a few voices worth listening to. Take your time; I care more about whether this sounds natural than whether it wins a benchmark. Does the rhythm feel right, or does it sound like someone reading a script?

Original WAV · Listening WAV

Chuckle and sigh

14.76 s generation · 8.52 s audio · RTF 1.733

Prompt and files

Well, that was subtle. [chuckle] The desktop survived, and nobody had to buy a new graphics card. [sigh] I could get used to this.

Original WAV · Listening WAV
More saved clips

Expression · Original WAV

Technical · Original WAV

Chatterbox Turbo · Alba

Conversation

19.30 s generation · 11.70 s audio · RTF 1.649

Prompt and files

Hey Aku. I found a few voices worth listening to. Take your time; I care more about whether this sounds natural than whether it wins a benchmark. Does the rhythm feel right, or does it sound like someone reading a script?

Original WAV · Listening WAV

Chuckle and sigh

14.80 s generation · 9.30 s audio · RTF 1.592

Prompt and files

Well, that was subtle. [chuckle] The desktop survived, and nobody had to buy a new graphics card. [sigh] I could get used to this.

Original WAV · Listening WAV
More saved clips

Expression · Original WAV

Technical · Original WAV

Chatterbox Turbo · Tara-reference

Conversation

19.01 s generation · 12.46 s audio · RTF 1.526

Prompt and files

Hey Aku. I found a few voices worth listening to. Take your time; I care more about whether this sounds natural than whether it wins a benchmark. Does the rhythm feel right, or does it sound like someone reading a script?

Original WAV · Listening WAV

Chuckle and sigh

12.70 s generation · 7.38 s audio · RTF 1.721

Prompt and files

Well, that was subtle. [chuckle] The desktop survived, and nobody had to buy a new graphics card. [sigh] I could get used to this.

Original WAV · Listening WAV
More saved clips

Expression · Original WAV

Technical · Original WAV