Voice Persona

Flash TTS (Expressive)
Ready. Adjust persona acoustic characteristics and trigger speech synthesis.

Spectral Oscilloscope & Model Telemetry

WebAudio Engine Online
REAL-TIME FFT 2048 · SPECTRAL ENERGY F0: 130 Hz | F1: 850 Hz
Time-to-First-Audio (TTFA)
68 ms
Fast streaming inference
Throughput / Speed
42.5x
Real-time factor (RTF: 0.02)
Audio Duration
6.4s
28 spoken words
Estimated Unit Cost
$0.00021
142 audio tokens
Documentary Narrator (Gemini 3.8 Flash)
Formant Warmth: 850Hz • Pitch Scale: 1.00x • Cadence: 1.00x

Voice Persona Profile & Acoustic Configuration

Export the complete model specification and synthetic test audio profile.

Architecture Breakdown: Gemini 3.8 Flash vs Flash-Lite TTS

Gemini 3.8 Flash TTS

Tuned for cinematic expression, rich emotional micro-inflections, and authentic accent contours. Ideal for audiobooks, gaming companions, narrative entertainment, and high-engagement conversational agents.

Gemini 3.8 Flash-Lite TTS

Architected for extreme efficiency and horizontal scaling. Offers ultra-low Time-to-First-Audio (<45ms) and halved compute footprint for high-volume customer service bots, interactive IVRs, and real-time live translations.

Parametric Formant Modeling

By shifting fundamental F0 frequencies and tuning formant envelope resonance filters, developers can reliably test persona archetypes directly in-browser prior to enterprise API deployment.

Enjoy this tool? Build your own with Super