Kokoro-82M
by HexgradThe two-paragraph Gettysburg excerpt, spoken with Kokoro-82M on Apple Silicon (MLX) or PyTorch (~23s).
Kokoro, Qwen, F5, Higgs, and Chatterbox are the supported speech backends. Each model narrates the identical two-paragraph Gettysburg excerpt so you can directly compare audio timbre, pacing, pronunciation, and latency.
Compare five open model architectures running on Apple Silicon (MLX) and AWS Batch:
The two-paragraph Gettysburg excerpt, spoken with Kokoro-82M on Apple Silicon (MLX) or PyTorch (~23s).
The same Gettysburg excerpt, spoken with Qwen3-TTS 0.6B CustomVoice with the Ryan preset (~48s).
The same Gettysburg excerpt, spoken with non-autoregressive F5-TTS via MLX on Apple Silicon.
The same Gettysburg excerpt, spoken with Boson AI's expressive 4B Higgs Audio v3 via MLX on Apple Silicon.
The same Gettysburg excerpt, spoken with Resemble AI's hybrid LLaMA-520M and Matcha-TTS flow matching.
The same Gettysburg excerpt, spoken with Fish Audio's dual-autoregressive model with natural prosody and breathing.
Customizing player appearance and article content extraction: