6 GB NVIDIA memory
FP32 · Chatterbox runtime
The official implementation uses CUDA when available; this is a practical provider baseline with room for synthesis.
Resemble AI
Audio model
An expressive English speech model with zero-shot voice cloning from Resemble AI.
Capabilities
The most useful reasons to choose this model, without making you read through its repository first.
Voice cloning
Narration
Expressive assistants
For providers
The provider requirements live here, separate from the information you need to choose and use the model.
6 GB NVIDIA memory
FP32 · Chatterbox runtime
The official implementation uses CUDA when available; this is a practical provider baseline with room for synthesis.
Source and license
OpenMayhem links back to the source repository so you can inspect the model card, files, license, limitations, and creator guidance directly.