LocalAI - Models

vibevoice

Links

https://github.com/microsoft/VibeVoice

Tags

kokoro

Kokoro is an open-weight TTS model with 82 million parametrs. Despite its lightweight architecture, it delivers comparable quality to larger models while being significantly faster and more cost-efficient. With Apache-licensed weights, Kokoro can be deployed anywhere from production environments to personal projects.

Links

https://github.com/hexgrad/kokoro

Tags

kitten-tts

Kitten TTS is an open-source realistic text-to-speech model with just 15 million parameters, designed for lightweight deployment and high-quality voice synthesis.

Links

https://github.com/KittenML/KittenTTS

Tags

chatterbox

Chatterbox, Resemble AI's first production-grade open source TTS model. Licensed under MIT, Chatterbox has been benchmarked against leading closed-source systems like ElevenLabs, and is consistently preferred in side-by-side evaluations.

Links

https://github.com/resemble-ai/chatterbox

Tags

parler-tts-mini-v0.1

Parler-TTS is a lightweight text-to-speech (TTS) model that can generate high-quality, natural sounding speech in the style of a given speaker (gender, pitch, speaking style, etc). It is a reproduction of work from the paper Natural language guidance of high-fidelity text-to-speech with synthetic annotations by Dan Lyth and Simon King, from Stability AI and Edinburgh University respectively.

Links

https://github.com/huggingface/parler-tts

Tags

Model Gallery

Filter by type:

Filter by tags:

vibevoice

kokoro

kitten-tts

chatterbox

parler-tts-mini-v0.1