Generate audiobooks from e-books, voice cloning & 1158+ languages!
-
Updated
Sep 25, 2026 - Python
Generate audiobooks from e-books, voice cloning & 1158+ languages!
Webui for using XTTS and for finetuning it
Turn PDFs and EPUBs into audiobooks; subtitles or videos into dubbed videos (including translation), and more. For free. Pandrator uses local models, including voice-cloning (instant, RVC-enhanced, XTTS fine-tuning) and LLM processing. It aspires to be a user-friendly app with a GUI, an installer and all-in-one packages.
A simple FastAPI Server to run XTTSv2
End-to-end platform for building voice first multimodal agents
Local-first, open-source YouTube dubbing: clones the original voice into another language, time-synced. Chatterbox voice cloning + faster-whisper + NLLB, with multi-voice, burned subtitles and optional Wav2Lip lip-sync. Runs free on your own machine via a simple CLI.
The world’s first game framework that lets you talk to AI in real time — locally and for free. Supports any custom voice.
Local-first CLI that turns Markdown scripts into multi-speaker podcast-style audio using Coqui XTTS v2.
Portable offline AI audio studio with web UI & local API – XTTS, Fish Speech, Kokoro, Stable Audio, ACE-Step, voice cloning, music gen (no install)
A local, multi-voice audiobook generator: an LLM casts every character in its own voice, kept consistent across a whole series — with voice design, per-line emotion, seven languages, and M4B / Audiobookshelf export. Runs on your own machine; you own the files. Kokoro · Qwen3-TTS.
This is an interface that will offline convert anything pdf document you give it into an interview between two people discussing it.
Generate audiobooks from e-books, voice cloning & 1158 languages!
Magic Chat 🎙️ is a #real-time AI voice chat app with expressive characters using #OpenAI, #ElevenLabs, XTTS, #Ollama, #Kokoro TTS, #WebRTC, and #Docker, supporting #games, #stories, and #local or #cloud models.
OhanashiGPT is an application that generates personalized children's stories based on parameters like age and preferences. It narrates these stories using an AI-generated voice that mimics a parent, trained on their audio samples. The app also creates illustrations to accompany each story, providing a unique and engaging experience for children.
📞 Локальный AI-секретарь, тех. поддержка и менеджер по продажам с клонированием голоса XTTS v2, real-time распознаванием речи (Vosk/Whisper) и offline LLM (vLLM + Qwen/Llama и тп). Полноценная админ-панель (Vue 3), Telegram-бот, виджет для сайта, fine-tuning pipeline. Self-hosted, приватность данных, СМС и телефонные звонки .
Local CPU-first voice cloning & dubbing runtime with a PySide6 desktop GUI, viXTTS/XTTS-v2 engines, safe voice profiles and isolated ML runtimes.
A highly customizable AI companion for Telegram. Create digital clones with unique personalities, voice, and vision directly from Google Colab.
Portable multi-GPU text-to-speech server for Windows — 10 AI models, gateway + worker architecture, 7-stage audio pipeline, Whisper verification, one-click install (no system Python, no Docker, no admin rights)
A python-based tool for using CoquiTTS to turn documents into Audiobooks.
To associate your repository with the xtts topic, visit your repo's landing page and select "manage topics."