Open Source Text-to-Speech Projects

Discover 31 open source Text To Speech repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. Text To Speech projects here are most often combined with Python, AI and Machine Learning. Last updated October 3, 2026.

31 repositories · updated October 3, 2026

BrowserAI: Run AI Models Directly in Your Browser

BrowserAI: Run AI Models Directly in Your Browser

BrowserAI is a TypeScript library for running language, speech, and audio models locally in a web browser. It suits developers building privacy-conscious AI features without server-side inference, provided users have a compatible WebGPU browser and hardware.

TypeScriptAILLM
Added Nov 21, 2025 View details
InfiniteTalk: Generate Audio-Driven Talking Videos

InfiniteTalk: Generate Audio-Driven Talking Videos

InfiniteTalk generates talking videos from audio and an image or existing video, synchronizing speech with facial expressions and body movement. It is aimed at creators and developers who need dubbed or audio-driven video, including long-form generation.

PythonAIGenerative AI
Added Nov 13, 2025 View details
podcastfy: Turn Multimodal Sources Into AI Podcasts

podcastfy: Turn Multimodal Sources Into AI Podcasts

Podcastfy is a Python package and CLI for turning websites, PDFs, images, YouTube videos, and topics into multilingual conversational audio. It suits developers and creators who want to customize or automate podcast generation with hosted or local language models.

PythonGenerative AIText To Speech
Added Nov 9, 2025 View details
captcha: Generate Image and Audio CAPTCHAs

captcha: Generate Image and Audio CAPTCHAs

captcha is a Python library for generating image and audio CAPTCHAs, with built-in voice and font data and support for custom assets. It suits applications that need to create their own CAPTCHA challenges and save them as image or audio files.

PythonLibrarySecurity
Added Nov 5, 2025 View details
xiaozhi-esp32-server: Run a Backend for ESP32 Voice Devices

xiaozhi-esp32-server: Run a Backend for ESP32 Voice Devices

A self-hosted backend for xiaozhi-esp32 devices, providing voice interaction, AI model integrations, device control, and an administration console. It suits ESP32 owners who want to deploy and configure their own service.

JavaScriptAISelf Hosted
Added Oct 28, 2025 View details
ChatTTS: Generate Expressive Speech for Dialogue

ChatTTS: Generate Expressive Speech for Dialogue

ChatTTS is a generative text-to-speech model built for dialogue, including assistant-style speech. It supports Chinese and English, multiple speakers, and controls for prosody such as pauses and laughter.

PythonAIMachine Learning
Added Oct 12, 2025 View details
chatterbox-vllm: Generate Speech with Chatterbox on vLLM

chatterbox-vllm: Generate Speech with Chatterbox on vLLM

A vLLM port of the Chatterbox text-to-speech model, built to improve GPU throughput and support batched generation. It suits developers with compatible Nvidia hardware who can work with an early, changing implementation.

PythonText To SpeechMachine Learning
Added Oct 11, 2025 View details
Previous Page 3 Next

Related topics

OS
OSRepos

Analysis and discovery of open source repositories. Find interesting projects and follow their updates.

Monitor your website with YourWebsiteScore

OSRepos shares public repositories for knowledge and discovery only. Any installation, execution, configuration, or use of third-party repository code is at your own risk. Always review source code, dependencies, licenses, and security implications before running anything.

© 2025 OSRepos. Built with Nuxt 3 and lots of ❤️