Offline inference engine for art, real-time voice conversations, LLM powered chatbots and automated workflows
-
Updated
Sep 19, 2026 - Python
Offline inference engine for art, real-time voice conversations, LLM powered chatbots and automated workflows
一个基于Indextts和Qwen3TTS的 AI 有声书制作工具。利用 LLM 自动拆解剧本与识别情绪,集成多角色 TTS 语音合成(可智能分析音色并使用Qwen3TTS语音设计模型从音色描述文本生成音色),支持音效(SFX)、背景音乐(BGM)混音及实时台词音频滤波器的自动插入和匹配,可直接在浏览器导出 wav 成品,本工具本体无需配置环境即可跨平台在浏览器使用。现已支持背景图片提示词生成功能,可一键导出带情节背景图片和故事音频的mp4视频。
🎙️ Deepfake Audio – A neural voice cloning studio powered by SV2TTS technology.
AI Voice Cloning Desktop Application that runs locally on your computer and doesn't cost anything to run
Pushing Deep Learning models into production using torchserve, kubernetes and react web app 😄
🔊 A fully basic voice synthesizer in vanillaJS
Full Stack Text To Speech App With React Native, Nativewind, Nodejs & Express 🔥
A curated list of AI audio generation APIs, SDKs, and tools including text-to-speech, speech synthesis, music generation, voice cloning, sound design, and generative AI platforms. Covers commercial services, open source models with APIs, and production-ready infrastructure for developers building audio applications.
Our powerful online tool allows you to easily transform written text into high-quality, natural-sounding speech in multiple languages.
Native macOS text-to-speech with a local expressive Kokoro voice.
Our powerful online tool allows you to easily transform written text into high-quality, natural-sounding speech in multiple languages.
A Text to Speech App for Qwen3-TTS Family Models to create custom voices, voice cloning with minimal effort.
OpenAI Text-to-Speech Interface
Python program to convert text to speech.
Text To Speech Converter using HTML, CSS and JavaScript
A festive interactive birthday web card, vibe-coded in Google AI Studio with Gemini. Guides the recipient through making a wish (real mic candle-blow detection), unwrapping a gift box, and a personal letter with confetti, synthesized music, and browser text-to-speech. Personalizable via URL params, no backend, pure client-side React.
CHATGPT Text-to-Speech Application
Real-time Urdu-to-English voice interpreter for macOS meetings
Open-source voice-cloning and voice-design studio with OmniVoice accent audition, Stitch Studio, and an OpenAI-compatible TTS API.
C# monorepo containing an ONNX-based TTS inference server (Piper/OpenVoice V2) with OpenAI API compatibility and real-time audio DSP. Zero Python.
To associate your repository with the text-to-speech-app topic, visit your repo's landing page and select "manage topics."