Stars
Cotton Music(棉花音乐): Player for Local file, Drive, Navidrome, Subsonic, Emby, Jellyfin, Plex. Supports Android, iOS, and PC
A feature-rich command-line audio/video downloader
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Official inference framework for 1-bit LLMs
Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
Optimized OpenAI's Whisper TFLite Port for Efficient Offline Inference on Edge Devices
Offline Speech Recognition with OpenAI Whisper and TensorFlow Lite for Android
ml-research / diffusers
Forked from huggingface/diffusers🤗 Diffusers: State-of-the-art diffusion models for image and audio generation in PyTorch
Bark Voice Cloning and Voice Cloning for Chinese Speech
Chinese version of GPT2 training code, using BERT tokenizer.
Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and…
StableLM: Stability AI Language Models
🔊 Text-Prompted Generative Audio Model
[CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation
CSDN自动展开全文chrome插件,安装后不用每次都要自己去手动点击阅读全文,不登录也可查看全文!
Segment Anything for Stable Diffusion WebUI
pytorch handbook是一本开源的书籍,目标是帮助那些希望和使用PyTorch进行深度学习开发和研究的朋友快速入门,其中包含的Pytorch教程全部通过测试保证可以成功运行
GFPGAN aims at developing Practical Algorithms for Real-world Face Restoration.
a fork that installs runs on pytorch cpu-only
Stable Diffusion cpuonly webui 2.0
AUTOMATIC1111 Stable Diffusion WebUI 1.5 + Kohya's Scripts
Samples and Unpacker of malicious backdoors and exploits developed and used by Pinduoduo
Stable Diffusion web UI
China province/city/country geoJSON data
LSPass: Bypass restrictions on non-SDK interfaces