Skip to content
View GuodongQi's full-sized avatar
🎯
Focusing
🎯
Focusing
  • ZheJiang University

Block or report GuodongQi

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A powerful proxy management tool, built on top of Xray-core, with a focus on simplicity and ease of use.

TypeScript 4,807 278 Updated Aug 6, 2026

High-performance panel based on V2board secondary development supporting new protocols and new features

PHP 4,626 1,311 Updated Aug 7, 2026

Xray panel supporting multi-protocol multi-user expire day & traffic & IP limit (Vmess, Vless, Trojan, ShadowSocks, Wireguard, Hysteria, Tunnel, Mixed, HTTP, Tun, MTProto)

Go 44,511 8,590 Updated Aug 6, 2026

Official code for "F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization"

Python 170 18 Updated Mar 3, 2026

Kimi-Audio, an open-source audio foundation model excelling in audio understanding, generation, and conversation

Python 4,716 373 Updated Jun 21, 2025

本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)

HTML 24,879 2,840 Updated Jul 19, 2026

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice…

Python 12,874 1,663 Updated Mar 17, 2026

MOSS-Audio-Tokenizer is a Causal Transformer-based audio tokenizer built on the CAT architecture. Trained on 3M hours of diverse audio, it supports streaming and variable bitrates, delivering SOTA …

Python 249 18 Updated Jun 16, 2026

Train the next generation of TTS systems.

Python 169 17 Updated Sep 13, 2024

Fast and memory-efficient exact attention

Python 24,656 2,972 Updated Aug 9, 2026

A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot text-to-speech

Python 961 75 Updated Apr 9, 2026

Open-Source Frontier Voice AI

Python 52,254 5,867 Updated Jul 24, 2026

The best ChatGPT that $100 can buy.

Python 57,067 7,908 Updated Aug 2, 2026

1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

Python 60,608 6,599 Updated Jul 22, 2026

pyright fork with various type checking improvements, improved vscode support and pylance features built into the language server

TypeScript 3,522 134 Updated Aug 8, 2026

12306接口抢票

Python 42 15 Updated Sep 19, 2025

Edit, preview and share mermaid charts/diagrams. New implementation of the live editor.

TypeScript 6,729 1,149 Updated Aug 8, 2026

An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System

Python 22,480 2,736 Updated Aug 4, 2026

Awesome speech/audio LLMs, representation learning, and codec models

1,243 74 Updated Jul 10, 2026

The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.

Python 948 49 Updated Aug 3, 2026

Your one-stop solution for voice dataset creation

Python 131 24 Updated Dec 10, 2023

Towards Human-Sounding Speech

Python 6,283 532 Updated Dec 5, 2025

SOTA Open Source TTS

Python 32,107 2,755 Updated Aug 3, 2026

Reverse Engineering of Supervised Semantic Speech Tokenizer (S3Tokenizer) proposed in CosyVoice

Python 525 69 Updated Dec 22, 2025

Text Normalization & Inverse Text Normalization

Python 807 115 Updated Jul 29, 2026

基于SparkTTS、OrpheusTTS等模型,提供高质量中文语音合成与声音克隆服务。

Python 609 76 Updated May 18, 2025

[ACL 2024] Official PyTorch code for extracting features and training downstream models with emotion2vec: Self-Supervised Pre-Training for Speech Emotion Representation

Python 1,173 89 Updated Dec 23, 2024

Official code for "EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting"

Python 123 13 Updated Oct 16, 2025

How to use our public wav2vec2 dimensional emotion model

Jupyter Notebook 556 50 Updated May 22, 2023
Next