Skip to content
View gwx314's full-sized avatar

Block or report gwx314

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 31 1 Updated Jun 17, 2026

MRSAudio: A Large-Scale Multimodal Recorded Spatial Audio Dataset with Refined Annotations

Python 43 2 Updated Oct 15, 2025

Scalable annotation pipeline for action-aglined fine-grained instruciton for Visual-language-Action model

Python 76 1 Updated Aug 3, 2026

Official repository for "Modeling and Benchmarking Spoken Dialogue Rewards with Modality and Colloquialness". An end-to-end multi-turn reward model for spoken dialogue systems.

Python 5 Updated Jun 17, 2026

Bridge local AI coding agents (Claude Code, Cursor, Gemini CLI, Codex) to messaging platforms (Feishu/Lark, DingTalk, Slack, Telegram, Discord, LINE, WeChat Work). Chat with your AI dev assistant f…

Go 2 Updated Aug 14, 2026

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice…

Python 12,949 1,680 Updated Mar 17, 2026

Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.

Jupyter Notebook 3,954 282 Updated Apr 23, 2026

PyTorch Implementation of TCSinger(EMNLP 2024): Zero-Shot Singing Voice Synthesis with Style Transfer and Multi-Level Style Control

Python 386 46 Updated Oct 7, 2025

PyTorch Implementation of StyleSinger(AAAI 2024): Style Transfer for Out-of-Domain Singing Voice Synthesis

Python 420 27 Updated Aug 15, 2025

PyTorch Implementation of TCSinger 2(ACL 2025): Customizable Multilingual Zero-shot Singing Voice Synthesis

Python 182 31 Updated Apr 19, 2026

Dataset and code of GTSinger(NeurIPS 2024 Spotlight): A Global Multi-Technique Singing Corpus with Realistic Music Scores for All Singing Tasks

Python 523 17 Updated Aug 15, 2025

pytorch implementation of openpose including Hand and Body Pose Estimation.

Jupyter Notebook 2,319 418 Updated Jul 9, 2024