Skip to content
View kingfener's full-sized avatar
  • Beijing

Block or report kingfener

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Fully automatic censorship removal for language models

Python 27,275 2,962 Updated Aug 7, 2026

Code and data for the paper: IntentionQA: A Benchmark for Evaluating Purchase Intention Comprehension Abilities of Large Language Models in E-commerce (https://arxiv.org/pdf/2406.10173)

Python 12 1 Updated Apr 27, 2024

Real-Time VLAs via Future-state-aware Asynchronous Inference.

Python 477 34 Updated Apr 22, 2026

PeRFlow: Piecewise Rectified Flow as Universal Plug-and-Play Accelerator (NeurIPS 2024)

Jupyter Notebook 538 32 Updated Sep 8, 2025

Towards Human-Sounding Speech

Python 6,285 532 Updated Dec 5, 2025

Local UI to run and train LLMs and diffusion models, including Kimi K3, MiniMax-H3, Gemma 4, Qwen3.6, DeepSeek-V4, FLUX and more.

Python 69,890 6,310 Updated Aug 11, 2026

Agent framework and applications built upon Qwen>=3.0, featuring Function Calling, MCP, Code Interpreter, RAG, Chrome extension, etc.

Python 16,949 1,690 Updated Mar 4, 2026

A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone

Python 26,152 2,041 Updated Aug 4, 2026

Command-line interface and Python library to transcribe pinyin to IPA. The tones are attached to the vowel of the syllable.

Python 53 10 Updated Apr 16, 2025

Target Speaker Extraction Toolkit

Python 304 47 Updated Oct 4, 2025

zero-shot voice conversion & singing voice conversion, with real-time support

Python 3,892 527 Updated Apr 20, 2025

[CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. 接近GPT-4o表现的开源多模态对话模型

Python 10,122 789 Updated Sep 22, 2025

Genshin Datasets For SVC/SVS/TTS

738 43 Updated Jan 11, 2026

video editing with vim/spreadsheet/sed/python. methodology inspired by BBC digital paper edit. "Excel-dit"

Python 101 6 Updated Jul 27, 2026

Build and share delightful machine learning apps, all in Python. 🌟 Star to support our work!

Python 43,331 3,575 Updated Aug 11, 2026

A dataset of 222 digital musical scores aligned with 1068 performances (more than 92 hours) of Western classical piano music.

Jupyter Notebook 39 5 Updated Feb 12, 2026

Foundational Models for State-of-the-Art Speech and Text Translation

Jupyter Notebook 11,840 1,176 Updated Jul 28, 2026

Official implementation of "Separate Anything You Describe"

Python 1,924 155 Updated Nov 26, 2024

Facebook AI Research Sequence-to-Sequence Toolkit written in Python.

Python 32,244 6,675 Updated Sep 30, 2025

Book_7_《机器学习》 | 鸢尾花书:从加减乘除到机器学习;欢迎批评指正

Jupyter Notebook 3,335 624 Updated May 1, 2026

🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production

Python 45,878 6,155 Updated Aug 16, 2024

State-of-the-art deep learning based audio codec supporting both mono 24 kHz audio and stereo 48 kHz audio.

Python 4,013 362 Updated Jan 4, 2024

A book about Text-to-Speech (TTS) in Chinese.

TeX 612 83 Updated Apr 19, 2022

Towards hot directions in industrial end to end speech recognition

329 40 Updated Nov 30, 2021

List of speech synthesis papers.

1,074 121 Updated Jul 24, 2023

PyTorch Implementation of Google's Parallel Tacotron 2: A Non-Autoregressive Neural TTS Model with Differentiable Duration Modeling

Python 191 44 Updated Nov 18, 2021

Mellotron: a multispeaker voice synthesis model based on Tacotron 2 GST that can make a voice emote and sing without emotive or singing training data

Jupyter Notebook 870 186 Updated Jul 22, 2023

an open-source implementation of sequence-to-sequence based speech processing engine

C++ 967 197 Updated Dec 2, 2022

Crack LeetCode, not only how, but also why.

Markdown 135,291 23,584 Updated Feb 28, 2026

中文语音识别; Mandarin Automatic Speech Recognition;

Python 1,969 477 Updated Jul 25, 2024
Next