Skip to content
View ivcylc's full-sized avatar
🏆
🏆

Block or report ivcylc

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

AI Image Detector - Check if the Image is AI

Python 303 7 Updated Jul 11, 2026

Vidu S1: A Real-Time Interactive Video Generation Model

226 6 Updated Jul 13, 2026

WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling

Python 211 6 Updated Jun 6, 2026

Open-source humanize text toolkit. Documents 4 humanization methodologies with reference implementations, plus a production pipeline combining LLM rewriting with a cross-engine translation chain. P…

Python 1,539 90 Updated Jul 27, 2026

MultiModal Audio Generation in Raw Waveform Space.

Python 154 10 Updated May 26, 2026

YAML-native agent workflow execution engine, written in Rust

Rust 1,225 9 Updated Apr 28, 2026

Official Cortex development plugin with native coding tools, workflow skills, project analysis, and git-aware automation.

Rust 96 Updated May 19, 2026

Enterprise-ready Spring AI platform for RAG, tool calling, async ingestion, JWT/RBAC security, and observability.

Java 183 13 Updated Jul 24, 2026

Omni2Sound — Your Multimodal Audio Generation Codebase (CVPR 2026 Highlight)

Python 143 3 Updated Apr 25, 2026

Being-H is BeingBeyond's family of human-centric embodied foundation models.

Python 1,111 59 Updated Jun 16, 2026

[RSS26'] Welcome to Psi-Zero, a Humanoid VLA towards Universal Humanoid Intelligence.

Python 2,740 83 Updated Jul 17, 2026

MiMo-Audio: Audio Language Models are Few-Shot Learners

Python 1,070 105 Updated Jun 17, 2026

The repository provides code for running inference with the Meta Segment Anything Audio Model (SAM-Audio), links for downloading the trained model checkpoints, and example notebooks that show how t…

Python 3,586 324 Updated May 26, 2026

Official repository of Myna: Masking-Based Contrastive Learning of Musical Representations

Python 17 2 Updated Mar 31, 2025

A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot text-to-speech

Python 956 72 Updated Apr 9, 2026
3 Updated Oct 26, 2025
HTML 1 Updated Aug 19, 2025

Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation

Python 452 30 Updated Nov 27, 2025
Python 1,744 202 Updated Nov 15, 2025

Vector (and Scalar) Quantization, in Pytorch

Python 3,992 336 Updated Jul 20, 2026

Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

Python 22,486 2,594 Updated May 25, 2026

[ICLR 2025] Official PyTorch implementation of our paper for general continual learning "Advancing Prompt-Based Methods for Replay-Independent General Continual Learning".

Python 18 Updated Dec 21, 2025

Official PyTorch implementation of our CVPR 2025 paper, "LoRA Subtraction for Drift-Resistant Space in Exemplar-Free Continual Learning."

Python 18 5 Updated Mar 28, 2025

Awesome Incremental Learning

4,502 628 Updated Jun 27, 2026

my homepage

JavaScript 1 Updated Jul 28, 2026

[ICLR 2025] Enhancing Self-Supervised Models with Audio Mixtures for Polyphonic Soundscapes

Python 79 5 Updated Oct 8, 2025

A library built for easier audio self-supervised training, downstream tasks evaluation

Python 140 11 Updated Sep 25, 2025

A Comprehensive Survey on Continual Learning in Generative Models.

163 10 Updated Jun 1, 2026

This is the official repository of the papers "Parameter-Efficient Transfer Learning of Audio Spectrogram Transformers" [IEEE MLSP 2024] and "Efficient Fine-tuning of Audio Spectrogram Transformers…

Python 41 4 Updated Jul 31, 2024

MutiModel paper reading (Visual, Audio)

22 1 Updated Nov 24, 2025
Next