Skip to content
View ilya16's full-sized avatar

Block or report ilya16

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

PianoKontext: Expressive Performance Rendering from Deadpan Context

Python 5 Updated Jun 13, 2026

Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.

Python 10,765 999 Updated May 16, 2026

SOTA Open Source TTS

Python 31,754 2,734 Updated Jul 26, 2026
Python 942 76 Updated Jun 26, 2026

Elucidated Text-To-Audio (ETTA) is a SOTA text-to-audio model with a holistic understanding of the design space and trained with synthetic captions.

Python 137 12 Updated Mar 3, 2026

Official repository for Aria-MIDI: a MIDI dataset of 1,186,253 transcribed solo-piano recordings.

98 4 Updated Jun 19, 2025

SyMuRBench: Benchmark for symbolic music representations

Python 19 1 Updated Nov 6, 2025
Python 270 14 Updated Feb 14, 2024
Python 169 9 Updated Nov 22, 2024

UniAudio 2.0: An audio fundation model for text, speech, sound, and music

Python 212 9 Updated Feb 14, 2026

The open source code of ALMTokenizer2: Towards Low bit-rate and Semantic-rich Audio Tokenizer with Flow-based Scalar Diffusion Transformer Decoder

Python 45 Updated Sep 5, 2025

Official code and pretrained models for Linear Consistency Autoencoders (Lin-CAE), a method to induce linearity in audio autoencoders via data augmentation.

Python 17 Updated Feb 12, 2026

[ICLR 2026] SmartDJ: declarative audio editing with audio langugae model.

Python 68 2 Updated Apr 25, 2026

Bridging Piano Transcription and Rendering via Disentangled Score Content and Style (ICLR 2026 accepted)

10 Updated Feb 8, 2026

The most powerful local music generation model that outperforms almost all commercial alternatives, supporting Mac, AMD, Intel, and CUDA devices.

Python 11,914 1,483 Updated Jul 25, 2026

C++ Library for tokenizing MIDI files, designed to be compatible with the MIDITok python library

C++ 51 3 Updated Jun 8, 2026

Step-Audio 2 is an end-to-end multi-modal large language model designed for industry-strength audio understanding and speech conversation.

Python 1,489 111 Updated Mar 16, 2026

Suno-like music generation studio for HeartMuLa/heartlib - AI-powered music creation with reference audio style transfer

TypeScript 617 101 Updated Feb 25, 2026

PersonaPlex code.

Python 10,277 1,429 Updated Mar 2, 2026

A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot text-to-speech

Python 956 72 Updated Apr 9, 2026

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice…

Python 12,687 1,642 Updated Mar 17, 2026

HeartMuLa Official Repo: The Most Powerful Open-Source Music Generation Model of 2026

Python 3,807 435 Updated Apr 10, 2026

This is the official implementation for the paper "Pianist Transformer: Towards Expressive Piano Performance Rendering via Scalable Self-Supervised Pre-Training".

Python 48 5 Updated Jun 25, 2026

Open-Source Frontier Voice AI

Python 51,606 5,725 Updated Jul 24, 2026

Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

Python 22,501 2,594 Updated May 25, 2026

An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System

Python 22,284 2,717 Updated Jul 14, 2026

Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audi…

Python 9,975 830 Updated Mar 25, 2026
Jupyter Notebook 93 4 Updated Feb 6, 2026
Next