Skip to content
View proshian's full-sized avatar
  • ITMO University

Highlights

  • Pro

Block or report proshian

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

C 37,541 4,065 Updated Sep 24, 2026

U-MusT: A Unified Framework for Cross-modal Translation of Score Images, Symbolic Music, and Performance Audio

Python 13 1 Updated Sep 9, 2026

Create acoustic diffusers with custom images!

C++ 27 2 Updated Sep 8, 2026

An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python depend…

C++ 3,010 341 Updated Sep 25, 2026

YingMusic-Singer-Plus: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance

Python 113 14 Updated Apr 12, 2026
Python 9,192 676 Updated Aug 15, 2026

Run the full 2.78-trillion-parameter Kimi K3 model, DeepSeek V4.1 Flash or GLM-5.3-Flash beyond available RAM by streaming activated weights directly from NVMe. A dependency-free, embeddable C infe…

C 2,445 186 Updated Sep 18, 2026

SyMuPe: Affective and Controllable Symbolic Music Performance (ACM MM '25, Outstanding Paper Award)

Python 27 1 Updated May 8, 2026

Convert any Repo into an RL Environment

Python 672 99 Updated Sep 24, 2026

A multi-instrument music transcription model developed by Kyutai and Mirelo.

Python 1,502 176 Updated Sep 4, 2026

RapidIn: Scalable Influence Estimation for Large Language Models (LLMs). The implementation for paper "Token-wise Influential Training Data Retrieval for Large Language Models" (Accepted on ACL 2024).

Python 22 5 Updated Mar 10, 2026
Python 3,847 777 Updated Sep 12, 2026

TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting.

Python 33,648 3,248 Updated Sep 15, 2026

Towards a general language-audio model for computational paralinguistic tasks

Python 31 1 Updated Dec 14, 2024

DEMON: Diffusion Engine for Musical Orchestrated Noise

Python 299 30 Updated Sep 24, 2026
Python 6,099 475 Updated Jun 15, 2026

Codec for paper: LLaSA: Scaling Train-time and Inference-time Compute for LLaMA-based Speech Synthesis

Python 367 55 Updated Jun 25, 2026

Official Release of CMI-RewardBench: Evaluating Music Reward Models with Compositional Multimodal Instruction

Python 31 1 Updated May 21, 2026

Move files and folders to the trash

JavaScript 1,415 39 Updated Sep 18, 2026

SA3 medium audio inpainter — MLX SAME-L decoder + FastAPI + vanilla Svelte UI

Python 23 4 Updated May 26, 2026
Python 430 48 Updated Jul 23, 2026

An open-source model for music captioning, lyrics transcription, structural analysis, and musical question answering

Python 180 11 Updated May 9, 2026

an architecture for neural network inference in real-time audio applications

C++ 232 11 Updated Sep 25, 2026

Implementation of Multiscreen proposed by Ken Nakanishi for "Screening is Enough"

Python 19 2 Updated May 13, 2026

Implementation of the Hierarchical Latent Action Model, proposed by Hanjung Kim et al. of Yonsei University

Python 17 Updated May 6, 2026

ICASSP2026 - Code for "Joint Estimation of Piano Dynamics and Metrical Structure". Estimate Piano dynamics markings from audio, with Bark-scale specific loudness feature extractor in PyTorch.

Python 7 Updated May 17, 2026

Implementation of Flow Matching model for MNIST to understand how it works

Jupyter Notebook 17 2 Updated Sep 19, 2025
Next