- Hong Kong
Highlights
- Pro
Stars
binary releases of VS Code without MS branding/telemetry/licensing
Text2midi is the first end-to-end model for generating MIDI files from textual descriptions. By leveraging pretrained large language models and a powerful autoregressive transformer decoder, text2m…
You like Beamer? You need Powerpoint? Fake Beamer is for you!
woct0rdho / SageAttention
Forked from thu-ml/SageAttentionFork of SageAttention for Windows wheels and easy installation
Official repository for "PosterO: Structuring Layout Trees to Enable Language Models in Generalized Content-Aware Layout Generation" (CVPR 2025).
ACE-Step: A Step Towards Music Generation Foundation Model
(CVPR2025) Learned Image Compression with Dictionary-based Entropy Model
Add virtual monitors to your windows 10/11 device! Works with VR, OBS, Sunshine, and/or any desktop sharing software.
Collection of audio-focused loss functions in PyTorch
GUI for a Vocal Remover that uses Deep Neural Networks.
[ISMIR 2019] A Bi-Directional Transformer for Musical Chord Recognition
The NES Music Database: use machine learning to compose music for the Nintendo Entertainment System!
Text To Video Synthesis Colab
NotaGen: Advancing Musicality in Symbolic Music Generation with Large Language Model Training Paradigms
Pytorch port of Google Research's VGGish model used for extracting audio features.
Algorithms for performing style transfer on videos.
This code extends the neural style transfer image processing technique to video by generating smooth transitions between several reference style images
QualityScaler - image/video AI upscaler app
Muzic: Music Understanding and Generation with Artificial Intelligence
[CoRL '23] Dexterous piano playing with deep reinforcement learning.
XMIDI Dataset: A large-scale symbolic music dataset with emotion and genre labels.
Second in series "Project StarBurst". RVC lesson stream downloader.