Lists (2)
Sort Name ascending (A-Z)
Stars
[NeurIPS 2023] Implementation of "PAPR: Proximity Attention Point Rendering"
Music repair method to convert lossy MP3 compressed music to lossless music.
Wan: Open and Advanced Large-Scale Video Generative Models
real time face swap and one-click video deepfake with only a single image
AI-Generated Presets for Faithful 4K Color Style Transfer in Real Time [CVPR 2023]
This is the open-source implement the paper "Color Transfer between Images" by Erik Reinhard, Michael Ashikhmin, Bruce Gooch and Peter Shirley.
Reference code for the paper HistoGAN: Controlling Colors of GAN-Generated and Real Images via Color Histograms (CVPR 2021).
Integrating ChatGPT into your browser deeply, everything you need is here
๐ฅ CNN for Watermark Removal using Deep Image Prior with Pytorch ๐ฅ.
singing voice change based on whisper, and lora for singing voice clone
AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head
A lightweight yet powerful audio-to-MIDI converter with pitch bend detection
Implementation of MusicLM, a text to music model published by Google Research, with a few modifications.
Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch
"Pop Music Transformer: Beat-based Modeling and Generation of Expressive Pop Piano Compositions", ACM Multimedia 2020
MusicTransformer written for MaestroV2 using the Pytorch framework for music generation
Global Rhythm Style Transfer Without Text Transcriptions
This repository contains the source code for the implementation of two deep learning models concerning the audio super resolution task.
In this project we combine techniques from neural voice cloning and musical instrument synthesis to achieve good results from as little as 16 seconds of target data.
Official implementation of DualCycleGAN for nonparallel audio super resolution
A LLM based research assistant that allows you to have a conversation with a research paper
AudioLDM: Generate speech, sound effects, music and beyond, with text.
Apply diffusion models using the new Hugging Face diffusers package to synthesize music instead of images.
NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates @ INTERSPEECH 2022
Tacotron 2 - PyTorch implementation with faster-than-realtime inference
A timeline of the latest AI models for audio generation, starting in 2023!
A collection of pre-trained audio models, in PyTorch.
Audio generation using diffusion models, in PyTorch.