Skip to content
View swayducky's full-sized avatar

Block or report swayducky

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

AudioLDM: Generate speech, sound effects, music and beyond, with text.

Python 2,905 271 Updated Jun 25, 2025

An open source implementation of Microsoft's VALL-E X zero-shot TTS model. Demo is available in https://plachtaa.github.io/vallex/

Python 7,935 783 Updated Feb 11, 2024

Automatically create prompts and make them fight each other to know which is the best

Vue 604 69 Updated Sep 1, 2023

structured outputs for llms

Python 13,726 1,185 Updated Aug 9, 2026

Code for robust monocular depth estimation described in "Ranftl et. al., Towards Robust Monocular Depth Estimation: Mixing Datasets for Zero-shot Cross-dataset Transfer, TPAMI 2022"

Python 5,424 725 Updated Aug 23, 2024

High accuracy RAG for answering questions from scientific documents with citations

Python 9,022 906 Updated Aug 12, 2026

Real-time end-to-end singing voice conversion system based on DDSP (Differentiable Digital Signal Processing)

Python 2,647 286 Updated Aug 12, 2026

Singing Voice Conversion via diffusion model

Jupyter Notebook 2,717 811 Updated Jun 6, 2026

PyTorch code for BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation

Jupyter Notebook 5,720 765 Updated Mar 3, 2026

An experimental open-source attempt to make GPT-4 fully autonomous.

Python 3 Updated Sep 21, 2023

Learning in public.

Python 3 Updated Sep 10, 2024

An unofficial PyTorch implementation of the audio LM VALL-E

Python 2,979 398 Updated May 10, 2023