Stars
[CVPR'25] SplineGS: Robust Motion-Adaptive Spline for Real-Time Dynamic 3D Gaussians from Monocular Video
[ICML'26] Code and website for Self-Flow: Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis
A research-friendly PyTorch Lightning toolkit for training, fine-tuning, and evaluating AutoencoderKL for Stable Diffusion and FLUX.
The context API to search, scrape, and interact with the web at scale. 🔥
An implementation of PSGD-QUAD optimizer for PyTorch
QuDAG Protocol (Quantum-Resistant DAG-Based Anonymous Communication System) - Claude Code implementation of a Test-Driven Development Implementation Plan for QuDAG Protocol with Claude Code
Simulation platform for general-purpose robotics & embodied AI learning.
Showing how the SDXL latent space corrections work
A family of compressed models obtained via pruning and knowledge distillation
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
Implementation of Agent Attention in Pytorch
nbardy / DRLX
Forked from CarperAI/DRLXDiffusion Reinforcement Learning Library
CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
Collection and Implementation of Mobile-based Vision Transformer in Pytorch
Implementation of a memory efficient multi-head attention as proposed in the paper, "Self-attention Does Not Need O(n²) Memory"
Fast and memory-efficient exact attention
Visual Taste Approximator (VTA) is a very simple tool that helps anyone create an automatic replica of themselves that can approximate their own personal visual taste
A small CLI app to scrap high-quality movie snapshots from various websites.
Implementation of Muse: Text-to-Image Generation via Masked Generative Transformers, in Pytorch
An official implementation of MobileStyleGAN in PyTorch
[SIGGRAPH'22] StyleGAN-XL: Scaling StyleGAN to Large Diverse Datasets