Skip to content
View sahilg06's full-sized avatar
💭
what i can't create, I don't understand.
💭
what i can't create, I don't understand.

Block or report sahilg06

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 2 1 Updated Aug 30, 2026

Benchmark to measure what the real knowledge cutoff of a model is

Python 21 2 Updated Sep 2, 2026

[ICML'26] Code and website for Self-Flow: Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis

Python 766 24 Updated May 23, 2026

[ICLR2026] The official code of "Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing Guidance"

Python 51 2 Updated Mar 23, 2026

A curated list of papers and selected technical blogs on Loop Models.

Python 409 14 Updated Sep 22, 2026

A theoretical reconstruction of the Claude Mythos architecture, built from first principles using the available research literature.

Python 14,907 3,296 Updated May 23, 2026

The best OSS video generation models, created by Genmo

Python 3,728 490 Updated Nov 14, 2025

[NeurIPS 2025 Oral] Representation Entanglement for Generation: Training Diffusion Transformers Is Much Easier Than You Think

Python 276 18 Updated Sep 9, 2026

🟣 LLMs interview questions and answers to help you prepare for your next machine learning and data science interview in 2026.

1,064 120 Updated Feb 17, 2026

🔥 LeetCode for PyTorch — practice implementing softmax, attention, GPT-2 and more from scratch with instant auto-grading. Jupyter-based, self-hosted or try online.

Jupyter Notebook 4,604 409 Updated Jul 31, 2026

Autonomous AI movie studio — turn a text prompt into a fully produced video. 100% local, no cloud, no API keys.

Python 205 44 Updated Jul 19, 2026
Python 45 3 Updated Mar 19, 2026

Official release of the benchmark in paper "VSP: Diagnosing the Dual Challenges of Perception and Reasoning in Spatial Planning Tasks for MLLMs"

Python 23 2 Updated Aug 1, 2025

Minimal and highly hackable implementation of Looped Transformers with GPT

Python 25 1 Updated Mar 8, 2026

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Python 166,547 34,658 Updated Sep 23, 2026

DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data (NeurIPS 2023 Spotlight) / / / / When Does Perceptual Alignment Benefit Vision Representations? (NeurIPS 2024)

Python 626 34 Updated Nov 24, 2025

[ICLR 2025] MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts

Python 280 14 Updated Oct 16, 2024

Curated list of methods that focuses on improving the efficiency of diffusion models

45 Updated Jul 9, 2024

A PyTorch implementation of the paper "All are Worth Words: A ViT Backbone for Diffusion Models".

Jupyter Notebook 1,109 77 Updated Mar 25, 2023

https://www.shoufachen.com/Awesome-Diffusion-Transformers/

HTML 147 8 Updated Mar 6, 2024

A PyTorch implementation of the paper "Revisiting Non-Autoregressive Transformers for Efficient Image Synthesis"

Python 47 2 Updated Jun 13, 2024

Minimal reproduction of DeepSeek R1-Zero

Python 13,242 1,575 Updated Feb 27, 2026

Denoising Diffusion Step-aware Models (ICLR2024)

Python 62 1 Updated Feb 6, 2024

[ICML 2024] CLLMs: Consistency Large Language Models

Python 418 25 Updated Nov 16, 2024

[CVPR2025 Highlight] PAR: Parallelized Autoregressive Visual Generation. https://yuqingwang1029.github.io/PAR-project

Python 185 4 Updated Mar 20, 2025

Multi-Agent System Powered by LLMs for End-to-end Multimodal ML Automation

Python 307 58 Updated Sep 1, 2026

[CVPR 2024] DeepCache: Accelerating Diffusion Models for Free

Python 969 51 Updated Jun 27, 2024

Official Implementation for our NeurIPS 2024 paper, "Don't Look Twice: Run-Length Tokenization for Faster Video Transformers".

Python 238 12 Updated Mar 29, 2025

Official codebase used to develop Vision Transformer, SigLIP, MLP-Mixer, LiT and more.

Jupyter Notebook 3,541 226 Updated May 19, 2025

High-Resolution Image Synthesis with Latent Diffusion Models

Jupyter Notebook 14,156 1,735 Updated Feb 29, 2024
Next