-
USC
- Los Angeles
Stars
CVPR 2026 Highlight: Pluggable Pruning with Contiguous Layer Distillation for Diffusion Transformers
M0rtzz / paper-notes
Forked from zhaoyang97/Paper-Notes📚 数千篇 AI、LLM、NLP、CV 顶会论文解读,每篇 5 分钟读懂核心思想。
AdaRefSR is a novel reference-based one-step diffusion super-resolution framework. Paper was accepted by ICLR2026.
Spec-driven development (SDD) for AI coding assistants.
[ECCV 2026] SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
🔥中文 prompt 精选🔥,ChatGPT 使用指南,提升 ChatGPT 可玩性和可用性!🚀
OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox. Unlock the magic 🪄: Generative-AI (AIGC), easy-to-use APIs, awsome model zoo, diffusion models, for text-to-image genera…
Enjoy the magic of Diffusion models!
LLM Frontend for Power Users.
An implementation of local windowed attention for language modeling
Production-ready FlashVSR implementation featuring Docker support, NVENC hardware acceleration, Low-VRAM tiling, and unified inference for real-time video super-resolution.
[CVPR 2026] Towards Real-Time Diffusion-Based Streaming Video Super-Resolution — An efficient one-step diffusion framework for streaming VSR with locality-constrained sparse attention and a tiny co…
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
A curated list of resources for video super-resolution using diffusion models.
A simple pip-installable Python tool to generate your HTML citation world map from your Google Scholar ID.
LLM framework for infrastructure as code generation & CloudFormation benchmark.
Iconic font aggregator, collection, & patcher. 3,600+ icons, 50+ patched fonts: Hack, Source Code Pro, more. Glyph collections: Font Awesome, Material Design Icons, Octicons, & more
Master programming by recreating your favorite technologies from scratch.
🙃 A delightful community-driven (with 2,500+ contributors) framework for managing your zsh configuration. Includes 300+ optional plugins (rails, git, macOS, hub, docker, homebrew, node, php, python…
Interactive roadmaps, guides and other educational content to help developers grow in their careers.
Official Implementation for the paper "Discovering Knowledge Deficiencies of Language Models on Massive Knowledge Base"
Inpaint anything using Segment Anything and inpainting models.
[ECCV 2024] Official implementation of the paper "Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection"
Official code of SmartEdit [CVPR-2024 Highlight]
Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation
[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". A…
PyTorch implementation of InstructAny2Pix: Flexible Visual Editing via Multimodal Instruction Following
A instruction data generation system for multimodal language models.