Lists (1)
Sort Name ascending (A-Z)
Starred repositories
This repository provides valuable reference for researchers in the field of multimodality, please start your exploratory travel in RL-based Reasoning MLLMs!
Differentiable Point Radiance Fields Rasteriser for Novel View Synthesis
[ECCV 2026] Official implementation of the paper "3D-Aware VLMs with Implicit and Explicit Geometries"
Stanford CS 269Q: Quantum Computer Programming
Research @ Stanford in historical Quantum Algorithms to replicate a Quantum Black Box
HilDA: Hierarchical Distillation with Diffusion for Advancing Self-Supervised LiDAR Pre-training
mehrsapo / Flower
Forked from annegnx/PnP-Flow[ICLR 2026] Flower: A Flow Matching Solver for Inverse Problems
[CVPR 2026 Best Paper Finalist] Pixel Diffusion Transformers for Image Generation
[ICLR 2025] SPA: 3D Spatial-Awareness Enables Effective Embodied Representation
About This repository is a curated collection of the most exciting and influential CVPR 2026 papers. 🔥 [Paper + Code + Demo]
World Model Self-Distillation project website
Official implementation of MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models
Official implementation of Surflo: Consistent 3D Surface Flow Model with Global State.
Office inference code for World Tracing (object/scene/dynamic). Live demos: https://haoz19.github.io/world-tracing-page/
Official code release for Reroute, Don’t Remove: Recoverable Visual Token Routing for Vision-Language Models.
[ICML 2026] PyTorch implementation of BudCache
Flex4DHuman turns monocular or sparse multi-view videos of dynamic subjects into synchronized dense multi-view videos.
Code for RepWAM: World Action Modeling with Representation Visual-Action Tokenizers
Envision4D: Envisioning Visual Futures via Feed-forward 4D Gaussian Splatting for Autonomous Driving
WorldOlympiad: Can Your World Model Survive a Triathlon?
Benchmarking MLLMs for Parametric 3D Generation and Structural Reasoning (Text-to-3D, Image-to-3D, Assembly-3D)
[CVPR 2025🔥] Official codebase for "Global-Local Tree Search in VLMs for 3D Indoor Scene Generation" and our arxiv 2026 extension
Official Pytorch implementation of the paper: "SAM-Flow: Source Anchored Masked Flow for Training-Free Image Editing"