Skip to content
View BJHYZJ's full-sized avatar

Block or report BJHYZJ

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 6,064 362 Updated Aug 15, 2026

[CVPR 2026] InfiniDepth: Arbitrary-Resolution and Fine-Grained Depth Estimation with Neural Implicit Fields

Python 1,069 47 Updated Apr 3, 2026

DGGT: Feedforward 4D Reconstruction of Dynamic Driving Scenes using Unposed Images

Python 567 61 Updated Jan 15, 2026

[ECCV 2026]Official implementation of the paper: "FoundationGeo: Learning Spatial Pixel-Wise Fields for Monocular Metric Geometry"

Python 104 1 Updated Jul 29, 2026

MapAnything: Universal Feed-Forward Metric 3D Reconstruction

Python 3,645 280 Updated Aug 7, 2026

Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models

Jupyter Notebook 521 62 Updated Mar 10, 2026

[CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations

Python 595 20 Updated Apr 22, 2026

Let us control diffusion models!

Python 34,066 3,019 Updated Feb 25, 2024

Cameras as Relative Positional Encoding

Python 746 13 Updated Dec 18, 2025

Official Implementation of OmniWeaving: Towards Unified Video Generation with Free-form Composition and Reasoning

Python 845 27 Updated Apr 11, 2026

SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation [CVPR 2026]

Python 106 1 Updated Jun 24, 2026

[CVPR 2024 Highlight] Feature 3DGS: Supercharging 3D Gaussian Splatting to Enable Distilled Feature Fields

C++ 678 50 Updated Oct 17, 2024

[ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning

Python 2,107 166 Updated Jul 3, 2026

NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

Jupyter Notebook 11,521 827 Updated Aug 14, 2026

Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the future state of the world in the form of video.

Python 1,346 190 Updated Jun 8, 2026

Official Implementation of "DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving"

Python 54 1 Updated Jul 19, 2026

News: the 10k dataset is ready for download.

HTML 659 17 Updated Feb 10, 2026

Project Lyra: Open Generative 3D World Models

Python 2,240 227 Updated Jul 20, 2026

ViPE: Video Pose Engine for Geometric 3D Perception

Python 2,070 169 Updated Jun 9, 2026

HorizonDrive: Self-Corrective Autoregressive World Model for Long-horizon Driving Simulation

Python 54 3 Updated Jun 16, 2026

SuperMap is a living spatial memory for embodied AI — it perceives the world, remembers its evolution, and supports reasoning and action.

275 8 Updated Jul 16, 2026

Infinite Worlds with Versatile Interactions

Python 1,515 109 Updated Jul 14, 2026

Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence

Python 920 45 Updated Aug 5, 2026

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

JavaScript 103,303 5,686 Updated Aug 7, 2026

A feed-forward 3D foundation model for reconstructing scenes from streaming data

Python 16,500 1,845 Updated Aug 12, 2026

Collect some World Models for Autonomous Driving (and Robotic, etc.) papers.

2,207 87 Updated Aug 10, 2026
Python 59 1 Updated Jun 23, 2026

[ICLR 2026] Official Implementation of "UniSplat: Unified Spatio-Temporal Fusion via 3D Latent Scaffolds for Dynamic Driving Scene Reconstruction""

Python 87 9 Updated May 22, 2026
Next