Skip to content
View jliu-ac's full-sized avatar
🎯
Focusing
🎯
Focusing
  • The University of Hong Kong
  • Hong Kong SAR

Organizations

@CVMI-Lab

Block or report jliu-ac

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Accompanying code for "Discovering State-of-the-art Reinforcement Algorithms" Nature publication

Python 721 60 Updated Dec 2, 2025

"LLaFEA: Frame-Event Complementary Fusion for Fine-Grained Spatiotemporal Understanding in LMMs", accepted by ICCV 2025

Python 12 Updated Jul 28, 2025
Python 6 1 Updated Dec 23, 2025

(NeurIPS 2025) Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation

Python 77 Updated May 21, 2026

(ICCV 2025) How Far are AI-generated Videos from Simulating the 3D Visual World: A Learned 3D Evaluation Approach

Python 6 Updated Nov 10, 2025

[NeurIPS'25] Official repository of Concerto: Joint 2D-3D Self-Supervised Learning Emerges Spatial Representations

Python 533 29 Updated Apr 7, 2026

ICCV 2025

17 Updated Mar 26, 2026

[ICLR 2026]QeRL enables RL for 32B LLMs on a single H100 GPU.

Python 514 52 Updated Mar 30, 2026

Long Video Gen Infrastructure

Python 2,534 242 Updated Aug 7, 2026

Implementation of Paper “GV-VAD : Exploring Video Generation for Weakly-Supervised Video Anomaly Detection”

Python 9 Updated Oct 9, 2025

Wan: Open and Advanced Large-Scale Video Generative Models

Python 16,813 3,262 Updated Mar 5, 2026

(ICCV 2025) Holistic Tokenizer for Autoregressive Image Generation

Python 34 1 Updated Oct 9, 2025
Python 369 50 Updated Mar 25, 2026

Official code and checkpoint release for mobile robot foundation models: GNM, ViNT, and NoMaD.

Python 1,290 198 Updated Sep 15, 2024

Official code for the CVPR 2025 paper "Navigation World Models".

Python 664 69 Updated Nov 24, 2025

The official implementation of the paper "UrbanWorld: An Urban World Model for 3D City Generation"

Python 57 5 Updated Nov 9, 2024
Python 160 14 Updated May 11, 2026

[NeurIPS 2025]Official repositories for "Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-Thought".

Python 32 Updated Jan 30, 2026

Official Implementation of "VAU-R1: Advancing Video Anomaly Understanding via Reinforcement Fine-Tuning".

Python 70 5 Updated Aug 4, 2026

[CVPR 2026] VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

Python 435 29 Updated Jul 15, 2026

[ICCV'25] Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Python 70 1 Updated Jul 22, 2025

Official codebase used to develop Vision Transformer, SigLIP, MLP-Mixer, LiT and more.

Jupyter Notebook 3,513 224 Updated May 19, 2025

OpenEQA Embodied Question Answering in the Era of Foundation Models

Jupyter Notebook 365 27 Updated Sep 20, 2024

Universal Monocular Metric Depth Estimation

Python 1,241 116 Updated May 18, 2025

[CVPR 2023 Highlight] Perspective Fields for Single Image Camera Calibration

Jupyter Notebook 314 23 Updated Nov 2, 2024

[NeurIPS 2024] Geometry-Aware Large Reconstruction Model for Efficient and High-Quality 3D Generation

Python 173 9 Updated Sep 30, 2024

[ICCV 2025] HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation

Python 259 16 Updated May 12, 2026

[CVPR 2025] UniGoal: Towards Universal Zero-shot Goal-oriented Navigation

Python 347 14 Updated Sep 16, 2025

[ECCV 2024] SGS-SLAM: Semantic Gaussian Splatting For Neural Dense SLAM

Jupyter Notebook 527 49 Updated Nov 20, 2025

[CVPR'24 Highlight & Best Demo Award] Gaussian Splatting SLAM

Python 2,135 230 Updated Aug 7, 2024
Next