-
The University of Hong Kong
- xh-liu.github.io
Stars
[NeurIPS 2025] Wan-Move: Motion-controllable Video Generation via Latent Trajectory Guidance
[SIGGRAPH Asia 2025] OmniPart: Part-Aware 3D Generation with Semantic Decoupling and Structural Cohesion
Math-VR Benchmark & CodePlot-CoT: Mathematical Visual Reasoning by Thinking with Code-Driven Images
Resources and paper list for "Thinking with Images for LVLMs". This repository accompanies our survey on how LVLMs can leverage visual information for complex reasoning, planning, and generation.
[ICLR26] GoT-R1: Unleashing Reasoning Capability of MLLM for Visual Generation with Reinforcement Learning
[CVPR2025 Highlight] PAR: Parallelized Autoregressive Visual Generation. https://yuqingwang1029.github.io/PAR-project
The official implementation of DiM: Diffusion Mamba for Efficient High-Resolution Image Synthesis
A collection of resources on controllable generation with text-to-image diffusion models.
[Neurips 2023 & TPAMI] T2I-CompBench (++) for Compositional Text-to-image Generation Evaluation
A unified framework for 3D content generation.
arXiv LaTeX Cleaner: Easily clean the LaTeX code of your paper to submit to arXiv
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
Segment-Anything + 3D. Let's lift anything to 3D.
General AI methods for Anything: AnyObject, AnyGeneration, AnyModel, AnyTask, AnyX
The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
Pointcept: Perceive the world with sparse points, a codebase for point cloud perception research. Latest works: Utonia (ICML'26), Concerto (NeurIPS'25), Sonata (CVPR'25 Highlight), PTv3 (CVPR'24 Oral)
(ෆ`꒳´ෆ) A Survey on Text-to-Image Generation/Synthesis.
my customized vimrc, using vundle to manage. mainly used for Python Development.
Ph.D. Thesis - Scalable Human Identification with Deep Learning
PyTorch open-source toolbox for unsupervised or domain adaptive object re-ID.
[ICLR-2020] Mutual Mean-Teaching: Pseudo Label Refinery for Unsupervised Domain Adaptation on Person Re-identification.
Research Framework for easy and efficient training of GANs based on Pytorch
Reading list for research topics in multimodal machine learning
PointRCNN: 3D Object Proposal Generation and Detection from Point Cloud, CVPR 2019.
PyTorch re-implementation of DeepLab v2 on COCO-Stuff / PASCAL VOC datasets
The author's officially unofficial PyTorch BigGAN implementation.
GANs with spectral normalization and projection discriminator