-
AIS, HKUST
- HKSAR, China
-
00:28
(UTC +08:00) - https://orcid.org/0009-0004-9865-4105
Highlights
- Pro
Stars
This repo is the official implementation of "τ0-VLA: a Hierarchical Robot Foundation Model with World-Model-Guided Test-Time Computation".
Official project page and open-source entry for Video-MME-Logical
EVA-Client: A Unified Framework for Deployment, Evaluation, and Data Collection on Real Robots
Official page of ImmerIris: A Large-Scale Dataset and Benchmark for Off-Axis and Unconstrained Iris Recognition in Immersive Applications.
A curated list of relighting papers, datasets, benchmarks, demos, and tools.
Implementation for paper "Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models".
[SIGGRAPH 2026 / TOG] Official code of the paper "UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors".
Steering Forward-Process Reinforcement Learning for Distilled Autoregressive Video Models
A unified inference and post-training framework for accelerated video generation.
NVIDIA FastGen: Fast Generation from Diffusion Models
[CVPR 2026 Highlight] Official implementation of BiCo: Composing Concepts from Images and Videos via Concept-prompt Binding
[CVPR 2025] Analyzing the Synthetic-to-Real Domain Gap in 3D Hand Pose Estimation
AnyTalker: Scaling Multi-person Talking Video Generation with Interactivity Refinement
[SIGGRAPH Asia 2024] I2VEdit: First-Frame-Guided Video Editing via Image-to-Video Diffusion Models
Xianghao's academic homepage with AcadHomepage
[HKUST MATH5470 25Fall Project1] This project focuses on building a robust credit default prediction model using the Home Credit Default Risk dataset (Kaggle).
A project page template for academic papers. Demo at https://eliahuhorwitz.github.io/Academic-project-page-template/
zero-shot voice conversion & singing voice conversion, with real-time support
Triton implementation of FlashAttention2 that adds Custom Masks.
Calculating the actual value of your job beyond just salary
Cosmos-Transfer1-DiffusionRenderer: High-quality video de-lighting and re-lighting based on Cosmos video diffusion framework
Repository of paper "Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis" (ACL 2025 Main)
Official repository of the video reasoning benchmark MMR-V. Can Your MLLMs "Think with Video"? [ICLR26]
Enjoy the magic of Diffusion models!
Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and…