-
The University of Tokyo
- Tokyo
- https://xg-chu.github.io
- in/xg-chu
- https://scholar.google.com/citations?user=yr4kSUsAAAAJ&hl=en
Highlights
- Pro
Lists (1)
Sort Name ascending (A-Z)
Stars
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
"OpenHarness: Open Agent Harness with a Built-in Personal Agent--Ohmo!"
A collection of DESIGN.md files analysis by popular brand design systems. Drop one into your project and let coding agents generate a matching UI.
[CVPR 2026] UniLS: End-to-End Audio-Driven Avatars for Unified Listening and Speaking
Master programming by recreating your favorite technologies from scratch.
A Semantically Consistent Dataset for Data-Efficient Query-Based Universal Sound Separation
Sharp Monocular View Synthesis in Less Than a Second
[NeurIPS 2025] I2-NeRF: Learning Neural Radiance Fields Under Physically-Grounded Media Interactions
The repository provides code for running inference with the SAM 3D Body Model (3DB), links for downloading the trained model checkpoints and datasets, and example notebooks that show how to use the…
An app that allows you to easily change icons of your favorite sites on Safari
Virtual whiteboard for sketching hand-drawn like diagrams
Official implementation of "Bilateral Normal Integration" (BiNI), ECCV 2022.
[NeurIPS 2025] Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
[Official Code] Pixel3DMM: Versatile Screen-Space Priors for Single-Image 3D Face Reconstruction
ICCV 2025 ACTalker: an end-to-end video diffusion framework for talking head synthesis that supports both single and multi-signal control (e.g., audio, expression).
FaceVerse: a Fine-grained and Detail-controllable 3D Face Morphable Model from a Hybrid Dataset (CVPR2022)
Official Pytorch Implementation of SMIRK: 3D Facial Expressions through Analysis-by-Neural-Synthesis (CVPR 2024)
Official implementation for the SIGGRAPH Asia 2024 paper SPARK: Self-supervised Personalized Real-time Monocular Face Capture
[CVPR 2025] Code for Deformable Radial Kernel Splatting
WebGL point cloud viewer for large datasets
ARTalk generates realistic 3D head motions (lip sync, blinking, expressions, head poses) from audio in ⚡ real-time ⚡.
A complete head tracking pipeline from videos to NeRF/3DGS-ready datasets.
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". A…
CUDA accelerated rasterization of gaussian splatting
Portrait4D: Learning One-Shot 4D Head Avatar Synthesis using Synthetic Data (CVPR 24); Portrait4D-v2: Pseudo Multi-View Data Creates Better 4D Head Synthesizer (ECCV 2024)