Skip to content
View lixiny's full-sized avatar
🎯
I may be slow to respond.
🎯
I may be slow to respond.

Highlights

  • Pro

Organizations

@MVIG-SJTU @oakink

Block or report lixiny

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A multi-round co-design skill for publication-ready paper framework diagrams and method overview figures.

1,861 105 Updated Jul 10, 2026

Official implementation of LaMP: Learning Vision-Language-Action Policy with 3D Scene Flow as Latent Motion Prior.

Python 7 Updated Aug 10, 2026

Official implementation of ChronoFlow-Policy: a diffusion-based visuomotor policy that jointly models past-current-future object-gripper interaction flows for robot manipulation.

6 Updated Jul 1, 2026

Open-source unified multimodal model

Python 6,144 547 Updated May 4, 2026

[ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"

Python 123 3 Updated Jan 10, 2026

PaperBanana: Automating Academic Illustration For AI Scientists

Python 6,911 516 Updated Jun 25, 2026

The official implementation of InfiniteVGGT

Python 382 21 Updated Apr 19, 2026

Accepted to ECCV 2026

Python 345 24 Updated Jul 6, 2026

A Pragmatic VLA Foundation Model

Python 1,741 180 Updated Jun 11, 2026

Official Repo of The Great March Project. https://www.rhos.ai/research/gm-100

22 1 Updated Jun 17, 2026

[CVPR 2026] InfiniDepth: Arbitrary-Resolution and Fine-Grained Depth Estimation with Neural Implicit Fields

Python 1,066 47 Updated Apr 3, 2026

🔥(CVPR 2025 Highlight) Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera

Python 302 37 Updated Mar 12, 2026

HaWoR: World-Space Hand Motion Reconstruction from Egocentric Videos

Python 324 39 Updated Apr 16, 2026

The official repo for [NeurIPS'22] "ViTPose: Simple Vision Transformer Baselines for Human Pose Estimation" and [TPAMI'23] "ViTPose++: Vision Transformer for Generic Body Pose Estimation"

Python 2,127 267 Updated Dec 25, 2025

ICCV 2025 | TesserAct: Learning 4D Embodied World Models

Python 405 20 Updated Aug 4, 2025

[ICRA 2026] VITRA: Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos

Python 472 33 Updated Jun 12, 2026

Code for "EgoX: Egocentric Video Generation from a Single Exocentric Video"

Python 746 51 Updated Jul 10, 2026

The repository provides code for running inference with the SAM 3D Body Model (3DB), links for downloading the trained model checkpoints and datasets, and example notebooks that show how to use the…

Python 3,433 408 Updated Feb 19, 2026

Momentum Human Rig is an anatomically-inspired parametric full-body digital human model developed at Meta. It includes: A parametric body skeletal model; A realistic 3D mesh skinned to the skeleton…

Python 809 75 Updated Aug 13, 2026
Python 496 51 Updated Jun 23, 2026

[ICLR 2026] A simple state update rule to enhance length generalization for CUT3R

Python 723 34 Updated May 11, 2026

[ICLR 2026] Trace Anything: Representing Any Video in 4D via Trajectory Fields

Python 544 21 Updated Oct 31, 2025

Cosmos-Transfer2.5, built on top of Cosmos-Predict2.5, produces high-quality world simulations conditioned on multiple spatial control inputs.

Python 718 124 Updated Jun 30, 2026

VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model

Python 2,282 208 Updated Mar 19, 2026

[CoRL 2025] TWIST: Teleoperated Whole-Body Imitation System

Python 805 74 Updated Nov 1, 2025

[ICLR 2026] An unified model for 4D human-scene reconstruction

Python 528 44 Updated Dec 30, 2025

A lightweight suite of motion imitation methods for training controllers.

Python 2,225 289 Updated Jun 23, 2026

[arXiv 2025] VisualMimic: Visual Humanoid Loco-Manipulation via Motion Tracking and Generation

Python 298 5 Updated Oct 3, 2025

Unfied World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets

Python 250 19 Updated Oct 8, 2025
Next