Skip to content
View Zhoues's full-sized avatar
😎
keep doing research!
😎
keep doing research!

Organizations

@camel-ai

Block or report Zhoues

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Skills for Real Engineers. Straight from my .agents directory.

Shell 213,649 18,440 Updated Aug 7, 2026
Python 139 6 Updated Jul 31, 2026

Official Codebase for "Do as I Do: Dexterous Manipulation Data from Everyday Human Videos"

Python 360 40 Updated Aug 2, 2026

A Web-Scale 4D Hand-Object Interaction Data Engine for Any-View Robot Retargeting and Video-to-Action Robot Learning

Python 105 5 Updated Jun 18, 2026

Provide with pre-build flash-attention 2 and 3 package wheels on Linux and Windows using GitHub Actions

Python 1,680 76 Updated Aug 11, 2026

A curated, continuously updated reading list, paper blogs, and resources for World Action Models (WAMs) in embodied AI.

HTML 1,262 33 Updated Aug 11, 2026

A paper list for Learning-based 3D Vision.

151 2 Updated May 2, 2026

A practical toolkit for process-level robot evaluation with Process Reward Models (PRMs).

Jupyter Notebook 128 2 Updated Jul 29, 2026

A paper list of Awesome Latent Space.

956 41 Updated Jul 13, 2026

📚 A curated collection of papers and open-source code repositories dedicated to the application of Vision-Language Models (VLMs) for streaming video.

191 5 Updated Jul 22, 2026

Panoramic Affordance Prediction (PAP) (ECCV 2026)

Python 46 Updated Jun 29, 2026

[RSS 2026] Interactive World Simulator for Robot Policy Training and Evaluation

Python 283 22 Updated Jun 4, 2026
Python 823 93 Updated May 6, 2026

[CVPR 2025] VideoWorld is a simple generative model that learns purely from unlabeled videos—much like how babies learn by observing their environment.

Python 794 40 Updated Feb 25, 2026

Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals

Python 2,534 222 Updated Apr 19, 2026

[NeurIPS 2025 Spotlight] Towards Understanding Camera Motions in Any Video

HTML 306 35 Updated Mar 5, 2026

CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning

Python 30 2 Updated May 23, 2026

[RSS 2026] Causal video-action world model for generalist robot control

Python 1,744 161 Updated Jul 9, 2026

Cosmos Policy

Python 852 96 Updated Jan 23, 2026

Being-H is BeingBeyond's family of human-centric embodied foundation models.

Python 1,117 61 Updated Aug 4, 2026

SPAgent, a foundation agent for understanding, reasoning over, and operating within the physical and spatial world.

Python 214 31 Updated Aug 9, 2026
Python 63 3 Updated Jul 6, 2025

Orient Anything V2, NeurIPS 2025 Spotlight

Python 240 10 Updated Jan 19, 2026
Python 139 2 Updated Jul 16, 2026

[ICCV 2025] VLM4D: Towards Spatiotemporal Awareness in Vision Language Models

Python 55 2 Updated Nov 20, 2025

Official code for "TraceGen: World Modeling in 3D Trace-Space Enables Learning from Cross-Embodiment Videos" (CVPR 2026)

Python 18 5 Updated Jan 31, 2026
Jupyter Notebook 75 7 Updated Apr 21, 2026

[NeurIPS 2025] Pixel-Perfect Depth

Python 1,059 38 Updated Feb 13, 2026

PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation

418 9 Updated Mar 11, 2026
Next