Skip to content
View ZebinX's full-sized avatar

Block or report ZebinX

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation

59 1 Updated Mar 25, 2026

iMac: Translating Actions into Motion and Contact Images for Embodied World Models

Python 33 Updated Jun 21, 2026

[ICLR 2026 Oral] Latent Particle World Models official repository

Jupyter Notebook 132 5 Updated Mar 19, 2026

[ICRA 2025] Towards Safe End-to-end Autonomous Driving via Online Map Uncertainty

Python 16 3 Updated Nov 14, 2024

Wan: Open and Advanced Large-Scale Video Generative Models

Python 17,021 2,145 Updated Mar 17, 2026

Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals

Python 2,527 221 Updated Apr 19, 2026

[CVPR 2026]

Python 95 12 Updated May 20, 2026
Python 4 Updated Aug 14, 2025

SE-Agent is a self-evolution framework for LLM Code agents. It enables trajectory-level evolution to exchange information across reasoning paths via Revision, Recombination, and Refinement, expandi…

Python 282 29 Updated Sep 23, 2025

[CVPR 2026 Oral] Learning to Drive via Real-World Simulation at Scale

Python 315 26 Updated Jul 24, 2026

Unofficial implementation of Titans, SOTA memory for transformers, in Pytorch

Python 1,972 206 Updated Jul 13, 2026

The official repo of "Mimir: Hierarchical Goal-Driven Diffusion with Uncertainty Propagation for End-to-End Autonomous Driving"

Python 20 2 Updated Apr 26, 2026

VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model

Python 2,275 206 Updated Mar 19, 2026

StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing

Python 3,417 440 Updated Aug 7, 2026
Jupyter Notebook 6 Updated Mar 13, 2025
Python 13 Updated Mar 16, 2025

[CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning

Python 348 14 Updated Apr 18, 2026

Official implementation for DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)

Python 217 30 Updated Aug 5, 2025

Official implementation of the paper "HUGSIM: A Real-Time, Photo-Realistic and Closed-Loop Simulator for Autonomous Driving"

Python 408 47 Updated Nov 8, 2025

MemGen: Weaving Generative Latent Memory for Self-Evolving Agents

Python 408 38 Updated Jun 10, 2026

A curated list of state-of-the-art research in embodied AI, focusing on vision-language-action (VLA) models, vision-language navigation (VLN), and related multimodal learning approaches.

3,445 162 Updated Aug 7, 2026

Implementation of [CVPR 2025] "DiffSensei: Bridging Multi-Modal LLMs and Diffusion Models for Customized Manga Generation"

Python 922 99 Updated Feb 5, 2025

[ICLR 2026] ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving

Python 580 71 Updated Jun 20, 2026

[CVPR2022] Remember Intentions: Retrospective-Memory-based Trajectory Prediction

Python 140 19 Updated Sep 11, 2022

HE-Drive: Human-Like End-to-End Driving with Vision Language Models

Python 253 16 Updated Aug 17, 2025
Python 284 36 Updated Feb 9, 2026

MOSS-TTSD is a spoken dialogue generation model designed for expressive multi-speaker synthesis. It features long-context modeling, flexible speaker control, and multilingual support, while enablin…

Python 1,378 134 Updated Jul 26, 2026

A curated list of awesome papers on Embodied AI and related research/industry-driven resources.

527 25 Updated Jun 3, 2025

Lumina Robotics Talent Call | Lumina社区具身智能招贤榜 | A list for Embodied AI / Robotics Jobs (PhD, RA, intern, etc

1,459 28 Updated Feb 25, 2026
Next