Stars
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM
SimWAM: A Simple World Action Model for End-to-End Autonomous Driving.
Implementation of the ROS Middleware (rmw) Interface using eProsima's Fast RTPS.
official implementation of [CE-Nav: Flow-Guided Reinforcement Refinement for Cross-Embodiment Local Navigation]
[IEEE RA-L'25] NavRL: Learning Safe Flight in Dynamic Environments (NVIDIA Isaac/Python/ROS1/ROS2)
A curated list of awesome robot descriptions (URDF, MJCF)
《深入理解 AI Agent:设计原理与工程实践》(李博杰 著)开源主仓库:全书正文、编译版 PDF 与按章配套代码
[CVPR 2025, Spotlight] SimLingo (CarLLava): Vision-Only Closed-Loop Autonomous Driving with Language-Action Alignment
Infinite Interactive World Rollout on a Single Desktop GPU
Denoising Diffusion Probabilistic Models
PyTorch implementation of JiT https://arxiv.org/abs/2511.13720
Official implementation of "ResAD: Normalized Residual Trajectory Modeling for End-to-End Autonomous Driving"
[CVPR 2025 Oral] Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
[ICLR2026] Official implementation for "JanusVLN: Decoupling Semantics and Spatiality with Dual Implicit Memory for Vision-Language Navigation"
InternRobotics' open platform for building generalized navigation foundation models.
[CVPR 2026] Official implementation of FantasyVLN: Unified Multimodal Chain-of-Thought Reasoning for Vision-and-Language Navigation
[ICLR 2026] ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
Code For AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving
Official codebase for "Diffusion Bridge Implicit Models" (ICLR 2025) and "Consistency Diffusion Bridge Models" (NeurIPS 2024)
Claude Code skills for academic papers: deep analysis, comics, summaries | 论文工艺:深度解读、漫画生成、速览总结
[ICLR'23 Spotlight & ECCV'24 & IJCV'24] MapTR: Structured Modeling and Learning for Online Vectorized HD Map Construction