-
Institute of Automation, Chinese Academy of Science
- Beijing, People's Republic of China
- https://lrmbbj.github.io/
Stars
A curated list of recent diffusion models for video generation, editing, and various other applications.
X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining
Post-training with Tinker
VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo
Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals
主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题
StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing
A paper list for spatial reasoning
A Curated List of Awesome Works in World Modeling, Aiming to Serve as a One-stop Resource for Researchers, Practitioners, and Enthusiasts Interested in World Modeling.
A curated list of state-of-the-art research in embodied AI, focusing on vision-language-action (VLA) models, vision-language navigation (VLN), and related multimodal learning approaches.
Curated visual catalog of 155+ vision-language model (VLM/MLLM) architectures: papers, diagrams, training recipes, datasets, and a release timeline for multimodal AI agents.
A comprehensive list of papers investigating physical cognition in video generation, including papers, codes, and related websites.
[ICLR 2025 Oral] The official implementation of "Diffusion-Based Planning for Autonomous Driving with Flexible Guidance"
[Lumina具身智能社区] 具身智能技术指南 Embodied-AI-Guide
A collection of awesome video generation studies.
⏰ Agenticly track worldwide conference deadlines (Website, Python Cli, Wechat Applet)
Curated list of papers and resources focused on 3D Gaussian Splatting, intended to keep pace with the anticipated surge of research in the coming months.
Realistic Progression One - Career mode for Realism Overhaul
Lightweight Armoury Crate alternative for Asus laptops with nearly the same functionality. Works with ROG Zephyrus, Flow, TUF, Strix, Scar, ProArt, Vivobook, Zenbook, Expertbook, ROG Ally, and many…
A curated list of awesome LLM/VLM/VLA/World Model for Autonomous Driving(LLM4AD) resources (continually updated)
[ICML 2024] The offical implementation of A2PR, a simple way to achieve SOTA in offline reinforcement learning with an adaptive advantage-guided policy regularization method, in Pytorch
[IJCAI'24 Oral] An index of algorithms, approaches, and systems on cross-domain policy transfer for embodied agents
A simple utility that cold patches dwm (uDWM.dll) in order to disable window rounded corners in Windows 11