Stars
Open Robot Product User Manual Management Warehouse
feng-7 / mycobot_ros2
Forked from elephantrobotics/mycobot_ros2myCobot ROS2 package
Open Robot Product User Manual Management Warehouse
A curated list of awesome LLM/VLM/VLA/World Model for Autonomous Driving(LLM4AD) resources (continually updated)
《Hello 算法》:动画图解、一键运行的数据结构与算法教程。支持简中、繁中、English、日本語,提供 Python, Java, C++, C, C#, JS, Go, Swift, Rust, Ruby, Kotlin, TS, Dart 等代码实现
[Embodied-AI-Survey-2025] Paper List and Resource Repository for Embodied AI
[ICLR & NeurIPS 2025] Repository for Show-o series, One Single Transformer to Unify Multimodal Understanding and Generation.
Pytorch implementation of Transfusion, "Predict the Next Token and Diffuse Images with One Multi-Modal Model", from MetaAI
PyTorch Implementation of Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
Automated dense category annotation engine that serves as the initial semantic labeling for the Segment Anything dataset (SA-1B).
[ECCV 2024] Official implementation of the paper "Semantic-SAM: Segment and Recognize Anything at Any Granularity"
Estimating the Focal Length of a Monocular Image
Tame a Wild Camera: In-the-Wild Monocular Camera Calibration
The repo for "Metric3D: Towards Zero-shot Metric 3D Prediction from A Single Image" and "Metric3Dv2: A Versatile Monocular Geometric Foundation Model..."
Code for "Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation"
PyTorch implementation of MAR+DiffLoss https://arxiv.org/abs/2406.11838
[WACV 2024 Survey Paper] Multimodal Large Language Models for Autonomous Driving
PyTorch code for Vision Transformers training with the Self-Supervised learning method DINO
a simple unofficial implementation of classifier-free diffusion guidance
Implementation of Denoising Diffusion Probabilistic Model in Pytorch
Pytorch implementation of Diffusion Models (https://arxiv.org/pdf/2006.11239.pdf)
A PyTorch implementation of MAGE: MAsked Generative Encoder to Unify Representation Learning and Image Synthesis
Implementation of TiTok, proposed by Bytedance in "An Image is Worth 32 Tokens for Reconstruction and Generation"
Vector (and Scalar) Quantization, in Pytorch