Skip to content
View ZiQi21's full-sized avatar

Block or report ZiQi21

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

UAV World Model & VLA & VLM Learning - 无人机领域的世界模型、VLA、VLM综述学习项目

30 4 Updated May 19, 2026
Python 21 Updated Apr 19, 2026

VLX-Flow: streaming VLM for real-time general vision intelligence

97 3 Updated Aug 13, 2026

Awesome World Models for Aerial Navigation

6 Updated Jun 27, 2026

Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations https://video-prediction-policy.github.io

Python 406 34 Updated May 17, 2025

Official PyTorch Implementation of Unified Video Action Model (RSS 2025)

Python 405 31 Updated Jul 23, 2025

[RSS 2026] Causal video-action world model for generalist robot control

Python 1,761 161 Updated Jul 9, 2026
Python 209 17 Updated Aug 1, 2025

🚦 macOS menu bar traffic light for Claude Code — red (working), yellow (waiting for input), green (idle)

Swift 2 1 Updated Jun 8, 2026

Building General-Purpose Robots Based on Embodied Foundation Model

Python 1,214 97 Updated Jul 21, 2026

A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models

Python 780 22 Updated Jun 15, 2026

ICLR 2026 Paper: Ctrl-World

Python 547 54 Updated Apr 8, 2026
Python 425 16 Updated Jul 30, 2026

Markdown + 证件照 → 精美简历(PDF/HTML/PNG)| AI-powered resume generator from Markdown

Python 4 2 Updated May 28, 2026

WAM-Diff: A Masked Diffusion VLA Framework with MoE and Online Reinforcement Learning for Autonomous Driving

Python 297 52 Updated Feb 1, 2026

A curated, continuously updated reading list, paper blogs, and resources for World Action Models (WAMs) in embodied AI.

HTML 1,282 35 Updated Aug 14, 2026

A Curated List of Vision-Language-Action (VLA) and World Action Models (WAM) Research and Beyond

966 36 Updated Aug 8, 2026
Python 370 50 Updated Mar 25, 2026

[ECCV 2026] HUGE-Bench: A Benchmark for High-Level UAV Vision-Language-Action Tasks

Python 44 Updated Aug 10, 2026
Python 273 34 Updated Jan 22, 2026

获取AirSim仿真惯性、视觉数据(解决官方提供的频率较低问题)

C++ 37 2 Updated Apr 28, 2022
Python 204 22 Updated Apr 14, 2026
Python 93 6 Updated Jun 9, 2026

A list of research papers, models, datasets, and other resources on Vision-Language-Action/Navigation (VLA/VLN) models for UAVs.

61 2 Updated Feb 28, 2026

ERNIE-Image is an open text-to-image generation model developed by the ERNIE-Image team at Baidu. It is built on a single-stream Diffusion Transformer (DiT), with only 8B DiT parameters, it reaches…

Python 499 34 Updated Apr 17, 2026

A rebuilt, fully functional version of Anthropic's Claude Code CLI

TypeScript 134 158 Updated Apr 3, 2026

Corpus of Annotations for Misspelings

29 3 Updated Jul 31, 2023
Next