-
Shanghai Innovation Institute
- artpli.github.io
- @artpli_
Stars
Code and project page for Human-Robot Partner Juggling
A curated collection of papers, research blogs, open-source tools, benchmarks, and community demos for robot-use agents, including demos powered by GPT-6 Astra. —— Explore our searchable website.
Google Workspace CLI — one command-line tool for Drive, Gmail, Calendar, Sheets, Docs, Chat, Admin, and more. Dynamically built from Google Discovery Service. Includes AI agent skills.
启智平台任务管理 CLI:资源查询、任务提交、日志查看和 MCP/agent workflow
Code and diagnostics for the manipulation benchmark audit paper.
GPT-6 Astra for embodied AI and robotics.
[RSS'26] HoMMI: Learning Whole-Body Mobile Manipulation from Human Demonstrations
A Python library for running thousands of MuJoCo simulations in parallel on CPU
In‑Context World‑Action Modeling from Human Videos for Open‑Ended Task Generalization
An embodied task agent that connects perception, action, verification, and learning in a continuous physical-world loop
Video-Action Models for Generalizable Robot Control Beyond VLAs
GR00T-VisualSim2Real: Open-source sim-to-real framework for humanoid visual loco-manipulation. Train in simulation, deploy zero-shot on real robots with RGB + proprioception for tasks like pick-and…
H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation
[CVPR 2025 highlight] Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision
HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos
egocentric humanoid manipulation benchmark
Open-source Unitree G1 Vision-Language-Action stack for teleop data collection, SonicLatent training, simulation, and real-time whole-body policy deployment(real world deployment TBD).
🎥 Python and OpenCV-based scene cut/transition detection program & library.
Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?
ROSA 🤖 is an AI Agent designed to interact with ROS1- and ROS2-based robotics systems using natural language queries. ROSA helps robot developers inspect, diagnose, understand, and operate robots.
[RSS 2026] The first framework enabling humanoid robots to learn whole-body loco-manipulation from egocentric human demos
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
GigaWorld-Policy: An Efficient Action-Centered World–Action Model
This is the official code repo for DiT4DiT, a Vision-Action-Model (VAM) framework that combines video generation model with flow-matching-based action prediction for generalizable robotic manipulat…
[CoRL'26] Welcome to SIMPLE, a full-stack simulation environment for humanoid loco-manipulation, built on AMO/SONIC, with integrated support for foundation models such as Psi0, Pi05, GR00T, DreamZe…