Highlights
- Pro
Stars
Developer Hub for NVIDIA Alpamayo, containing ready-to-use recipes for fine-tuning, reinforcement-learning post-training, quantization, and deployment.
A system for building 3D Scene Graphs from sensor data in real-time
AI agents running research on single-GPU nanochat training automatically
An agentic skills framework & software development methodology that works.
GMMap: Memory-Efficient Continous Occupancy Map Using Gaussian Mixture Model
FULL Augment Code, Claude Code, Cluely, CodeBuddy, Comet, Cursor, Devin AI, Junie, Kiro, Leap.new, Lovable, Manus, NotionAI, Orchids.app, Perplexity, Poke, Qoder, Replit, Same.dev, Trae, Traycer AI…
Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepowe…
Official repository for OmniVLA training and inference code
A feed-forward 3D foundation model for reconstructing scenes from streaming data
The GEP-powered self-evolving engine for AI agents. Auditable evolution with Genes, Capsules, and Events. | evomap.ai
Opinionated skills for AI coding agents to create stunning diagrams and visualizations directly in Markdown. These skills extend agent capabilities across diagram generation, data visualization, an…
The simplest, fastest repository for training/finetuning medium-sized GPTs.
Minimal AI coding agent (~1,000 lines of Python) inspired by Claude Code. Works with any LLM. Think NanoGPT for coding agents. Formerly NanoCoder.
MiniMax LLM + Pi05 VLA Robot Agent Demo
高德地图 JSAPI Skills 是一套专为 AI IDE 设计的 AI 编程技能包。它将高德地图 JavaScript API v2.0 的官方文档、最佳实践和代码模板整合为结构化的技能文件,使 Cursor、Claude、Cline 的 AI Codiing 工具能够: 精准理解高德地图 API 的使用方法 自动生成符合官方规范的地图代码 主动避免常见的开发陷阱和安全问题 提供经过验证…
GigaBrain-0: A World Model-Powered Vision-Language-Action Model
Dimensional is the agentic operating system for physical space. Command humanoids, quadrupeds, drones, and other hardware platforms in natural language and build multi-agent systems that work seaml…
Visualize, query, and stream to train on multimodal robotics data.
[ICRA 2026] Official implementation of the paper: "StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling"
Ultra-lightweight, open-source, self-hosted personal AI agent framework in Python with WebUI, tools, memory, MCP, multi-agent workflows, automation, and chat apps
A PyTorch Library for Meta-learning Research
A simple toolkit for training semantic segmentation models from autolableled data
Codebase for "Less is More 🍋: Scalable Visual Navigation from Limited Data"