Skip to content
View agwmon's full-sized avatar

Highlights

  • Pro

Block or report agwmon

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[Tech Report] Context Scaling: Scaling Properties of Text Conditioning in Visual Generation

Python 143 1 Updated Aug 7, 2026

HY-SOAR:Self-Correction for Optimal Alignment and Refinement in Diffusion Models

Python 741 64 Updated Apr 21, 2026

ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing?

Python 149 10 Updated Jul 30, 2026

OSCAR — public inference release

Python 71 6 Updated Jun 16, 2026

NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

Jupyter Notebook 11,425 815 Updated Aug 7, 2026

Official Implemenation for RAEv2: Improved Baselines with Representation Autoencoders

Python 315 13 Updated May 21, 2026

Official repository of LIBERO-plus, a generalized benchmark for in-depth robustness analysis of vision-language-action models.

Python 411 35 Updated Jan 21, 2026

[ICML'26] Code and website for Self-Flow: Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis

Python 722 22 Updated May 23, 2026

Official PyTorch Implementation of "Flow Map Distillation Without Data"

Python 129 10 Updated Nov 25, 2025

This is the official code repo for DiT4DiT, a Vision-Action-Model (VAM) framework that combines video generation model with flow-matching-based action prediction for generalizable robotic manipulat…

Python 423 32 Updated Jun 30, 2026

Video-Action Models for Generalizable Robot Control Beyond VLAs

Python 295 31 Updated Jun 26, 2026

VLS: Steering Pretrained Robot Policies via Vision–Language Models

Python 67 8 Updated Mar 29, 2026

[ECCV 2026] VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model

Python 525 46 Updated May 2, 2026

Single-stage End-to-End Training for Tokenization and Generation

Python 118 2 Updated Mar 24, 2026

🔥 LeetCode for PyTorch — practice implementing softmax, attention, GPT-2 and more from scratch with instant auto-grading. Jupyter-based, self-hosted or try online.

Jupyter Notebook 4,464 390 Updated Jul 31, 2026

Official Code Repository for the paper "Score-based Generative Modeling of Graphs via the System of Stochastic Differential Equations" (ICML 2022)

Python 193 27 Updated Nov 16, 2023

GigaWorld-Policy: An Efficient Action-Centered World–Action Model

Python 1,385 107 Updated Jul 21, 2026

An end-to-end open ecosystem for robot learning

Python 435 63 Updated Aug 8, 2026

Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Python 1,270 158 Updated Apr 3, 2026

[RSS 2026] Causal video-action world model for generalist robot control

Python 1,738 159 Updated Jul 9, 2026

ACTSmooth extends ACT with prefix conditioning and async inference, eliminating inter-chunk discontinuities and inference latency stalls.

Python 12 Updated Mar 29, 2026

Unfied World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets

Python 249 19 Updated Oct 8, 2025

[ICLR 2026] Official implementation for What matters for Representation Alignment: Global Information or Spatial Structure?

Python 260 14 Updated Dec 15, 2025

ICLR 2026 Paper: Ctrl-World

Python 546 52 Updated Apr 8, 2026

VLA model interpretability tools

Python 180 8 Updated Mar 30, 2026

The first multiplayer video world model in Minecraft

Python 221 9 Updated Mar 3, 2026

(ICML2026) Official implementation of VLANeXt.

Python 222 10 Updated Aug 6, 2026

A simulation evaluation platform for DROID

Python 238 40 Updated Mar 16, 2026
Next