Skip to content
View pmj110119's full-sized avatar
🤒
Outside
🤒
Outside

Block or report pmj110119

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Understanding R1-Zero-Like Training: A Critical Perspective

Python 1,268 60 Updated Aug 27, 2025

Muon is an optimizer for hidden layers in neural networks

Python 2,728 131 Updated May 24, 2026

PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.

Python 13,602 2,162 Updated Jul 24, 2026

Reference PyTorch implementation and models for DINOv3

Jupyter Notebook 11,000 910 Updated Jul 15, 2026

Official repository for "LIV: Language-Image Representations and Rewards for Robotic Control" (ICML 2023)

Python 135 14 Updated Oct 19, 2023

Implementation of advantage-weighted regression.

Python 211 39 Updated May 30, 2020
Python 62 7 Updated Apr 18, 2025

Open-source 3D digital assets for simulation and robot training. Free for non-commercial use.

170 4 Updated Dec 17, 2025

Official Implementation of RoboCLIP (NeurIPS 2023)

Python 53 15 Updated Aug 12, 2024

RoboBrain 2.5: Advanced version of RoboBrain. Depth in Sight, Time in Mind. 🎉🎉🎉

Python 1,115 111 Updated Feb 28, 2026

This repository summarizes recent advances in the VLA + RL paradigm and provides a taxonomic classification of relevant works.

424 5 Updated Oct 10, 2025

[ECCV 2024] 💐Official implementation of the paper "Diffusion Reward: Learning Rewards via Conditional Video Diffusion"

Python 122 11 Updated Jul 2, 2024

Collections of robotics environments geared towards benchmarking multi-task and meta reinforcement learning

Python 1,856 346 Updated Jul 19, 2026

Code for Neural Dynamics Augmented Diffusion Policy, ICRA 2025

Python 10 1 Updated May 26, 2025

PyTorch implementation of YAY Robot

Python 165 9 Updated Apr 7, 2024

Source code for Video-Language Critic: Transferable Reward Functions for Language-Conditioned Robotics

Python 10 3 Updated Dec 2, 2024

Open-source implementation of AlphaEvolve

Python 6,785 1,090 Updated Jul 18, 2026

Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Python 1,311 190 Updated Sep 9, 2025

A final sanity checklist to help your CS paper get accepted, not desk rejected.

1,601 145 Updated May 25, 2026

[CVPR 2025] RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete. Official Repository.

Python 559 44 Updated Oct 13, 2025

Pytorch implementation of Evolutionary Policy Optimization, from Wang et al. of the Robotics Institute at Carnegie Mellon University

Python 110 4 Updated May 18, 2026

MM-EUREKA: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Python 770 29 Updated Sep 7, 2025

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 73,480 8,980 Updated Jul 24, 2026

Recipes for shrinking, optimizing, customizing cutting edge vision models. 💜

Jupyter Notebook 1,964 154 Updated May 26, 2026

Programmer's guide about how to cook at home.

101,341 11,036 Updated Jul 24, 2026

Lets make video diffusion practical!

Python 17,133 1,728 Updated Oct 16, 2025

Single-file implementation to advance vision-language-action (VLA) models with reinforcement learning.

Python 446 20 Updated Nov 8, 2025

Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation

Python 178 4 Updated Jul 17, 2025

AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World | CoRL 2025

Python 99 8 Updated Mar 26, 2026
Python 18 Updated May 24, 2026
Next