Stars
ethanhharrison / swissai-dsrl
Forked from lasgroup/swissai-dsrlOfficial implementation for pi0 steering via DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)
Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals
This project includes code for using the AsyncWebRL and WebGym frameworks to train web agent models.
A high-throughput and memory-efficient inference and serving engine for LLMs
Various Gradescope autograder templates.
gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
Minimal reproduction of DeepSeek R1-Zero
A benchmark for offline goal-conditioned RL and offline RL
Re-implementation of pi0 vision-language-action (VLA) model from Physical Intelligence
Simulation platform for general-purpose robotics & embodied AI learning.
Examples and guides for using the Gemini API
Open Overleaf/ShareLaTex projects in vscode, with full collaboration support.
An LLM-powered knowledge curation system that researches a topic and generates a full-length report with citations.
A framework for Reinforcement Learning research.
The repo of paper `RoboMamba: Multimodal State Space Model for Efficient Robot Reasoning and Manipulation`
AlphaFold 3 inference pipeline.
🤖 The Full Process Python Package for Robot Learning from Demonstration and Robot Manipulation
Official inference framework for 1-bit LLMs
[arXiv 2023] Set-of-Mark Prompting for GPT-4V and LMMs
A plotting tool that outputs Line Rider maps, so you can watch a man on a sled scoot down your loss curves. 🎿
Implementation of Diffusion Transformer (DiT) in JAX
This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
Large World Model -- Modeling Text and Video with Millions Context
Minimal, clean code for the Byte Pair Encoding (BPE) algorithm commonly used in LLM tokenization.