Skip to content
View prestonfu's full-sized avatar

Highlights

  • Pro

Block or report prestonfu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Official implementation for pi0 steering via DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)

Jupyter Notebook 1 Updated Jul 15, 2026

NanoGPT (124M) in 90 seconds

Python 5,603 852 Updated Jul 28, 2026

Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals

Python 2,508 216 Updated Apr 19, 2026

This project includes code for using the AsyncWebRL and WebGym frameworks to train web agent models.

Python 46 6 Updated Jun 9, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 87,666 20,045 Updated Jul 30, 2026

Various Gradescope autograder templates.

Python 2 Updated Mar 3, 2025

gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI

Python 20,273 2,131 Updated Jul 24, 2026
Python 141 8 Updated Dec 9, 2025

Minimal reproduction of DeepSeek R1-Zero

Python 13,209 1,582 Updated Feb 27, 2026

A benchmark for offline goal-conditioned RL and offline RL

Python 445 96 Updated Jan 14, 2026

Re-implementation of pi0 vision-language-action (VLA) model from Physical Intelligence

Python 1,508 108 Updated Jan 31, 2025

Simulation platform for general-purpose robotics & embodied AI learning.

Python 29,666 2,821 Updated Jul 30, 2026

Examples and guides for using the Gemini API

Jupyter Notebook 17,584 2,712 Updated Jul 30, 2026

Open Overleaf/ShareLaTex projects in vscode, with full collaboration support.

TypeScript 1,621 67 Updated Jul 11, 2026

An LLM-powered knowledge curation system that researches a topic and generates a full-length report with citations.

Python 30,399 2,842 Updated Sep 30, 2025

A framework for Reinforcement Learning research.

Python 271 23 Updated Jul 28, 2026

The repo of paper `RoboMamba: Multimodal State Space Model for Efficient Robot Reasoning and Manipulation`

Python 151 13 Updated Dec 22, 2024

AlphaFold 3 inference pipeline.

Python 8,378 1,315 Updated Jul 28, 2026

🤖 The Full Process Python Package for Robot Learning from Demonstration and Robot Manipulation

Python 717 62 Updated May 19, 2025

Official inference framework for 1-bit LLMs

C++ 39,791 3,657 Updated Jul 27, 2026

[arXiv 2023] Set-of-Mark Prompting for GPT-4V and LMMs

Python 1,551 111 Updated Aug 19, 2024

A plotting tool that outputs Line Rider maps, so you can watch a man on a sled scoot down your loss curves. 🎿

Python 333 6 Updated Aug 23, 2024

Implementation of Diffusion Transformer (DiT) in JAX

Python 320 12 Updated Jun 11, 2024

LLM training in simple, raw C/CUDA

Cuda 30,677 3,713 Updated Jun 26, 2025

This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.

Python 12,155 1,061 Updated Mar 8, 2026

Large World Model -- Modeling Text and Video with Millions Context

Python 7,426 562 Updated Oct 19, 2024

Minimal, clean code for the Byte Pair Encoding (BPE) algorithm commonly used in LLM tokenization.

Python 10,649 1,081 Updated Jul 1, 2024
2 Updated Jan 13, 2024
Next