Skip to content
View jpli02's full-sized avatar
🎯
Focusing
🎯
Focusing
  • UIUC CS
  • Earth
  • 02:44 (UTC -06:00)

Block or report jpli02

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

JavaScript 240,061 36,436 Updated Aug 13, 2026

This repo is meant to serve as a guide for Machine Learning/AI technical interviews.

Jupyter Notebook 8,756 1,524 Updated Jul 30, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,258 81,188 Updated Aug 14, 2026

📊 An infographics generator with 30+ plugins and 300+ options to display stats about your GitHub account and render them as SVG, Markdown, PDF or JSON!

JavaScript 17,054 2,253 Updated May 29, 2026

RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI

Python 4,532 653 Updated Aug 13, 2026

dLLM: Simple Diffusion Language Modeling

Python 2,669 281 Updated Jul 17, 2026

A unified inference and post-training framework for accelerated video generation.

Python 3,944 402 Updated Aug 14, 2026

Flexible and Pluggable Serving Engine for Diffusion LLMs

Python 151 18 Updated Jul 13, 2026

An interface library for RL post training with environments.

Python 2,499 425 Updated Aug 13, 2026

A Lightweight LLM Post-Training Library

Python 2,405 331 Updated Aug 14, 2026

Multi-Turn RL Training System with AgentTrainer for Language Model Game Reinforcement Learning

Python 66 12 Updated Dec 18, 2025

Training Large Language Model to Reason in a Continuous Latent Space

Python 1,683 187 Updated Jul 2, 2026

An open-source AI coding agent that lives in your terminal.

TypeScript 26,990 2,856 Updated Aug 14, 2026

A benchmark for LLMs on complicated tasks in the terminal

Python 2,533 563 Updated Jul 11, 2026

[NeurIPS 2025 Spotlight] Reasoning Environments for Reinforcement Learning with Verifiable Rewards

Python 1,485 124 Updated Apr 17, 2026

Checkpoint-engine is a simple middleware to update model weights in LLM inference engines

Python 996 102 Updated Aug 12, 2026

Our library for RL environments + evals

Python 4,508 643 Updated Aug 14, 2026
Python 652 67 Updated Aug 28, 2025

Implementation for FP8/INT8 Rollout for RL training without performence drop.

Python 309 23 Updated Nov 7, 2025

Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL

Python 5,293 479 Updated Nov 13, 2025

[NeurIPS 2025] Simple extension on vLLM to help you speed up reasoning model without training.

Python 232 33 Updated May 31, 2025
Python 890 52 Updated Sep 15, 2025

gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI

Python 20,308 2,136 Updated Jul 24, 2026

The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

Python 5,666 574 Updated Aug 14, 2026

Simple RL training for reasoning

Python 3,869 285 Updated Dec 23, 2025

SkyRL: A Modular Full-stack RL Library for LLMs

Python 2,151 403 Updated Aug 14, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,362 305 Updated Aug 14, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,920 1,137 Updated Aug 14, 2026
Next