Skip to content
View jpli02's full-sized avatar
🎯
Focusing
🎯
Focusing
  • UIUC CS
  • Earth
  • 11:18 (UTC -06:00)

Block or report jpli02

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

JavaScript 240,247 36,461 Updated Aug 15, 2026

This repo is meant to serve as a guide for Machine Learning/AI technical interviews.

Jupyter Notebook 8,764 1,525 Updated Jul 30, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,370 81,207 Updated Aug 15, 2026

📊 An infographics generator with 30+ plugins and 300+ options to display stats about your GitHub account and render them as SVG, Markdown, PDF or JSON!

JavaScript 17,060 2,255 Updated May 29, 2026

RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI

Python 4,544 656 Updated Aug 15, 2026

dLLM: Simple Diffusion Language Modeling

Python 2,668 281 Updated Jul 17, 2026

A unified inference and post-training framework for accelerated video generation.

Python 3,943 402 Updated Aug 15, 2026

Flexible and Pluggable Serving Engine for Diffusion LLMs

Python 151 18 Updated Jul 13, 2026

An interface library for RL post training with environments.

Python 2,502 427 Updated Aug 13, 2026

A Lightweight LLM Post-Training Library

Python 2,406 332 Updated Aug 15, 2026

Multi-Turn RL Training System with AgentTrainer for Language Model Game Reinforcement Learning

Python 66 12 Updated Dec 18, 2025

Training Large Language Model to Reason in a Continuous Latent Space

Python 1,684 187 Updated Jul 2, 2026

An open-source AI coding agent that lives in your terminal.

TypeScript 27,038 2,863 Updated Aug 15, 2026

A benchmark for LLMs on complicated tasks in the terminal

Python 2,533 563 Updated Jul 11, 2026

[NeurIPS 2025 Spotlight] Reasoning Environments for Reinforcement Learning with Verifiable Rewards

Python 1,485 123 Updated Apr 17, 2026

Checkpoint-engine is a simple middleware to update model weights in LLM inference engines

Python 996 102 Updated Aug 12, 2026

Our library for RL environments + evals

Python 4,518 646 Updated Aug 15, 2026
Python 652 67 Updated Aug 28, 2025

Implementation for FP8/INT8 Rollout for RL training without performence drop.

Python 309 23 Updated Nov 7, 2025

Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL

Python 5,296 479 Updated Nov 13, 2025

[NeurIPS 2025] Simple extension on vLLM to help you speed up reasoning model without training.

Python 232 33 Updated May 31, 2025
Python 890 52 Updated Sep 15, 2025

gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI

Python 20,309 2,137 Updated Jul 24, 2026

The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

Python 5,667 574 Updated Aug 15, 2026

Simple RL training for reasoning

Python 3,869 285 Updated Dec 23, 2025

SkyRL: A Modular Full-stack RL Library for LLMs

Python 2,153 403 Updated Aug 14, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,362 305 Updated Aug 15, 2026

slime is an LLM post-training framework for RL Scaling.

Python 8,032 1,147 Updated Aug 14, 2026
Next