Skip to content
View jonhue's full-sized avatar

Organizations

@who-wrote-that @tony-lang @lasgroup

Block or report jonhue

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A general framework for strategically scaling evaluation-driven discovery loops, discovering state-of-the-art solutions on 21 open-ended problems.

C++ 163 15 Updated Jul 30, 2026

Reimplementation of TTT-Discover without the Tinker dependency

Python 14 1 Updated Jun 21, 2026

CaOPD: Calibration-Aware On-Policy Distillation

Python 15 3 Updated Jun 2, 2026

Lightweight coding agent that runs in your terminal

Rust 104,974 15,894 Updated Aug 10, 2026

Bash Line Editor―a line editor written in pure Bash with syntax highlighting, auto suggestions, vim modes, etc. for Bash interactive sessions.

Shell 4,594 140 Updated Jul 11, 2026

AI-Driven Scientific and Algorithmic Discovery

Python 595 87 Updated Jun 14, 2026

Framework for evaluating and improving agents

Python 4,046 1,503 Updated Aug 9, 2026

OpenClaw-RL: Train any agent simply by talking

Python 5,627 608 Updated May 23, 2026

The official codebase for "Experiential Reinforcement Learning" - https://arxiv.org/pdf/2602.13949v1

Python 76 8 Updated Jul 2, 2026
Python 35 6 Updated Apr 3, 2026

PyTorch library for Active Fine-Tuning

Python 100 9 Updated Sep 27, 2025

Aligning Language Models from User Interactions via Self-Distillation

Python 29 6 Updated Mar 31, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 385,706 81,069 Updated Aug 10, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,822 1,130 Updated Aug 7, 2026

Continual Learning as a Service

Python 59 6 Updated Mar 7, 2026

pip install continualcode

Python 43 4 Updated Feb 10, 2026

Reinforcement Learning via Self-Distillation (SDPO)

Python 1,044 123 Updated Jul 1, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 88,614 20,486 Updated Aug 10, 2026

Official JAX implementation of End-to-End Test-Time Training for Long Context

Python 631 47 Updated Feb 15, 2026

Hydra is a framework for elegantly configuring complex applications

Python 10,586 944 Updated Aug 9, 2026

Storing long contexts in tiny caches with self-study

Python 309 41 Updated Mar 23, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,886 4,365 Updated Aug 10, 2026

Test-Time Curricula for Targeted RL

Python 10 Updated Oct 7, 2025

Specialization after Generalization

Python 8 Updated Mar 31, 2026

The official repository of paper "Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models''

Python 113 5 Updated Aug 15, 2025

[ICLR 2026] On the Generalization of SFT: A Reinforcement Learning Perspective with Reward Rectification.

Python 1,095 26 Updated Aug 1, 2026

Official repository for "Test-time Offline Reinforcement Learning on Goal-related Experience" (Preprint)

Python 7 Updated Jul 29, 2025

Hierarchical Reasoning Model Official Release

Python 12,615 1,827 Updated Mar 31, 2026

The codebase and some introductions of FineMed.

Jupyter Notebook 31 1 Updated Sep 11, 2025
Next