Skip to content
View jmhb0's full-sized avatar

Block or report jmhb0

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Procedural data generators for verifiable reasoning, synthetic pretraining, post-training, evaluation, and RL.

Python 48 4 Updated Aug 19, 2026

A skill to stop your coding agent from burying the answer. ADHD-friendly output.

Python 22,139 1,401 Updated Aug 19, 2026

Implementation of the paper: Beyond Correlation: The impact of human uncertainty in measuring the effectiveness of automatic evaluation and LLM-as-a-judge

Python 13 4 Updated Feb 17, 2025

MORPH: PDE Foundation Models with Arbitrary Data Modality

Jupyter Notebook 28 6 Updated Jun 18, 2026

This repo is meant to serve as a guide for Machine Learning/AI technical interviews.

Jupyter Notebook 9,364 1,647 Updated Aug 19, 2026

Official Repo of "$X$-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding"

Shell 34 1 Updated Jun 18, 2026

Code for Negation Neglect

Python 16 5 Updated May 22, 2026
Python 57 3 Updated Jun 8, 2026

[CVPR 2026 Findings] V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think

Python 57 2 Updated Apr 28, 2026

tLLM is an test-time training extension of vLLM

Python 46 Updated Apr 26, 2026

Code, Data, and Model Outputs for the paper "This Treatment Works, Right? Evaluating LLM Sensitivity to Patient Question Framing in Medical QA"

HTML 1 Updated Jun 1, 2026

Elucidating the Design Space of Flow Matching for Cellular Microscopy

Python 5 Updated May 19, 2026

[ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.

Python 39 1 Updated Mar 5, 2026

[ICML 2025] Teaching Language Models to Critique via Reinforcement Learning

Python 126 7 Updated May 6, 2025

Unison file synchronizer

OCaml 5,449 271 Updated Aug 6, 2026

PyTorch building blocks for the OLMo ecosystem

Python 1,475 304 Updated Aug 19, 2026

Modeling, training, eval, and inference code for OLMo

Python 6,635 793 Updated Nov 24, 2025

Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours

Python 520 59 Updated Aug 13, 2026
1 Updated Feb 12, 2026
Python 14 1 Updated Feb 27, 2026

Implement a ChatGPT-like LLM in PyTorch from scratch, step by step

Jupyter Notebook 102,987 15,780 Updated Aug 10, 2026
Python 11 Updated Jun 25, 2025

A framework bridging cognitive science and LLM reasoning research to diagnose and improve how large language models reason, based on analysis of 192K model traces and 54 human think-aloud traces.

Python 43 9 Updated Nov 26, 2025
Jupyter Notebook 9 2 Updated May 6, 2026

Fully Open-source Multimodal Language Models for Science Discovery

Python 167 11 Updated Mar 20, 2026

[EACL 2026] PaperSearchQA. Data generation pipeline for QA over scientific papers, suitable for RL training search agents

Python 36 1 Updated Feb 4, 2026
Next