Skip to content
View jkkjjj's full-sized avatar
🎀
JK lover
🎀
JK lover

Block or report jkkjjj

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[ICML 2026] Milestone-Guided Policy Learning for Long-Horizon Language Agents

Python 42 2 Updated May 29, 2026
TypeScript 45 3 Updated Aug 6, 2026

The papers and code related to deep generative models for material discovery.

42 4 Updated Aug 10, 2026
Python 75 4 Updated Jun 10, 2025

An active paper-reading skill that reconstructs author reasoning, explains methods mechanistically, stress-tests assumptions, and generates follow-up research ideas.

456 23 Updated Aug 3, 2026
Python 10 Updated Jun 3, 2026

ICML 2026 autonomous AI agent for end-to-end spatial proteomics analysis, with SP-Bench for agentic multiplexed-imaging workflows.

Python 174 22 Updated Jul 13, 2026

Framework for evaluating and improving agents

Python 4,215 1,556 Updated Aug 13, 2026

【ICML2026 Spotlight】 T2PO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement Learning

Python 54 Updated May 27, 2026

Scholar All-In-One: A research infrastructure for AI agents

Python 559 75 Updated Jul 29, 2026

Collection of Summer, Fall, Spring 2027 tech internships!

8,809 305 Updated Aug 6, 2026

A collection of full time roles in SWE, Quant, and PM for new grads.

17,661 1,302 Updated Aug 14, 2026
Python 143 9 Updated Jun 18, 2026

[CVPR 25] A framework named B^2-DiffuRL for RL-based diffusion model fine-tuning.

Python 57 4 Updated Mar 31, 2025

The offical repo for "Parallel-Probe: Towards Efficient Parallel Thinking via 2D Probing"

Python 19 1 Updated Feb 3, 2026
Python 55 14 Updated May 3, 2026
Python 22 4 Updated Jun 2, 2026

This codebase is to reproduce the results of the paper "Grounded Test-Time Adaptation for LLM Agents".

Python 18 1 Updated Mar 4, 2026
Python 3 Updated May 13, 2026

Offical implementation of "Life-Harness"

Python 214 14 Updated Jul 14, 2026

Skill package for ML/CV/NLP paper writing, curated and adapted from Prof. Peng Sida's open notes for Codex, Claude Code, and Gemini.

6,023 293 Updated Jun 23, 2026

A collection of reinforcement learning environments for simple web interaction tasks

HTML 396 58 Updated Aug 13, 2026

​TextWorld is a sandbox learning environment for the training and evaluation of reinforcement learning (RL) agents on text-based games.

Jupyter Notebook 1,438 203 Updated Jan 30, 2026
C++ 7 Updated Mar 3, 2026

Evolutionary Generation of Multi-Agent Systems; Yuntong Hu, Yuting Zhang, Matthew Trager, Yi Zhang, Shuo Yang, Wei Xia, Stefano Soatto, ICML 2026

Python 7 3 Updated May 29, 2026

CATArena is an engineering-level tournament evaluation platform for Large Language Model-driven code agents (LLM-driven code agents), based on an iterative competitive peer learning framework.

Python 67 10 Updated Dec 25, 2025

[ICML 2026] Principle-Evolvable Scientific Discovery via Uncertainty Minimization

Python 34 7 Updated Jul 4, 2026

[ICML'26] MemEvolve & EvolveLab

Python 259 28 Updated May 5, 2026

(ICML 2026) Self‑Evolving LLM Agents through an Experience‑Driven Lifecycle

Python 106 7 Updated May 8, 2026
Next