Skip to content
View kykim0's full-sized avatar

Organizations

@sisl @JuliaPOMDP @StanfordVL

Block or report kykim0

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

A curated list for Self-Improvement in Foundation Model Based Agentic Systems.

TeX 217 23 Updated Jul 25, 2026

Master programming by recreating your favorite technologies from scratch.

Markdown 531,528 50,277 Updated Jul 14, 2026

The end of web parsing. The beginning of scalable pixel-native search. link: https://pixelrag.ai/

Python 7,181 599 Updated Jul 16, 2026

Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning

Python 1,571 105 Updated Jul 21, 2026

An Open-Source Large-Scale Reinforcement Learning Project for Search Agents

Python 602 39 Updated Nov 26, 2025

MiroThinker is a deep research agent optimized for complex research and prediction tasks. Our latest models, MiroThinker-1.7, achieves 74.0 and 75.3 on the BrowseComp and BrowseComp Zh, respectively.

Python 8,349 645 Updated Jul 6, 2026

A Tree Search Library with Flexible API for LLM Inference-Time Scaling

Python 557 75 Updated Feb 5, 2026
Python 117 20 Updated Jun 30, 2025

OpenClaw-RL: Train any agent simply by talking

Python 5,606 607 Updated May 23, 2026

The official repository of "A Comprehensive Survey on Reinforcement Learning-based Agentic Search: Foundations, Roles, Optimizations, Evaluations, and Applications".

279 10 Updated Jul 21, 2026

Self-referential self-improving agents that can optimize for any computable task

Python 2,651 346 Updated May 9, 2026

Research on Coding Agents

12,177 19,644 Updated Apr 1, 2026

Curated academic CV templates and guidelines for PhD students, researchers, and faculty job applicants.

TeX 1,201 136 Updated Apr 1, 2026

AI agents running research on single-GPU nanochat training automatically

Python 92,033 13,144 Updated Mar 26, 2026

LLM Chess - evaluating Large Language Models' reasoning and instruction-following abilities by simulating chess games

Python 112 10 Updated Jul 25, 2026

A collection of various llm pruning implementations, training code for GPUs & TPUs, and evaluation script.

Python 69 8 Updated Apr 20, 2026

CATArena is an engineering-level tournament evaluation platform for Large Language Model-driven code agents (LLM-driven code agents), based on an iterative competitive peer learning framework.

Python 67 10 Updated Dec 25, 2025

"AI-Trader: 100% Fully-Automated Agent-Native Trading"

Python 21,030 3,202 Updated Jun 11, 2026

Synthetic data curation for post-training and structured data extraction

Python 1,704 143 Updated Jul 22, 2026

Benchmark LLM reasoning capability by solving chess puzzles.

Python 91 6 Updated Apr 26, 2025

World model reasoning RL for multi-turn VLM agents

Python 488 61 Updated Jul 23, 2026

Harsh Jhamtani*, Varun Gangal*, Eduard Hovy, Graham Neubig, Taylor Berg-Kirkpatrick. Learning to Generate Move-by-Move Commentary for Chess Games from Large-Scale Social Forum Data. ACL 2018

OpenEdge ABL 48 11 Updated Jul 21, 2022

Open source neural network chess engine with GPU acceleration and broad hardware support.

C++ 3,160 600 Updated May 5, 2026

A Text-Based Environment for Interactive Debugging

Python 301 41 Updated Jul 20, 2026

This is the official GitHub repository for our survey paper "Beyond Single-Turn: A Survey on Multi-Turn Interactions with Large Language Models".

Python 201 6 Updated Jul 11, 2026

Fully open reproduction of DeepSeek-R1

Python 26,415 2,446 Updated Apr 2, 2026

[ICLR 2026] Learning to Reason without External Rewards

Python 418 44 Updated Jan 26, 2026

A library for generative social simulation

Python 1,576 346 Updated Jul 22, 2026

AI paper trading project inspired by nof1 Alpha Arena, using cctx for quotation.

Python 596 146 Updated Nov 21, 2025

Procgen Benchmark: Procedurally-Generated Game-Like Gym-Environments

C++ 1,177 222 Updated Mar 27, 2026
Next