Skip to content
View seolhokim's full-sized avatar
🌴
On vacation
🌴
On vacation

Block or report seolhokim

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

An agentic skills framework & software development methodology that works.

Shell 272,558 24,373 Updated Aug 13, 2026

A developer vanished. The clues are in the code. Only Claude Code can find them.

TypeScript 8 1 Updated Mar 5, 2026

Python Backtesting library for trading strategies

Python 22,855 5,241 Updated Aug 19, 2024

A TTS model capable of generating ultra-realistic dialogue in one pass.

Python 19,368 1,692 Updated Nov 19, 2025

Code and datasets for "Character-LLM: A Trainable Agent for Role-Playing"

Python 643 50 Updated Oct 29, 2024

AllenAI's post-training codebase

Python 3,828 576 Updated Aug 15, 2026

OpenChat: Advancing Open-source Language Models with Imperfect Data

Python 5,487 432 Updated Sep 13, 2024

Accurate answers and instant citations for your documents.

Python 1,634 787 Updated May 29, 2024

Offline, privacy-first grammar checker. Fast, open-source, Rust-powered

Rust 14,479 563 Updated Aug 16, 2026

NanoGPT (124M) in 90 seconds

Python 5,673 865 Updated Aug 9, 2026

DIAMOND (DIffusion As a Model Of eNvironment Dreams) is a reinforcement learning agent trained in a diffusion world model. NeurIPS 2024 Spotlight.

Python 2,089 163 Updated Dec 6, 2024

Code for "Learning to Model the World with Language." ICML 2024 Oral.

Python 421 31 Updated Jan 7, 2026

🤖 AgentVerse 🪐 is designed to facilitate the deployment of multiple LLM-based agents in various applications, which primarily provides two frameworks: task-solving and simulation

JavaScript 5,109 511 Updated Sep 9, 2024

IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures

Python 6 Updated Apr 11, 2024

[ECCV'24] SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic

Python 209 12 Updated Jul 14, 2025

A curated list of awesome exploration RL resources (continually updated)

722 26 Updated May 21, 2026

METRA: Scalable Unsupervised RL with Metric-Aware Abstraction (ICLR 2024)

Python 95 14 Updated Oct 15, 2023

Basic constrained RL agents used in experiments for the "Benchmarking Safe Exploration in Deep Reinforcement Learning" paper.

Python 464 111 Updated Apr 2, 2023

A continually updated list of literature on Reinforcement Learning from AI Feedback (RLAIF)

205 7 Updated Aug 6, 2025

This repo is built to facilitate the training and analysis of autoregressive transformers on maze-solving tasks.

Jupyter Notebook 35 7 Updated Oct 28, 2025

Gymnasium extension for DarkSouls III, Elden Ring, and other Souls games

Python 156 17 Updated Aug 9, 2026

High-quality single-file implementations of SOTA Offline and Offline-to-Online RL algorithms: AWAC, BC, CQL, DT, EDAC, IQL, SAC-N, TD3+BC, LB-SAC, SPOT, Cal-QL, ReBRAC

Python 657 39 Updated Feb 10, 2024

PyTorch implementation of Contrastive Learning methods

Python 1,993 180 Updated Oct 4, 2023

[NeurIPS 2023 Spotlight] LightZero: A Unified Benchmark for Monte Carlo Tree Search in General Sequential Decision Scenarios (awesome MCTS)

Python 1,632 197 Updated Aug 14, 2026
Python 741 79 Updated Jun 20, 2023

Initiative to read research papers

181 37 Updated Dec 30, 2023

Resources on various topics being worked on at IvLabs

356 63 Updated Nov 17, 2023

MetaDrive: Lightweight driving simulator for everyone

Python 1,232 197 Updated Aug 15, 2025
Next