Skip to content
View seolhokim's full-sized avatar
🌴
On vacation
🌴
On vacation

Block or report seolhokim

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

An agentic skills framework & software development methodology that works.

Shell 269,301 24,056 Updated Aug 8, 2026

A developer vanished. The clues are in the code. Only Claude Code can find them.

TypeScript 8 1 Updated Mar 5, 2026

Python Backtesting library for trading strategies

Python 22,771 5,230 Updated Aug 19, 2024

A TTS model capable of generating ultra-realistic dialogue in one pass.

Python 19,368 1,690 Updated Nov 19, 2025

Code and datasets for "Character-LLM: A Trainable Agent for Role-Playing"

Python 642 49 Updated Oct 29, 2024

AllenAI's post-training codebase

Python 3,820 573 Updated Aug 8, 2026

OpenChat: Advancing Open-source Language Models with Imperfect Data

Python 5,486 432 Updated Sep 13, 2024

Accurate answers and instant citations for your documents.

Python 1,635 789 Updated May 29, 2024

Offline, privacy-first grammar checker. Fast, open-source, Rust-powered

Rust 14,261 553 Updated Aug 8, 2026

NanoGPT (124M) in 90 seconds

Python 5,649 857 Updated Aug 2, 2026

DIAMOND (DIffusion As a Model Of eNvironment Dreams) is a reinforcement learning agent trained in a diffusion world model. NeurIPS 2024 Spotlight.

Python 2,089 161 Updated Dec 6, 2024

Code for "Learning to Model the World with Language." ICML 2024 Oral.

Python 421 31 Updated Jan 7, 2026

🤖 AgentVerse 🪐 is designed to facilitate the deployment of multiple LLM-based agents in various applications, which primarily provides two frameworks: task-solving and simulation

JavaScript 5,099 510 Updated Sep 9, 2024

IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures

Python 6 Updated Apr 11, 2024

[ECCV'24] SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic

Python 210 12 Updated Jul 14, 2025

A curated list of awesome exploration RL resources (continually updated)

720 26 Updated May 21, 2026

METRA: Scalable Unsupervised RL with Metric-Aware Abstraction (ICLR 2024)

Python 96 14 Updated Oct 15, 2023

Basic constrained RL agents used in experiments for the "Benchmarking Safe Exploration in Deep Reinforcement Learning" paper.

Python 464 111 Updated Apr 2, 2023

A continually updated list of literature on Reinforcement Learning from AI Feedback (RLAIF)

205 7 Updated Aug 6, 2025

This repo is built to facilitate the training and analysis of autoregressive transformers on maze-solving tasks.

Jupyter Notebook 35 7 Updated Oct 28, 2025

Gymnasium extension for DarkSouls III, Elden Ring, and other Souls games

Python 156 17 Updated Aug 8, 2026

High-quality single-file implementations of SOTA Offline and Offline-to-Online RL algorithms: AWAC, BC, CQL, DT, EDAC, IQL, SAC-N, TD3+BC, LB-SAC, SPOT, Cal-QL, ReBRAC

Python 654 39 Updated Feb 10, 2024

PyTorch implementation of Contrastive Learning methods

Python 1,993 181 Updated Oct 4, 2023

[NeurIPS 2023 Spotlight] LightZero: A Unified Benchmark for Monte Carlo Tree Search in General Sequential Decision Scenarios (awesome MCTS)

Python 1,630 197 Updated Aug 8, 2026
Python 741 79 Updated Jun 20, 2023

Initiative to read research papers

180 37 Updated Dec 30, 2023

Resources on various topics being worked on at IvLabs

356 63 Updated Nov 17, 2023

MetaDrive: Lightweight driving simulator for everyone

Python 1,231 197 Updated Aug 15, 2025
Next