Skip to content
View cw-yancey's full-sized avatar

Highlights

  • Pro

Block or report cw-yancey

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Reinforcement learning resources curated

9,917 1,945 Updated May 25, 2023

A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models

726 25 Updated Aug 15, 2026

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.

204,374 20,948 Updated Apr 20, 2026

A high-capacity ride-sharing simulator calibrated by real request datasets and road netwoeks

Python 21 6 Updated May 2, 2023

AI agents running research on single-GPU nanochat training automatically

Python 94,278 13,323 Updated Mar 26, 2026

The awesome collection of OpenClaw skills. 5,400+ skills filtered and categorized from the official OpenClaw Skills Registry.🦞

52,088 5,002 Updated Aug 20, 2026

The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

Python 5,682 579 Updated Aug 20, 2026

"AI-Trader: 100% Fully-Automated Agent-Native Trading"

Python 21,473 3,281 Updated Jun 11, 2026

The best ChatGPT that $100 can buy.

Python 57,356 7,972 Updated Aug 2, 2026

Tongyi Deep Research, the Leading Open-source Deep Research Agent

Python 19,856 1,511 Updated Feb 27, 2026

Code for Posterior Sampling for Deep Reinforcement Learning, ICML 2023

Python 28 6 Updated Mar 7, 2024

A project to assess the costs of flexible vehicle routing strategies

Jupyter Notebook 20 8 Updated Apr 24, 2023
Python 1 2 Updated Sep 17, 2024

Puffing up reinforcement learning

C 6,291 542 Updated Aug 20, 2026

Exploring the Roles of Large Language Models in Reshaping Transportation Systems: A Survey, Framework, and Roadmap

77 5 Updated Jul 15, 2025

OpenAI Baselines: high-quality implementations of reinforcement learning algorithms

Python 16,755 4,932 Updated Aug 1, 2024

Paper list of Reinforcement Learning (RL) applied on transportation

107 25 Updated Oct 28, 2021

An educational resource to help anyone learn deep reinforcement learning.

Python 11,905 2,465 Updated Aug 5, 2024

This repository contains a collection of resources and papers on Diffusion Models for RL, accompanying the paper "Diffusion Models for Reinforcement Learning: A Survey"

670 31 Updated Nov 29, 2024

A curated list of Diffusion Model in RL resources (continually updated)

1,633 78 Updated May 30, 2026

A final sanity checklist to help your CS paper get accepted, not desk rejected.

1,611 146 Updated May 25, 2026

Official implementation of "Graph Neural Network Reinforcement Learning for Autonomous Mobility-on-Demand

Python 84 19 Updated Apr 28, 2021

为OPC/中小微企业量身打造的自媒体获客智能体

Python 8,440 1,431 Updated Aug 15, 2026

Educational framework exploring ergonomic, lightweight multi-agent orchestration. Managed by OpenAI Solution team.

Python 21,909 2,329 Updated Apr 15, 2026

Meal Delivery Routing Problem Test Instance Library

Python 76 25 Updated Feb 28, 2018

Course to get into Large Language Models (LLMs) with roadmaps and Colab notebooks.

81,881 9,521 Updated Feb 5, 2026
Python 22 7 Updated Jun 23, 2026

[Lumina具身智能社区] 具身智能技术指南 Embodied-AI-Guide

15,542 994 Updated Aug 2, 2026
Next