Skip to content
View klvntagoe's full-sized avatar

Organizations

@competemcgill

Block or report klvntagoe

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Algorithm powering the For You feed on X

Rust 26,958 4,589 Updated May 15, 2026

Textbook on reinforcement learning from human feedback

Python 2,271 245 Updated Aug 7, 2026

A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.

Python 3,513 474 Updated Aug 9, 2026

Code from the Deep Reinforcement Learning in Action book from Manning, Inc

Jupyter Notebook 852 349 Updated Apr 22, 2024
Python 41 7 Updated Jun 8, 2026

Waymo Open Dataset

Python 3,384 698 Updated Jan 8, 2026

A data-driven, fast driving simulator for multi-agent coordination under partial observability.

Python 299 34 Updated Jun 18, 2024
C++ 513 54 Updated Nov 3, 2025

TensorZero is an open-source LLMOps platform that unifies an LLM gateway, observability, evaluation, optimization, and experimentation.

Rust 11,724 961 Updated Jun 11, 2026

Deep Q-learning for playing tetris game

Python 531 117 Updated Apr 3, 2023

Deep reinforcement learning without experience replay, target networks, or batch updates.

Python 294 34 Updated Mar 18, 2025

Dopamine is a research framework for fast prototyping of reinforcement learning algorithms.

Jupyter Notebook 10,892 1,393 Updated Mar 24, 2026

A technical report on convolution arithmetic in the context of deep learning

TeX 14,695 2,309 Updated Jun 8, 2023

Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more

Python 36,125 3,726 Updated Aug 9, 2026

ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator

C++ 21,321 4,108 Updated Aug 9, 2026

AG2 (formerly AutoGen): The Open-Source AgentOS.Join us at: https://discord.gg/sNGSwQME3x

Python 4,847 697 Updated Aug 9, 2026

A programming framework for agentic AI

Python 60,330 9,088 Updated Apr 15, 2026

You like pytorch? You like micrograd? You love tinygrad! ❤️

Python 33,432 4,248 Updated Aug 9, 2026

MCTS project for Tetris

Python 348 35 Updated Oct 9, 2024

A deep reinforcement learning bot that plays tetris

Python 327 73 Updated Sep 1, 2024

🦁 A research-friendly codebase for fast experimentation of multi-agent reinforcement learning in JAX

Python 926 122 Updated May 26, 2026

Deep Reinforcement Learning for mobile robot navigation in ROS Gazebo simulator. Using Twin Delayed Deep Deterministic Policy Gradient (TD3) neural network, a robot learns to navigate to a random g…

Python 1,350 192 Updated Dec 13, 2025

A standard API for single-agent reinforcement learning environments, with popular reference environments and related utilities (formerly Gym)

Python 12,302 1,406 Updated Aug 5, 2026

A standard API for multi-agent reinforcement learning environments, with popular reference environments and related utilities

Python 3,489 515 Updated Aug 3, 2026

1 million FPS multi-agent driving simulator

Jupyter Notebook 613 87 Updated Dec 1, 2025

Explorer is a PyTorch reinforcement learning framework for exploring new ideas.

Python 98 14 Updated Jul 17, 2026

Pytorch implementation of convolutional neural network visualization techniques

Python 8,225 1,503 Updated Jan 1, 2025

A fully configurable Gymnasium compatible Tetris environment

Python 47 7 Updated Jun 1, 2026

Minimal and Clean Reinforcement Learning Examples

Python 3,660 737 Updated Jun 12, 2026
Next