Skip to content
View aryamanan's full-sized avatar
💭
I may be slow to respond.
💭
I may be slow to respond.

Block or report aryamanan

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

pytorch like baby neural net and reinforcement learning algorithms with scratch implementations

Python 8 1 Updated Aug 18, 2025

Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!

Python 10,603 972 Updated Aug 19, 2026

UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation

Python 889 30 Updated Dec 23, 2025

Master classic RL, deep RL, distributional RL, inverse RL, and more using OpenAI Gym and TensorFlow with extensive Math

Jupyter Notebook 476 141 Updated Apr 1, 2021

FULL Augment Code, Claude Code, Cluely, CodeBuddy, Comet, Cursor, Devin AI, Junie, Kiro, Leap.new, Lovable, Manus, NotionAI, Orchids.app, Perplexity, Poke, Qoder, Replit, Same.dev, Trae, Traycer AI…

142,938 34,844 Updated Aug 11, 2026

From-scratch implementation of DPO / SFT on GPT-2.

Python 8 1 Updated Dec 14, 2024

A chatbot using the RAG pipeline built with Django, Next.js 14, WebSockets, and powered by the LangChain framework. It leverages the llama2 model for processing user queries and generating responses.

TypeScript 9 4 Updated Jul 3, 2025

Repo for the Deep Reinforcement Learning Nanodegree program

Jupyter Notebook 5,175 2,373 Updated Jul 8, 2026

This is the official repository for The Hundred-Page Language Models Book by Andriy Burkov

Jupyter Notebook 2,176 367 Updated Feb 8, 2026

A roadmap for "generative AI" learning resources

CSS 309 37 Updated Feb 2, 2026

Question paper of courses taught at IISC as part of MTech AI curriculum

120 18 Updated Dec 1, 2024

🐭 A tiny single-file implementation of Group Relative Policy Optimization (GRPO) as introduced by the DeepSeekMath paper

Python 43 3 Updated Jun 28, 2025

One click away from a locally downloaded, fine-tuned model, hosted on hugging face, with inference built in. In two hours.

Jupyter Notebook 24 4 Updated Aug 13, 2026

PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.

Python 13,710 2,171 Updated Aug 17, 2026

FinRL®: Financial Reinforcement Learning. 🔥

Jupyter Notebook 16,045 3,470 Updated Jul 13, 2026

Minimal reproduction of DeepSeek R1-Zero

Python 13,224 1,578 Updated Feb 27, 2026

Fully open reproduction of DeepSeek-R1

Python 26,438 2,446 Updated Apr 2, 2026

100 days of building GPU kernels!

Cuda 629 78 Updated Apr 27, 2025

A curated list of reinforcement learning with human feedback resources (continually updated)

4,422 258 Updated May 20, 2026

Reference implementation for DPO (Direct Preference Optimization)

Python 2,907 236 Updated Aug 11, 2024

Code for NeurIPS 2024 paper - The GAN is dead; long live the GAN! A Modern Baseline GAN - by Huang et al.

Python 870 46 Updated Jan 23, 2025

PyTorch implementations of Generative Adversarial Networks.

Python 17,446 4,064 Updated Jun 18, 2024

Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities

Python 22,193 2,705 Updated Jan 23, 2026

First-principle implementations of groundbreaking AI algorithms using a wide range of deep learning frameworks, accompanied by supporting research papers and demos.

Jupyter Notebook 183 12 Updated Jul 30, 2026

FLUX, Stable Diffusion, SDXL, SD3, LoRA, Fine Tuning, DreamBooth, Training, Automatic1111, Forge WebUI, SwarmUI, DeepFake, TTS, Animation, Text To Video, Tutorials, Guides, Lectures, Courses, Comfy…

HTML 2,757 371 Updated Aug 14, 2026

A nanoGPT pipeline packed in a spreadsheet

2,159 128 Updated Jun 17, 2024

The most cited deep learning papers

TeX 26,178 4,417 Updated Jan 18, 2024

Deep Learning papers reading roadmap for anyone who are eager to learn this amazing tech!

Python 39,547 7,267 Updated Nov 27, 2022

just me trying to implement deep learning concepts in code

Jupyter Notebook 232 15 Updated Nov 8, 2025
Next