Skip to content
View Dhawgupta's full-sized avatar

Highlights

  • Pro

Organizations

@rlai-lab

Block or report Dhawgupta

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

A collection of 100+ specialized Claude Code subagents covering a wide range of development use cases

Shell 24,277 2,817 Updated Aug 12, 2026

A curated list for Self-Improvement in Foundation Model Based Agentic Systems.

TeX 360 41 Updated Aug 12, 2026

NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

Jupyter Notebook 11,488 824 Updated Aug 13, 2026
Python 1 Updated Aug 4, 2026

AI agents running research on single-GPU nanochat training automatically

Python 93,807 13,304 Updated Mar 26, 2026

Fully autonomous & self-evolving research from idea to paper. Chat an Idea. Get a Paper. 🦞

Python 14,018 1,638 Updated Jul 13, 2026

Open-source implementation of AlphaEvolve

Python 6,958 1,114 Updated Jul 18, 2026

A Survey of Reinforcement Learning for Large Reasoning Models

TeX 2,479 132 Updated Aug 1, 2026

Leetcode for Pytorch

Jupyter Notebook 2,401 301 Updated Jun 14, 2026

The best ChatGPT that $100 can buy.

Python 57,174 7,921 Updated Aug 2, 2026

Fine-tune LLM agents with online reinforcement learning

Python 1,255 65 Updated Mar 19, 2024

NumPy+Jax with named axes and an uncompromising attitude

Jupyter Notebook 23 1 Updated Mar 4, 2025

An open-source alternative to OpenAI and Gemini's deep research.

TypeScript 881 112 Updated Feb 12, 2025

A template for writing academic papers in Markdown.

Makefile 10 Updated Mar 21, 2025

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 88,990 20,652 Updated Aug 13, 2026

Elegant easy-to-use neural networks + scientific computing in JAX. https://docs.kidger.site/equinox/

Python 2,947 210 Updated Aug 10, 2026

Online Goal-Conditioned Reinforcement Learning in JAX. ICLR 2025 Spotlight.

Python 275 52 Updated Aug 7, 2026

🏛️A research-friendly codebase for fast experimentation of single-agent reinforcement learning in JAX • End-to-End JAX RL

Python 418 48 Updated Mar 18, 2026

Really Fast End-to-End Jax RL Implementations

Python 1,095 87 Updated Sep 9, 2024

PDF references add-on for Zotero.

JavaScript 2,798 83 Updated Mar 27, 2026

Implements QuickLook in Zotero

JavaScript 62 3 Updated Jun 13, 2022

Implements QuickLook in Zotero

JavaScript 787 82 Updated Jun 16, 2023

Scripts to build a trimmed-down Windows 11 image.

PowerShell 19,357 1,486 Updated Sep 12, 2025

Safe Reinforcement Learning with Natural Language Constraints

17 1 Updated Oct 24, 2021

High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)

Python 10,264 1,152 Updated Apr 20, 2026
Python 4 2 Updated Aug 1, 2020

PyTorch implementation of Advantage Actor-Critic (A2C), Asynchronous Advantage Option-Critic (A2OC), Proximal Policy Optimization (PPO) and Scalable trust-region method for deep reinforcement learn…

Jupyter Notebook 8 2 Updated Oct 7, 2018
Python 47 4 Updated Feb 8, 2024
Jupyter Notebook 1 Updated May 1, 2024
Next