Skip to content
View amfarahmand's full-sized avatar

Highlights

  • Pro

Block or report amfarahmand

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)

Python 1 Updated Apr 1, 2026

Characteristic Value Iteration (CVI): frequency-domain reinforcement learning using characteristic functions on tabular MDPs

Jupyter Notebook 2 Updated Jan 29, 2026

[Adage Lab version] Official Code for "Relative Entropy Pathwise Policy Optimization"

Python 1 Updated Apr 25, 2026

Official code for the paper "Calibrated Value-Aware Model Learning with Stochastic Environment Models

Python 2 Updated Jun 13, 2025

[ICML'25] PANDAS: Improving Many-shot Jailbreaking via Positive Affirmation, Negative Demonstration, and Adaptive Sampling

Python 8 2 Updated Jun 13, 2025

Official Code for the paper: Deflated Dynamics Value Iteratioon

Python 1 Updated May 4, 2025

Official package for the MDOT-TNT algorithm for discrete optimal transport.

Python 6 3 Updated Mar 5, 2026

Code to reproduce the results in the NeurIPS 2023 paper Distributional Model Equivalence for Risk-Sensitive Reinforcement Learning.

Python 5 Updated Dec 13, 2023

[ECCV'24] Improving Adversarial Transferability via Model Alignment

Python 10 1 Updated Jul 25, 2024

An application of ideas from control theory to hopefully accelerate the dynamics of TD learning.

Python 4 Updated Sep 1, 2024

Experiments for policy-aware model learning

Jupyter Notebook 9 Updated Oct 22, 2020