Skip to content
View veds12's full-sized avatar
😀
Meta Learning
😀
Meta Learning

Highlights

  • Pro

Organizations

@mila-iqia @SforAiDl

Block or report veds12

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A simple reference implementation of the single-worker MuLoCo optimizer in Jax & PyTorch. MuLoCo-1 has been shown to outperfrom Muon and have larger critical batch sizes.

Python 32 Updated Feb 27, 2026

Reproduction of Hilbert Prover

Python 6 Updated Nov 5, 2025

Post-training with Tinker

Python 4,018 509 Updated Aug 13, 2026

Automatically generating mathematical questions through a multi-agent system that the original LLM cannot solve.

Python 8 Updated Jul 5, 2025

Simple RL training for reasoning

Python 3,869 285 Updated Dec 23, 2025
Python 1,177 58 Updated Jan 10, 2026

GFlowNet library specialized for graph & molecular data

Python 295 54 Updated May 21, 2026

📋 A list of open LLMs available for commercial use.

12,848 985 Updated Feb 13, 2025
Python 17 2 Updated May 1, 2023

An offline deep reinforcement learning library

Python 1,676 267 Updated Sep 10, 2025

Generative Flow Networks

Python 685 80 Updated Feb 28, 2023

A playbook for systematically maximizing the performance of deep learning models.

30,282 2,420 Updated Jun 18, 2024

Implementation of 🦩 Flamingo, state-of-the-art few-shot visual question answering attention net out of Deepmind, in Pytorch

Python 1,269 66 Updated Oct 18, 2022

A modular, easy to extend GFlowNet library

Jupyter Notebook 313 57 Updated Aug 13, 2026

Implementation of Memorizing Transformers (ICLR 2022), attention net augmented with indexing and retrieval of memories using approximate nearest neighbors, in Pytorch

Python 646 49 Updated Jul 17, 2023

This repository contains code for the paper "Stateful Active Facilitator: Coordination and Environmental Heterogeneity in Cooperative Multi-Agent Reinforcement Learning". https://arxiv.org/abs/2210…

Python 6 6 Updated Apr 27, 2023

A curated list of resources about generative flow networks (GFlowNets).

503 36 Updated Oct 1, 2024

Transformers are Sample-Efficient World Models. ICLR 2023, notable top 5%.

Python 898 92 Updated Oct 14, 2024

Pytorch reimplementation of the Vision Transformer (An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale)

Jupyter Notebook 2,161 398 Updated Jun 7, 2022

A collection of meta-learning algorithms in Jax

Python 24 3 Updated Sep 3, 2022

Rainbow: Combining Improvements in Deep Reinforcement Learning

Python 1,671 294 Updated Jan 13, 2022

Official release of CompoSuite, a compositional RL benchmark

Python 51 4 Updated Jan 27, 2024

A PyTorch implementation of Perceiver, Perceiver IO and Perceiver AR with PyTorch Lightning scripts for distributed training

Python 535 46 Updated Jan 2, 2024

BabyAI platform. A testbed for training agents to understand and execute language commands.

Python 766 154 Updated Oct 1, 2023

Change Python code while it's running without losing state

Python 1,129 33 Updated Jun 15, 2024

GPU Accelerated t-SNE for CUDA with Python bindings

Cuda 1,940 138 Updated Jul 22, 2026

A suite of test scenarios for multi-agent reinforcement learning.

Python 862 157 Updated Aug 5, 2026

Code for "Data-Efficient Reinforcement Learning with Self-Predictive Representations"

Python 167 31 Updated Dec 21, 2021
Next