Skip to content
View ryyzn9's full-sized avatar
πŸ–€
moody
πŸ–€
moody

Block or report ryyzn9

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Combining Double Soft-Min Critics with Adaptive KL Thresholding for Alignmet

Python 1 1 Updated Feb 7, 2026

H100-Optimized Goal-Conditioned Contrastive Self-RL for ARC Puzzles

Python 2 Updated Jan 18, 2026

Physics-based language model: O(n log n) wave field attention with linear-wave content routing. V4.1 achieves PPL 543 on WikiText-2.

Python 21 7 Updated Mar 17, 2026

The best-benchmarked open-source AI memory system. And it's free.

Python 58,251 7,488 Updated Aug 8, 2026
Python 3 Updated Apr 7, 2026

qqr is an RL training framework for open-ended agents.

Python 1 Updated Feb 13, 2026

James' cookbook of evaluations and finetuning experiments

Python 32 5 Updated Feb 19, 2026

AuON( Alternative Unit-norm momentum-updates by Normal- ized nonlinear scaling), a linear-time optimizer that achieves remarkable perfor- mance at linear time without producing semi-orthogonal matr…

Python 4 Updated Jan 18, 2026

Quick illustration of how one can easily read books together with LLMs. It's great and I highly recommend it.

Python 3,827 487 Updated Nov 18, 2025

The official repository for AdaMuon

Python 39 4 Updated Aug 27, 2025

Bootstrapping ARC

Python 164 25 Updated Nov 20, 2024

This Repo Contains Script To Fine Tune Open Source Models Using Unsloth by using UI with simple click and progress

Python 12 8 Updated Oct 3, 2024

Materials for ConceptARC paper

120 9 Updated Feb 10, 2026

VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo

Python 2,135 247 Updated Aug 7, 2026

Classify the morphologies of distant galaxies with cnn-vit

Python 2 Updated Dec 10, 2023

[CVPR2025] Breaking the Low-Rank Dilemma of Linear Attention

Python 44 2 Updated Mar 11, 2025

Leo optimizer, variation of Muon that runs faster

Python 58 4 Updated Sep 6, 2025

GLU Attention provide nearly cost-free performance boost for transformers with a simple mechanism that applies Gated Linear Unit to the values in Attention.

Jupyter Notebook 9 2 Updated Jul 17, 2025

Awesome resources on normalizing flows.

Python 1,638 130 Updated Jul 31, 2026
JavaScript 4,318 1,938 Updated Jun 21, 2024

codes for R-Zero: Self-Evolving Reasoning LLM from Zero Data (https://www.arxiv.org/pdf/2508.05004)

Python 1 Updated Aug 8, 2025

Stella Nera is the first Maddness accelerator achieving 15x higher area efficiency (GMAC/s/mm^2) and 25x higher energy efficiency (TMAC/s/W) than direct MatMul accelerators in the same technology

Python 2 1 Updated Apr 24, 2024

Unison file synchronizer

OCaml 5,430 272 Updated Aug 6, 2026

Continuous Thought Machines, because thought takes time and reasoning is a process.

Python 2,016 306 Updated Dec 29, 2025

American Sign Language Recognition System that translates Signs into their respective alphabets in real time

Python 8 1 Updated Aug 1, 2022

πŸ€— LeRobot: Making AI for Robotics more accessible with end-to-end learning

Python 26,517 5,345 Updated Aug 9, 2026

Lecture notes on the RL series provided by Stanford.

Jupyter Notebook 1 2 Updated May 1, 2024

This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."

MATLAB 17,411 1,657 Updated Aug 5, 2026
Next