Skip to content
View aatkinson's full-sized avatar
👨‍🔬
...
👨‍🔬
...

Organizations

@Maluuba

Block or report aatkinson

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A guidance language for controlling large language models.

Jupyter Notebook 21,713 1,200 Updated May 21, 2026

Mirage Persistent Kernel: Compiling LLMs into a MegaKernel

Cuda 2,423 238 Updated Aug 4, 2026

A simplified implementation for experimenting with RLVR on GSM8K, This repository provides a starting point for exploring reasoning.

Python 172 16 Updated Feb 6, 2025

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,099 9,069 Updated Aug 13, 2026

[ICML 2024 Spotlight] Differentially Private Synthetic Data via Foundation Model APIs 2: Text

Python 61 17 Updated Jan 11, 2025

Helpful tools and examples for working with flex-attention

Python 1,225 77 Updated Aug 14, 2026

Official repository for ACL 2025 paper "ProcessBench: Identifying Process Errors in Mathematical Reasoning"

Python 192 18 Updated May 20, 2025

A comprehensive Rust translation of the code from Sebastian Raschka's Build an LLM from Scratch book.

Rust 330 39 Updated Aug 14, 2026

Pytorch optimiser for training ANNs with exponentiated gradient desent

Jupyter Notebook 20 3 Updated Mar 20, 2025

Democratizing Reinforcement Learning for LLMs

Python 5,784 605 Updated Aug 14, 2026

Prepare for DeekSeek R1 inference: Benchmark CPU, DRAM, SSD, iGPU, GPU, ... with efficient code.

C 73 2 Updated Feb 2, 2025

Simple RL training for reasoning

Python 3,869 285 Updated Dec 23, 2025

AutoAWQ implements the AWQ algorithm for 4-bit quantization with a 2x speedup during inference. Documentation:

Python 2,349 307 Updated May 11, 2025

An ML Systems Onboarding list

1,114 44 Updated Feb 19, 2026

The official repository of Quamba1 [ICLR 2025] & Quamba2 [ICML 2025]

Python 70 16 Updated Jun 19, 2025

Large Reasoning Models

Python 803 47 Updated Dec 3, 2024

Demonstrations of Loss of Plasticity and Implementation of Continual Backpropagation

Python 389 84 Updated Jul 14, 2026

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downs…

Python 3,442 545 Updated Aug 14, 2026

Hackable and optimized Transformers building blocks, supporting a composable construction.

Python 10,546 781 Updated Aug 7, 2026

Code for the paper "The Impact of Positional Encoding on Length Generalization in Transformers", NeurIPS 2023

Python 139 7 Updated Apr 30, 2024

The calflops is designed to calculate FLOPs、MACs and Parameters in all various neural networks, such as Linear、 CNN、 RNN、 GCN、Transformer(Bert、LlaMA etc Large Language Model)

Python 945 44 Updated Jun 27, 2024

A bibliography and survey of the papers surrounding o1

TeX 1,215 50 Updated Jul 7, 2026

Accelerating your LLM training to full speed! Made with ❤️ by ServiceNow Research

Python 331 49 Updated Jul 20, 2026

A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance…

Python 3,491 803 Updated Aug 14, 2026

Never use print for debugging again

Python 16,579 961 Updated Jun 8, 2026

📰 Must-read papers on KV Cache Compression (constantly updating 🤗).

733 28 Updated Apr 15, 2026

FlashInfer: Kernel Library for LLM Serving

Python 6,161 1,277 Updated Aug 14, 2026

Profiling and inspecting memory in pytorch

Python 1,078 40 Updated Aug 12, 2026

Efficient Triton Kernels for LLM Training

Python 6,568 581 Updated Aug 14, 2026
Next