Stars
[NeurIPS 2025 Spotlight] Reasoning Environments for Reinforcement Learning with Verifiable Rewards
The Entropy Mechanism of Reinforcement Learning for Large Language Model Reasoning.
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
Awesome LLM pruning papers all-in-one repository with integrating all useful resources and insights.
Official FUTO Keyboard Issue Tracker and Source Mirror of https://gitlab.futo.org/keyboard/latinime
An AI-powered research assistant that performs iterative, deep research on any topic by combining search engines, web scraping, and large language models. The goal of this repo is to provide the siโฆ
From Chain-of-Thought prompting to OpenAI o1 and DeepSeek-R1 ๐
Efficient Triton Kernels for LLM Training
LLMs can generate feedback on their work, use it to improve the output, and repeat this process iteratively.
Create web-based user interfaces with Python. The nice way.
Ultra-fast, low latency LLM prompt injection/jailbreak detection โ๏ธ
๐ python package to calculate readability statistics of a text object - paragraphs, sentences, articles.
A Python library for calculating a large variety of metrics from text
A unified evaluation framework for large language models
[NeurIPS 2023] MeZO: Fine-Tuning Language Models with Just Forward Passes. https://arxiv.org/abs/2305.17333
A guidance language for controlling large language models.
Accessible large language models via k-bit quantization for PyTorch.
Train transformer language models with reinforcement learning.
A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.
AirLLM 70B inference with single 4GB GPU
Automatically split your PyTorch models on multiple GPUs for training & inference
Simple, minimal implementation of the Mamba SSM in one file of PyTorch.