Stars
DeepTeam is a framework to red team LLMs and AI agents.
slime is an LLM post-training framework for RL Scaling.
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
Repository hosting code for "Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative Recommendations" (https://arxiv.org/abs/2402.17152).
Sparse Transition Matrix-Accelerated Trie Index for Constrained Decoding (https://arxiv.org/abs/2602.22647)
Public quant internship repository, maintained by NUFT but available for everyone.
Understanding R1-Zero-Like Training: A Critical Perspective
A framework for few-shot evaluation of language models.
The guide to online assessments and interviews
Open-source implementation of AlphaEvolve
AI molecular design tool for de novo design, scaffold hopping, R-group replacement, linker design and molecule optimization.
Uncertainty quantification with PyTorch
Lightning-UQ-Box: Uncertainty Quantification for Neural Networks with PyTorch and Lightning
Official repository Flash Local Linear Attention
Examining how large language models (LLMs) perform across various synthetic regression tasks when given (input, output) examples in their context, without any parameter update
A regression-alike loss to improve numerical reasoning in language models - ICML 2025
🚀 MassGen is an open-source multi-agent scaling system that runs in your terminal, autonomously orchestrating frontier models and agents to collaborate, reason, and produce high-quality results. | …
official code for paper Probing the Decision Boundaries of In-context Learning in Large Language Models. https://arxiv.org/abs/2406.11233 [NeurIPS 2024]
First Open-Source Industry-Specific Model for Semiconductors
Official PyTorch implementation for "Large Language Diffusion Models"
MMaDA - Open-Sourced Multimodal Large Diffusion Language Models (dLLMs with block diffusion, mixed-CoT, unified RL)