Skip to content
View duterscmy's full-sized avatar
🎯
Focusing
🎯
Focusing
  • Beijing
  • 02:07 (UTC +01:00)

Block or report duterscmy

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A user-friendly & efficient knowledge distillation framework for LLMs, supporting off-policy, on-policy (OPD), cross-tokenizer, multimodal, and on-policy self-distillation.

Python 236 17 Updated Aug 13, 2026
Python 24 4 Updated May 17, 2026

The official implemention of "Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning"

Python 3 1 Updated May 31, 2026

Official PyTorch implementation of Dynamic-dLLM (ICLR 2026).

Python 6 1 Updated May 26, 2026
Python 4 Updated Feb 15, 2026

[ICLR 2026] Learning to Parallel: Accelerating Diffusion Large Language Models via Learnable Parallel Decoding

Python 34 2 Updated Jan 27, 2026

Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

Python 927 66 Updated Jun 29, 2026

A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models

666 21 Updated Aug 14, 2026

A curated collection of papers and resources on On-Policy Distillation for Large Language Models.

Python 513 10 Updated Aug 12, 2026
Python 19 1 Updated Jul 24, 2026

🏡 GitHub Pages template for personal academic homepage

HTML 691 406 Updated Jul 1, 2026
Python 33 2 Updated Aug 21, 2025

Official PyTorch implementation of the paper "Accelerating Diffusion Large Language Models with SlowFast Sampling: The Three Golden Principles" (Slow Fast Sampling).

Python 45 2 Updated Jul 18, 2025

Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"

Python 1,072 141 Updated May 30, 2026

Official implementation of "Diffusion Language Models Know the Answer Before Decoding"

Jupyter Notebook 60 1 Updated Apr 28, 2026

SDAR (Synergy of Diffusion and AutoRegression), a large diffusion language model(1.7B, 4B, 8B, 30B)

Python 483 33 Updated Jul 29, 2026

Automatically crawl arXiv papers daily and summarize them using AI. Illustrating them using GitHub Pages.

JavaScript 2,938 1,091 Updated Aug 14, 2026

Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends

Python 2,517 528 Updated Aug 11, 2026

[NeurIPS 2025] TTRL: Test-Time Reinforcement Learning

Python 1,110 82 Updated Apr 15, 2026

A project page template for academic papers. Demo at https://eliahuhorwitz.github.io/Academic-project-page-template/

JavaScript 5,149 1,148 Updated Sep 4, 2025

[ICML2026] Official Pytorch Implement for "Search or Accelerate: Confidence-Switched Position Beam Search for Diffusion Language Models"

Python 9 Updated Jul 6, 2026

[ICML 2026 Outstanding Paper] Minimalist RL for Diffusion LLMs. 89.1% on GSM8K.

Python 268 10 Updated Jul 6, 2026

[ICML 2026] d3LLM: Ultra-Fast Diffusion LLM 🚀

Python 149 10 Updated May 1, 2026

Fully open reproduction of DeepSeek-R1

Python 26,430 2,446 Updated Apr 2, 2026

The official GitHub repo for the survey paper "A Survey on Diffusion Language Models".

1,174 59 Updated May 29, 2026

dLLM: Simple Diffusion Language Modeling

Python 2,669 281 Updated Jul 17, 2026

[ICLR 2026] Official repository of "Beyond Fixed: Training-Free Variable-Length Denoising for Diffusion Large Language Models"

Python 174 7 Updated Feb 16, 2026

Official PyTorch implementation for "Large Language Diffusion Models"

Python 3,937 272 Updated Jul 15, 2026

Official PyTorch implementation of CD-MOE

Python 12 Updated Mar 18, 2026

Official Pytorch Implementation of "Outlier-weighed Layerwise Sampling for LLM Fine-tuning" by Pengxiang Li, Lu Yin, Xiaowei Gao, Shiwei Liu

Python 35 7 Updated Jun 3, 2025
Next