Skip to content
View duterscmy's full-sized avatar
🎯
Focusing
🎯
Focusing
  • Beijing
  • 23:02 (UTC +01:00)

Block or report duterscmy

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A user-friendly & efficient knowledge distillation framework for LLMs, supporting off-policy, on-policy (OPD), cross-tokenizer, multimodal, and on-policy self-distillation.

Python 236 16 Updated Aug 6, 2026
Python 22 3 Updated May 17, 2026

The official implemention of "Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning"

Python 3 1 Updated May 31, 2026

Official PyTorch implementation of Dynamic-dLLM (ICLR 2026).

Python 6 1 Updated May 26, 2026
Python 4 Updated Feb 15, 2026

[ICLR 2026] Learning to Parallel: Accelerating Diffusion Large Language Models via Learnable Parallel Decoding

Python 34 2 Updated Jan 27, 2026

Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

Python 911 63 Updated Jun 29, 2026

A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models

649 20 Updated Aug 9, 2026

A curated collection of papers and resources on On-Policy Distillation for Large Language Models.

Python 506 8 Updated Aug 9, 2026
Python 19 1 Updated Jul 24, 2026

🏡 GitHub Pages template for personal academic homepage

HTML 685 398 Updated Jul 1, 2026
Python 33 2 Updated Aug 21, 2025

Official PyTorch implementation of the paper "Accelerating Diffusion Large Language Models with SlowFast Sampling: The Three Golden Principles" (Slow Fast Sampling).

Python 43 2 Updated Jul 18, 2025

Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"

Python 1,070 140 Updated May 30, 2026

Official implementation of "Diffusion Language Models Know the Answer Before Decoding"

Jupyter Notebook 60 1 Updated Apr 28, 2026

SDAR (Synergy of Diffusion and AutoRegression), a large diffusion language model(1.7B, 4B, 8B, 30B)

Python 443 32 Updated Jul 29, 2026

Automatically crawl arXiv papers daily and summarize them using AI. Illustrating them using GitHub Pages.

JavaScript 2,934 1,089 Updated Aug 10, 2026

Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends

Python 2,514 525 Updated Aug 10, 2026

[NeurIPS 2025] TTRL: Test-Time Reinforcement Learning

Python 1,111 82 Updated Apr 15, 2026

A project page template for academic papers. Demo at https://eliahuhorwitz.github.io/Academic-project-page-template/

JavaScript 5,143 1,144 Updated Sep 4, 2025

[ICML2026] Official Pytorch Implement for "Search or Accelerate: Confidence-Switched Position Beam Search for Diffusion Language Models"

Python 8 Updated Jul 6, 2026

[ICML 2026 Outstanding Paper] Minimalist RL for Diffusion LLMs. 89.1% on GSM8K.

Python 261 10 Updated Jul 6, 2026

[ICML 2026] d3LLM: Ultra-Fast Diffusion LLM 🚀

Python 149 10 Updated May 1, 2026

Fully open reproduction of DeepSeek-R1

Python 26,431 2,445 Updated Apr 2, 2026

The official GitHub repo for the survey paper "A Survey on Diffusion Language Models".

1,173 58 Updated May 29, 2026

dLLM: Simple Diffusion Language Modeling

Python 2,666 282 Updated Jul 17, 2026

[ICLR 2026] Official repository of "Beyond Fixed: Training-Free Variable-Length Denoising for Diffusion Large Language Models"

Python 174 7 Updated Feb 16, 2026

Official PyTorch implementation for "Large Language Diffusion Models"

Python 3,932 273 Updated Jul 15, 2026

Official PyTorch implementation of CD-MOE

Python 12 Updated Mar 18, 2026

Official Pytorch Implementation of "Outlier-weighed Layerwise Sampling for LLM Fine-tuning" by Pengxiang Li, Lu Yin, Xiaowei Gao, Shiwei Liu

Python 35 7 Updated Jun 3, 2025
Next