Skip to content
View zzzyzh's full-sized avatar

Highlights

  • Pro

Organizations

@PRIS-CV

Block or report zzzyzh

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Code for kai0, including training, inference and data collection.

C++ 420 30 Updated Mar 9, 2026

Official style files for papers submitted to venues of the Association for Computational Linguistics

BibTeX Style 1,948 377 Updated Jun 29, 2026
Python 19 Updated Jun 23, 2026

Official Repo of "Flow-OPD: On-Policy Distillation for Flow Matching Models"

Python 280 4 Updated Jun 24, 2026
27 Updated May 31, 2026

[ACM MM 2026] On the Limitations of Cross-Lingual Consistency in Multilingual Text-to-Image Generation

Python 3 1 Updated Jul 11, 2026

Explain in Your Own Words: Improving Reasoning via Token-Selective Dual Knowledge Distillation, The Fourteenth International Conference on Learning Representations (ICLR) 2026, Accepted

Python 12 2 Updated May 9, 2026

OPRD: On-Policy Representation Distillation (https://arxiv.org/abs/2606.06021)

Python 69 2 Updated Jul 16, 2026

openpi-RLT is an openpi-based real-robot RL system with RL-token-guided action refinement.

Python 211 15 Updated May 16, 2026

[RSS 2026] Code for RISE: Self-Improving Robot Policy with Compositional World Model

Python 340 20 Updated Jul 6, 2026

Codex-native Academic Research Skills suite for human-in-the-loop academic research workflows

Python 8,811 411 Updated Aug 18, 2026

Academic Research Skills for Claude Code: research → write → review → revise → finalize

Python 42,931 3,413 Updated Aug 18, 2026

[ECCV 2026] EgoSim: Egocentric World Simulator for Embodiment Interaction Generation

Python 67 3 Updated Jun 26, 2026

NEO Series: Native Vision-Language Models from First Principles

Python 883 31 Updated Jul 27, 2026

The implementation of Teachability-Aware On-Policy Distillation (TA-OPD), the method introduced in Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation.

Python 14 Updated May 25, 2026

Everything about the SmolLM and SmolVLM family of models

Python 3,875 307 Updated May 26, 2026

On Policy Distillation Build on top of Verl

Python 94 9 Updated May 25, 2026

Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

Python 934 66 Updated Jun 29, 2026

Code, data and weights for the paper **What drives success in physical planning with Joint-Embedding Predictive World Models?**

Python 452 50 Updated Apr 11, 2026

Official repository for the paper "Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation"

Python 283 14 Updated May 28, 2026

Post-training with Tinker

Python 4,031 516 Updated Aug 18, 2026

HY-Embodied: Embodied Foundation Models for Real-World Agents

Python 854 16 Updated Jul 15, 2026

Awesome List for On-Policy Distillation

837 18 Updated Jul 31, 2026

Official implementation of "OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning"

Python 237 14 Updated May 30, 2025

Source code for the Refined Policy Distillation paper.

Python 21 Updated Jul 18, 2025

Vision-OPD is a regional-to-global on-policy self-distillation framework that transfers a model's own privileged crop-conditioned perception to its full-image policy, enabling fine-grained visual u…

Python 280 11 Updated Jul 17, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 89,366 20,865 Updated Aug 18, 2026
Next