Skip to content
View zhao1iang's full-sized avatar
🎯
Focusing
🎯
Focusing
  • Institution of Computational Linguistics at Peking University

Block or report zhao1iang

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Scalable toolkit for efficient model reinforcement

Python 1,913 517 Updated Aug 18, 2026

Backend that powers the dataset viewer on Hugging Face dataset pages through a public API.

Python 898 123 Updated Aug 18, 2026

Democratizing Reinforcement Learning for LLMs

Python 5,788 607 Updated Aug 18, 2026

Automatic evals for LLMs

HTML 604 88 Updated Feb 24, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 23,019 4,426 Updated Aug 18, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 32,021 7,982 Updated Aug 18, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 89,347 20,854 Updated Aug 18, 2026

A framework for few-shot evaluation of language models.

Python 13,702 3,495 Updated Aug 14, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,197 9,077 Updated Aug 18, 2026

O1 Replication Journey

2,001 61 Updated Jan 14, 2025

A flexible and efficient training framework for large-scale alignment tasks

Python 451 40 Updated Oct 23, 2025

This is the repository that contains the source code for the Self-Evaluation Guided MCTS for online DPO.

Jupyter Notebook 332 38 Updated Jan 29, 2026

A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.

6,899 369 Updated Dec 17, 2025

Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Python 8,659 1,120 Updated Sep 14, 2024

The official repository of our survey paper: "Towards a Unified View of Preference Learning for Large Language Models: A Survey"

192 4 Updated Oct 28, 2024

✨✨Latest Advances on Multimodal Large Language Models

17,980 1,133 Updated Aug 14, 2026

[ECCV2024] Video Foundation Models & Data for Multimodal Understanding

Python 2,362 159 Updated Jul 2, 2026

Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.

Jupyter Notebook 19,803 1,833 Updated Jan 30, 2026

VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Python 1,306 89 Updated Jan 23, 2025

A collection of awesome video generation studies.

TeX 781 43 Updated Mar 31, 2026

A Doctor for your data

Python 3,485 256 Updated Jun 16, 2026

Train transformer language models with reinforcement learning.

Python 19,098 2,914 Updated Aug 18, 2026

Skywork-MoE: A Deep Dive into Training Techniques for Mixture-of-Experts Language Models

140 6 Updated Jun 12, 2024

The official Meta Llama 3 GitHub site

Python 29,252 3,528 Updated Jan 26, 2025

SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding challenges. [NeurIPS 2024]

Python 20,076 2,200 Updated Aug 17, 2026

Linter for C++ based on Google's style guide

Python 1,843 312 Updated Aug 14, 2026

⚙️ A curated list of static analysis (SAST) tools and linters for all programming languages, config files, build tools, and more. The focus is on tools which improve code quality.

Rust 14,737 1,505 Updated Jun 10, 2026

Static Code Analysis - 静态代码分析

Python 1,842 305 Updated Nov 3, 2025

Code for the curation of The Stack v2 and StarCoder2 training data

Jupyter Notebook 145 14 Updated Apr 11, 2024
Next