-
Beijing Jiaotong University
- Beijing
- https://xl2248.github.io/
Stars
Official repository for the paper "Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation"
Code for "Think Natively: Unlocking Multilingual Reasoning with Consistency-Enhanced Reinforcement Learning".
UR2: Unify RAG and Reasoning through Reinforcement Learning
Learning to Generate STRUCTURED Output with Schema Reinforcement Learning
Multilingual and Multiculture Benchmark and LLM
Official repository for DAC: A Dynamic Attention-aware Approach for Task-Agnostic Prompt Compression
Learn Globally, Speak Locally: Bridging the Gaps in Multilingual Reasoning
The official repo of One RL to See Them All: Visual Triple Unified Reinforcement Learning
RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.
Scalable RL solution for advanced reasoning of language models
a benckmark for evaluating logical reasoning of LLMs
GAOGAO-Bench-Updates is a supplement to the GAOKAO-Bench, a dataset to evaluate large language models.
华中科学大学数学分析与高代代数考研真题
LLM evaluation on 2024 Chinese Gaokao Mathematics — zero-contamination benchmark with dual prompt formats
The most comprehensive database of Chinese poetry 🧶最全中华古诗词数据库, 唐宋两朝近一万四千古诗人, 接近5.5万首唐诗加26万宋诗. 两宋时期1564位词人,21050首词。
[ICML 2024] Memory-Space Visual Prompting for Efficient Vision-Language Fine-Tuning
A family of open-sourced Mixture-of-Experts (MoE) Large Language Models
Fast inference engine for Transformer models
MNBVC(Massive Never-ending BT Vast Chinese corpus)超大规模中文语料集。对标chatGPT训练的40T数据。MNBVC数据集不但包括主流文化,也包括各个小众文化甚至火星文的数据。MNBVC数据集包括新闻、作文、小说、书籍、杂志、论文、台词、帖子、wiki、古诗、歌词、商品介绍、笑话、糗事、聊天记录等一切形式的纯文本中文数据。
This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
✨✨Latest Advances on Multimodal Large Language Models
Implementation of AudioLM, a SOTA Language Modeling Approach to Audio Generation out of Google Research, in Pytorch