Skip to content
View heyLinsir's full-sized avatar

Block or report heyLinsir

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

OpenClaw 中文官方技能库 | 翻译自 Clawdbot 官方技能,按场景分类整理,支持中文自然语言调用

4,142 Updated May 26, 2026

The awesome collection of OpenClaw skills. 5,400+ skills filtered and categorized from the official OpenClaw Skills Registry.🦞

51,822 4,992 Updated Aug 7, 2026

[ARCHIVED] Old repository for ACL 2025 Paper 'MaXIFE'. The official code is now in the new repository linked below.

1 Updated Aug 20, 2025

Paper list for Efficient Reasoning.

900 47 Updated May 29, 2026

[TMLR 2025] Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

786 40 Updated Feb 28, 2026

GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models

Python 4,411 469 Updated Feb 1, 2026

A collection of awesome-prompt-datasets, awesome-instruction-dataset, to train ChatLLM such as chatgpt 收录各种各样的指令数据集, 用于训练 ChatLLM 模型。

739 43 Updated Jun 17, 2026

[ACL 2025] We introduce ScaleQuest, a scalable, novel and cost-effective data synthesis method to unleash the reasoning capability of LLMs.

Python 69 7 Updated Oct 27, 2024

语言学竞赛集成 / Collection on Linguistics Olympiad (Chinese version only)

30 3 Updated Oct 26, 2024

Must-read Papers on Knowledge Editing for Large Language Models.

1,245 78 Updated Jun 25, 2026

A curated list of reinforcement learning with human feedback resources (continually updated)

4,421 256 Updated May 20, 2026

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,896 995 Updated Jul 14, 2026
JavaScript 18 2 Updated Feb 29, 2024

Reference BLEU implementation that auto-downloads test sets and reports a version string to facilitate cross-lab comparisons

Python 1,254 175 Updated Jul 17, 2026

👨‍💻 An awesome and curated list of best code-LLM for research.

1,291 74 Updated Dec 10, 2024

Google Research

Jupyter Notebook 38,498 8,462 Updated Aug 8, 2026

Reference implementation for DPO (Direct Preference Optimization)

Python 2,903 237 Updated Aug 11, 2024

Code for "Learning to summarize from human feedback"

Python 1,062 153 Updated Sep 5, 2023

ChatGLM3 series: Open Bilingual Chat LLMs | 开源双语对话语言模型

Python 13,663 1,584 Updated Jan 13, 2025

Example models using DeepSpeed

Python 6,835 1,120 Updated Aug 4, 2026

A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)

Python 4,752 486 Updated Jan 8, 2024

Fine-tuning ChatGLM-6B with PEFT | 基于 PEFT 的高效 ChatGLM 微调

Python 3,717 462 Updated Oct 12, 2023

Perspective is an API that uses machine learning models to score the perceived impact a comment might have on a conversation. See https://developers.perspectiveapi.com for more information.

924 117 Updated Mar 25, 2021
Python 283 24 Updated Jan 6, 2025

Secrets of RLHF in Large Language Models Part I: PPO

Python 1,426 104 Updated Mar 3, 2024

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

Python 42,886 4,922 Updated Aug 7, 2026

OpenChat: Advancing Open-source Language Models with Imperfect Data

Python 5,486 432 Updated Sep 13, 2024

Fast and memory-efficient exact attention

Python 24,654 2,970 Updated Aug 7, 2026

ChatGLM2-6B: An Open Bilingual Chat LLM | 开源双语对话语言模型

Python 15,535 1,792 Updated Jun 27, 2024

基于ChatGLM-6B、ChatGLM2-6B、ChatGLM3-6B模型,进行下游具体任务微调,涉及Freeze、Lora、P-tuning、全参微调等

Python 2,772 308 Updated Dec 12, 2023
Next