Skip to content
View cqduan's full-sized avatar

Block or report cqduan

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

SkillOpt is a text-space optimizer that trains reusable natural-language skills for frozen LLM agents through trajectory-driven edits, validation-gated updates, and deployable best_skill.md artifacts.

Python 15,862 1,477 Updated Aug 9, 2026

A collection of reinforcement learning notes with formula derivations, problem records, algorithm source codes and exported PDFs in Markdown & LaTeX format.

TeX 34 1 Updated Jun 9, 2026

A series of math-specific large language models of our Qwen2 series.

Python 1,083 161 Updated Jan 11, 2025

Sharing both practical insights and theoretical knowledge about LLM evaluation that we gathered while managing the Open LLM Leaderboard and designing lighteval!

Jupyter Notebook 2,138 124 Updated Dec 3, 2025

Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型

Python 6,651 583 Updated Oct 24, 2024

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,905 4,371 Updated Aug 10, 2026

An open-source AI coding agent that lives in your terminal.

TypeScript 26,901 2,837 Updated Aug 11, 2026

xLAM: A Family of Large Action Models to Empower AI Agent Systems

Python 637 60 Updated Jun 2, 2026

A powerful tool for creating datasets for LLM fine-tuning 、RAG and Eval

JavaScript 14,768 1,521 Updated May 1, 2026

Repair malformed JSON from LLMs, APIs, logs, and user input in Python.

Python 5,066 208 Updated Aug 5, 2026

O1 Replication Journey

2,001 61 Updated Jan 14, 2025

A Survey of Reinforcement Learning for Large Reasoning Models

TeX 2,477 132 Updated Aug 1, 2026

心理健康大模型 (LLM x Mental Health), Pre & Post-training & Dataset & Evaluation & Depoly & RAG, with InternLM / Qwen / Baichuan / DeepSeek / Mixtral / LLama / GLM series models

Python 1,775 224 Updated Jun 18, 2026

No fortress, purely open ground. OpenManus is Coming.

Python 57,912 10,062 Updated Feb 11, 2026

Implement a ChatGPT-like LLM in PyTorch from scratch, step by step

Jupyter Notebook 102,311 15,682 Updated Aug 10, 2026

FinGLM: 致力于构建一个开放的、公益的、持久的金融大模型项目,利用开源开放来促进「AI+金融」。

HTML 2,254 315 Updated May 8, 2024

Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.

Go 178,238 17,343 Updated Aug 10, 2026

北京航空航天大学大数据高精尖中心自然语言处理研究团队开展了智能问答的研究与应用总结。包括基于知识图谱的问答(KBQA),基于文本的问答系统(TextQA),基于表格的问答系统(TableQA)、基于视觉的问答系统(VisualQA)和机器阅读理解(MRC)等,每类任务分别对学术界和工业界进行了相关总结。

1,816 260 Updated Apr 6, 2023

Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and…

Python 38,537 6,263 Updated Nov 10, 2025

文本挖掘和预处理工具(文本清洗、新词发现、情感分析、实体识别链接、关键词抽取、知识抽取、句法分析等),无监督或弱监督方法

Python 2,625 340 Updated May 13, 2024
Python 16 4 Updated May 19, 2023

中文nlp解决方案(大模型、数据、模型、训练、推理)

Jupyter Notebook 3,832 443 Updated Aug 5, 2025

A GPT-4 AI Tutor Prompt for customizable personalized learning experiences.

29,597 3,289 Updated Sep 30, 2025

LangGPT: Empowering everyone to become a prompt expert! 🚀 📌 结构化提示词(Structured Prompt)提出者 📌 元提示词(Meta-Prompt)发起者 📌 最流行的提示词落地范式 | Language of GPT The pioneering framework for structured & meta-prompt…

Jupyter Notebook 12,417 939 Updated Jul 16, 2026

pCLUE: 1000000+多任务提示学习数据集

Jupyter Notebook 509 60 Updated Oct 4, 2022

中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)

Python 18,940 1,842 Updated Apr 19, 2026

A playbook for systematically maximizing the performance of deep learning models.

30,276 2,421 Updated Jun 18, 2024

闻达:一个LLM调用平台。目标为针对特定环境的高效内容生成,同时考虑个人和中小企业的计算资源局限性,以及知识安全和私密性问题

JavaScript 6,161 786 Updated Jan 23, 2025

The agent engineering platform.

Python 143,921 23,974 Updated Aug 11, 2026

Examples and guides for using the OpenAI API

Jupyter Notebook 75,196 12,709 Updated Aug 11, 2026
Next