Skip to content
View Ericva's full-sized avatar
  • Beijing Institute of Technology
  • Beijing

Block or report Ericva

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

将博导十年科研经验炼化为可直接调用的 AI 技能。从 Idea 构思到论文投稿,你的 AI 科研副导师。

Python 5,650 382 Updated Aug 7, 2026

Harness engineering beginner tutorial, from 0 to 1

TypeScript 11,833 1,262 Updated Aug 6, 2026

The agent that grows with you

Python 231,905 46,170 Updated Aug 17, 2026

给 Claude Code 装上完整联网能力的 skill:三层通道调度 + 浏览器 CDP + 并行分治

JavaScript 8,680 620 Updated May 16, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,534 81,217 Updated Aug 17, 2026
Python 3,639 757 Updated May 28, 2026

This guide is designed for OpenClaw itself (Agent-facing), not as a traditional human-only hardening checklist.

Shell 2,856 196 Updated Apr 6, 2026

Deep Reinforcement Learning

4,683 677 Updated Dec 10, 2022

Nano vLLM

Python 15,034 2,468 Updated Apr 26, 2026

Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving

Python 449 32 Updated Aug 19, 2025

This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."

MATLAB 17,488 1,663 Updated Aug 10, 2026

Tools for merging pretrained large language models.

Python 7,293 778 Updated Jun 17, 2026

Ray tutorials from Anyscale

Jupyter Notebook 637 208 Updated Jun 25, 2026

这是一个从头训练大语言模型的项目,包括预训练、微调和直接偏好优化,模型拥有1B参数,支持中英文。

Python 864 112 Updated Feb 18, 2025

Train transformer language models with reinforcement learning.

Python 19,088 2,913 Updated Aug 17, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,999 4,422 Updated Aug 17, 2026

Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.

Python 43,537 7,929 Updated Aug 17, 2026

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

Python 14,398 2,665 Updated Aug 17, 2026

My learning notes for ML SYS.

HTML 6,880 479 Updated Aug 17, 2026

No fortress, purely open ground. OpenManus is Coming.

Python 57,996 10,073 Updated Aug 16, 2026

RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.

Python 2,771 227 Updated Jul 24, 2026

Muon is Scalable for LLM Training

1,533 101 Updated Aug 3, 2025

The Official Python Client for Lamini's API

Python 2,533 152 Updated Apr 7, 2025

Simple introduction to LLM Agents

Jupyter Notebook 137 40 Updated Mar 16, 2024

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,166 9,077 Updated Aug 13, 2026

Embark on the "Reinforcement Learning from Human Feedback" course and align Large Language Models (LLMs) with human values.

Jupyter Notebook 13 9 Updated Jan 31, 2024

The official GitHub page for the survey paper "A Survey of Large Language Models".

Python 12,206 931 Updated Mar 11, 2025

https://hrl.boyuai.com/

Jupyter Notebook 4,933 829 Updated Nov 22, 2022

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Python 164,186 34,259 Updated Aug 17, 2026

采用MegEngine实现的各种主流深度学习模型

Python 306 96 Updated Dec 7, 2022
Next