Skip to content
View Ericva's full-sized avatar
  • Beijing Institute of Technology
  • Beijing

Block or report Ericva

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

将博导十年科研经验炼化为可直接调用的 AI 技能。从 Idea 构思到论文投稿,你的 AI 科研副导师。

Python 5,609 379 Updated Aug 7, 2026

Harness engineering beginner tutorial, from 0 to 1

TypeScript 11,440 1,237 Updated Aug 6, 2026

The agent that grows with you

Python 231,232 45,956 Updated Aug 16, 2026

给 Claude Code 装上完整联网能力的 skill:三层通道调度 + 浏览器 CDP + 并行分治

JavaScript 8,670 617 Updated May 16, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,428 81,209 Updated Aug 16, 2026
Python 3,633 757 Updated May 28, 2026

This guide is designed for OpenClaw itself (Agent-facing), not as a traditional human-only hardening checklist.

Shell 2,856 196 Updated Apr 6, 2026

Deep Reinforcement Learning

4,681 677 Updated Dec 10, 2022

Nano vLLM

Python 15,012 2,463 Updated Apr 26, 2026

Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving

Python 449 32 Updated Aug 19, 2025

This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."

MATLAB 17,471 1,663 Updated Aug 10, 2026

Tools for merging pretrained large language models.

Python 7,292 779 Updated Jun 17, 2026

Ray tutorials from Anyscale

Jupyter Notebook 638 208 Updated Jun 25, 2026

这是一个从头训练大语言模型的项目,包括预训练、微调和直接偏好优化,模型拥有1B参数,支持中英文。

Python 864 112 Updated Feb 18, 2025

Train transformer language models with reinforcement learning.

Python 19,083 2,911 Updated Aug 16, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,978 4,410 Updated Aug 15, 2026

Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.

Python 43,526 7,928 Updated Aug 16, 2026

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

Python 14,392 2,663 Updated Aug 16, 2026

My learning notes for ML SYS.

HTML 6,871 479 Updated Aug 16, 2026

No fortress, purely open ground. OpenManus is Coming.

Python 57,984 10,066 Updated Feb 11, 2026

RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.

Python 2,770 228 Updated Jul 24, 2026

Muon is Scalable for LLM Training

1,532 101 Updated Aug 3, 2025

The Official Python Client for Lamini's API

Python 2,533 152 Updated Apr 7, 2025

Simple introduction to LLM Agents

Jupyter Notebook 137 40 Updated Mar 16, 2024

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,138 9,072 Updated Aug 13, 2026

Embark on the "Reinforcement Learning from Human Feedback" course and align Large Language Models (LLMs) with human values.

Jupyter Notebook 13 9 Updated Jan 31, 2024

The official GitHub page for the survey paper "A Survey of Large Language Models".

Python 12,206 931 Updated Mar 11, 2025

https://hrl.boyuai.com/

Jupyter Notebook 4,933 827 Updated Nov 22, 2022

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Python 164,133 34,249 Updated Aug 15, 2026

采用MegEngine实现的各种主流深度学习模型

Python 306 96 Updated Dec 7, 2022
Next