Skip to content
View Ericva's full-sized avatar
  • Beijing Institute of Technology
  • Beijing

Block or report Ericva

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Harness engineering beginner tutorial, from 0 to 1

TypeScript 11,130 1,207 Updated Aug 6, 2026

The agent that grows with you

Python 228,286 44,859 Updated Aug 10, 2026

给 Claude Code 装上完整联网能力的 skill:三层通道调度 + 浏览器 CDP + 并行分治

JavaScript 8,617 614 Updated May 16, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 385,783 81,078 Updated Aug 10, 2026
Python 3,599 753 Updated May 28, 2026

This guide is designed for OpenClaw itself (Agent-facing), not as a traditional human-only hardening checklist.

Shell 2,856 196 Updated Apr 6, 2026

Deep Reinforcement Learning

4,681 677 Updated Dec 10, 2022

Nano vLLM

Python 14,935 2,441 Updated Apr 26, 2026

Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving

Python 449 31 Updated Aug 19, 2025

This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."

MATLAB 17,424 1,658 Updated Aug 10, 2026

Tools for merging pretrained large language models.

Python 7,287 777 Updated Jun 17, 2026

Ray tutorials from Anyscale

Jupyter Notebook 638 208 Updated Jun 25, 2026

这是一个从头训练大语言模型的项目,包括预训练、微调和直接偏好优化,模型拥有1B参数,支持中英文。

Python 862 112 Updated Feb 18, 2025

Train transformer language models with reinforcement learning.

Python 19,039 2,900 Updated Aug 10, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,900 4,369 Updated Aug 10, 2026

Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.

Python 43,488 7,910 Updated Aug 10, 2026

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

Python 14,351 2,651 Updated Aug 10, 2026

My learning notes for ML SYS.

HTML 6,847 474 Updated Aug 8, 2026

No fortress, purely open ground. OpenManus is Coming.

Python 57,911 10,061 Updated Feb 11, 2026

RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.

Python 2,765 230 Updated Jul 24, 2026

Muon is Scalable for LLM Training

1,528 100 Updated Aug 3, 2025

The Official Python Client for Lamini's API

Python 2,533 152 Updated Apr 7, 2025

Simple introduction to LLM Agents

Jupyter Notebook 137 40 Updated Mar 16, 2024

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 73,967 9,050 Updated Aug 10, 2026

Embark on the "Reinforcement Learning from Human Feedback" course and align Large Language Models (LLMs) with human values.

Jupyter Notebook 13 9 Updated Jan 31, 2024

The official GitHub page for the survey paper "A Survey of Large Language Models".

Python 12,204 932 Updated Mar 11, 2025

https://hrl.boyuai.com/

Jupyter Notebook 4,926 828 Updated Nov 22, 2022

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Python 163,540 34,174 Updated Aug 10, 2026

采用MegEngine实现的各种主流深度学习模型

Python 306 96 Updated Dec 7, 2022

This tool provides an efficient implementation of the continuous bag-of-words and skip-gram architectures for computing vector representations of words. These representations can be subsequently us…

C 1,742 627 Updated May 10, 2021
Next