Skip to content
View zhouwx666's full-sized avatar

Block or report zhouwx666

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

The paper list of the 86-page SCIS cover paper "The Rise and Potential of Large Language Model Based Agents: A Survey" by Zhiheng Xi et al.

8,172 495 Updated Sep 12, 2025

This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."

MATLAB 17,470 1,663 Updated Aug 10, 2026

FastAPI framework, high performance, easy to learn, fast to code, ready for production

Python 101,616 9,785 Updated Aug 15, 2026

Use ChatGPT to summarize the arXiv papers. 全流程加速科研,利用chatgpt进行论文全文总结+专业翻译+润色+审稿+审稿回复

Python 19,753 1,935 Updated Mar 2, 2026

A live stream development of RL tunning for LLM agents

Python 4,149 591 Updated May 5, 2026

No fortress, purely open ground. OpenManus is Coming.

Python 1 Updated May 8, 2025

An Easy-to-use, Scalable and High-performance RLHF Framework based on Ray (PPO & GRPO & REINFORCE++ & LoRA & vLLM & RFT)

Python 1 Updated May 12, 2025

LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and vector stores, and makes implementing …

Java 12,874 2,459 Updated Aug 14, 2026