Stars
The paper list of the 86-page SCIS cover paper "The Rise and Potential of Large Language Model Based Agents: A Survey" by Zhiheng Xi et al.
This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."
FastAPI framework, high performance, easy to learn, fast to code, ready for production
Use ChatGPT to summarize the arXiv papers. 全流程加速科研,利用chatgpt进行论文全文总结+专业翻译+润色+审稿+审稿回复
A live stream development of RL tunning for LLM agents
zhouwx666 / OpenManus
Forked from FoundationAgents/OpenManusNo fortress, purely open ground. OpenManus is Coming.
zhouwx666 / OpenRLHF
Forked from OpenRLHF/OpenRLHFAn Easy-to-use, Scalable and High-performance RLHF Framework based on Ray (PPO & GRPO & REINFORCE++ & LoRA & vLLM & RFT)
LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and vector stores, and makes implementing …