Skip to content
View xxyQwQ's full-sized avatar
  • The Chinese University of Hong Kong
  • Hong Kong SAR, China
  • 01:50 (UTC +08:00)

Highlights

  • Pro

Block or report xxyQwQ

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Implementation for the paper "StraTA: Incentivizing Agentic Reinforcement Learning with Strategic Trajectory Abstraction".

Python 46 9 Updated May 8, 2026

AI handles execution, humans own the direction, and every run becomes an inspectable research artifact on disk.

Python 805 25 Updated Aug 16, 2026

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.

Rust 195,063 109,097 Updated Aug 16, 2026

LatentMem: Customizing Latent Memory for Multi-Agent Systems

Python 50 8 Updated Feb 9, 2026

🦄️ 🎃 👻 Clash Premium 规则集(RULE-SET),兼容 ClashX Pro、Clash for Windows 等基于 Clash Premium 内核的客户端。

28,029 2,184 Updated Aug 15, 2026

分流规则、重写写规则及脚本。

JavaScript 27,562 4,051 Updated Aug 15, 2026

Elevate your AI research writing, no more tedious polishing ✨

33,003 2,429 Updated May 18, 2026

⏰ Agenticly track worldwide conference deadlines (Website, Python Cli, Wechat Applet)

Rust 9,244 619 Updated Aug 14, 2026

verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

Python 2,221 215 Updated Jun 9, 2026

Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows

Python 168 4 Updated Jun 2, 2026

Democratizing Reinforcement Learning for LLMs

Python 5,784 606 Updated Aug 16, 2026
Python 245 29 Updated Jul 25, 2025

A Framework for LLM-based Multi-Agent Reinforced Training and Inference

Python 547 50 Updated Apr 14, 2026

🏝️ OASIS: Open Agent Social Interaction Simulations with One Million Agents.

Python 5,029 618 Updated Aug 14, 2026

A simple tool to update bib entries with their official information (e.g., DBLP or the ACL anthology).

Python 3,026 164 Updated Aug 10, 2026

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,919 1,003 Updated Aug 13, 2026

A very simple GRPO implement for reproducing r1-like LLM thinking.

Python 1,702 133 Updated Nov 21, 2025

User-friendly AI Interface (Supports Ollama, OpenAI API, ...)

Python 148,948 21,679 Updated Aug 15, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,141 9,072 Updated Aug 13, 2026

Official Repo for Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning

Jupyter Notebook 417 41 Updated Dec 15, 2024

Code and implementations for the ACL 2025 paper "AgentGym: Evolving Large Language Model-based Agents across Diverse Environments" by Zhiheng Xi et al.

Python 829 115 Updated May 30, 2026

The paper list of the 86-page SCIS cover paper "The Rise and Potential of Large Language Model Based Agents: A Survey" by Zhiheng Xi et al.

8,172 495 Updated Sep 12, 2025

Implementation for the paper "ComfyBench: Benchmarking LLM-based Agents in ComfyUI for Autonomously Designing Collaborative AI Systems".

Python 205 10 Updated Dec 24, 2025

A project page template for academic papers. Demo at https://eliahuhorwitz.github.io/Academic-project-page-template/

JavaScript 5,154 1,148 Updated Sep 4, 2025

A repo lists papers related to LLM based agent

Python 2,335 155 Updated Jul 12, 2025

Must-read Papers on LLM Agents.

3,101 184 Updated Jul 27, 2026

😎 Awesome lists about all kinds of interesting topics

496,451 36,420 Updated Jun 30, 2026

The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.

Python 127,923 15,063 Updated Aug 16, 2026
Next