Skip to content
View WujiangXu's full-sized avatar
🏠
Working from home
🏠
Working from home

Block or report WujiangXu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results
Python 7 Updated Apr 29, 2026

[preprint] sparsity

Python 23 1 Updated Jul 26, 2026

Code and implementations for the paper "AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning" by Zhiheng Xi et al.

Python 845 85 Updated Feb 15, 2026

The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".

Python 21 3 Updated Jun 2, 2026

MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory

Python 50 5 Updated May 17, 2026

Agentic memory using knowledge graphs

Rust 36 3 Updated May 23, 2026

This is a library to use with Robinhood Financial App. It currently supports trading crypto-currencies, options, and stocks. In addition, it can be used to get real time ticker information, assess …

Python 2,109 544 Updated Feb 11, 2026

[NeurIPS 2025 Spotlight] OpenCUA: Open Foundations for Computer-Use Agents

Python 820 107 Updated May 25, 2026

A transparent, minimal, and hackable agent framework. ~300 lines of readable code. Full control, no magic.

Python 454 45 Updated Jan 2, 2026

An Open-Source Large-Scale Reinforcement Learning Project for Search Agents

Python 607 39 Updated Nov 26, 2025

The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench verified!

Python 6,527 898 Updated Aug 10, 2026

The code for paper "EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning"

Python 40 2 Updated Jul 13, 2026

Awesome List for Agentic RL

HTML 1,780 71 Updated Aug 11, 2026

An extremely fast Python package and project manager, written in Rust.

Rust 88,766 3,487 Updated Aug 15, 2026

[EACL'26] DeepSieve: Information Sieving via LLM-as-a-Knowledge-Router

Python 108 11 Updated Jan 4, 2026
Python 328 21 Updated Jan 3, 2026

[ICML 2025 Oral] Official repo of EmbodiedBench, a comprehensive benchmark designed to evaluate MLLMs as embodied agents.

Python 328 40 Updated May 30, 2026
Python 4 Updated Jun 11, 2025

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,973 4,410 Updated Aug 15, 2026

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,917 1,002 Updated Aug 13, 2026

[ICML'25] Our study systematically investigates massive values in LLMs' attention mechanisms. First, we observe massive values are concentrated in low-frequency dimensions across different attentio…

Python 87 3 Updated Jun 20, 2025

[ICLR 2025] Benchmarking Agentic Workflow Generation

Python 156 10 Updated Feb 19, 2025

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,131 9,070 Updated Aug 13, 2026

The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthr…

Python 56,431 10,617 Updated Aug 16, 2026

User Profile-Based Long-Term Memory for AI Chatbot Applications.

Python 2,841 226 Updated Jan 11, 2026

Code for EACL 26 Findings paper "I-MCTS: Enhancing Agentic AutoML via Introspective Monte Carlo Tree Search"

HTML 13 4 Updated Jan 28, 2026

AIOS: AI Agent Operating System

Python 6,233 890 Updated Jul 20, 2026
Next