Skip to content
View longxudou's full-sized avatar

Organizations

@HIT-SCIR @sail-sg @sea-sailor @terminal-agent

Block or report longxudou

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 26 1 Updated Aug 7, 2026

Code and Data for paper "GameCraft-Bench: Can Agents Build Playable Games End-to-End in a Real Game Engine?"

Shell 191 12 Updated Aug 5, 2026

A Call of Duty-quality FPS in Three.js, built from a single prompt.

JavaScript 2,948 437 Updated Jul 25, 2026

A Survey on Large Language Model-Based Game Agents (ACM CSUR)

944 33 Updated Jun 7, 2026

Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.

HTML 21,009 1,432 Updated Aug 7, 2026
Python 95 6 Updated Aug 6, 2026

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents

Python 216 9 Updated Jul 31, 2026

A runtime substrate that turns an agent's execution into a reversible, Git-like trace, so meta-agents can observe, fork, replay, and revert any run. Couples agent and environments in a copy-on-writ…

Python 1,714 135 Updated Jul 21, 2026

Turn any agent into a life science expert with NVIDIA BioNeMo skills.

Python 412 62 Updated Aug 7, 2026

AI demo for playing ARPG/Soul-like game with RL frame

Python 401 71 Updated Sep 24, 2024

AgentSims is an easy-to-use infrastructure for researchers from all disciplines to test the specific capacities they are interested in.

Python 960 120 Updated Nov 18, 2023

The official repo for "CodeScaler: Scaling Code LLM Training and Test-Time Inference via Execution-Free Reward Models"

Python 35 1 Updated Mar 26, 2026

💻 Terminal-Agent with Human-in-the-Loop Learning

Python 41 2 Updated Jan 16, 2026

Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL training

Python 72 15 Updated Jul 28, 2025

The official repository for "Rongsheng Wang's Arxiv Template"

TeX 65 9 Updated May 7, 2025

SkyRL: A Modular Full-stack RL Library for LLMs

Python 2,133 400 Updated Aug 6, 2026

Interleaving Reasoning: Next-Generation Reasoning Systems for AGI

281 12 Updated Jun 5, 2026

Defeating the Training-Inference Mismatch via FP16

Python 198 17 Updated Nov 14, 2025

[ICML'26] Scaling Long-Horizon LLM Agent via Context-Folding

Python 179 12 Updated May 18, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,800 1,125 Updated Aug 7, 2026
C 15 Updated Oct 13, 2025

User Profile-Based Long-Term Memory for AI Chatbot Applications.

Python 2,833 226 Updated Jan 11, 2026

A tool for exploring each layer in a docker image

Go 54,425 1,992 Updated Dec 15, 2025

Docker image registry for SWE-bench, created by Epoch AI.

Python 19 2 Updated Aug 21, 2025

Fast, Flexible and Portable Structured Generation

C++ 1,813 181 Updated Aug 6, 2026

Cost-efficient and pluggable Infrastructure components for GenAI inference

Go 4,996 640 Updated Aug 6, 2026

A curated list of papers related to constrained decoding of LLM, along with their relevant code and resources.

365 17 Updated Jul 27, 2026

☑️ A simple and extensible shell script for managing your todo.txt file.

Shell 6,154 736 Updated Aug 7, 2026

A Tool to Visualize Claude Code's LLM Interactions

JavaScript 2,408 409 Updated Aug 26, 2025

The official github repo for "Diffusion Language Models are Super Data Learners".

Python 228 8 Updated Nov 6, 2025
Next