Skip to content
View langfengQ's full-sized avatar

Highlights

  • Pro

Block or report langfengQ

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 230 8 Updated Jul 17, 2026
Python 111 2 Updated Jul 1, 2026

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

JavaScript 239,884 36,412 Updated Aug 12, 2026

SkillsBench evaluates how well skills work and how effective agents are at using them.

PDDL 1,678 356 Updated Jul 23, 2026

[ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting RL-driven self-compression

Python 44 3 Updated Mar 1, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,167 81,168 Updated Aug 13, 2026

[ICLR'26] Stronger-MAS: A RL Framework for multi LLM agent system; [arxiv] MetaAgent-X: End-to-End Reinforcement Learning Automatic Multi-Agent Systems

Python 208 28 Updated May 15, 2026

Dr. MAS is an end-to-end RL training framework for multi-agent LLM systems, supporting the co-training of multiple (heterogeneous) LLMs.

Python 153 10 Updated Jul 27, 2026

An RL Recipe for Building Agentic LLMs via Self-Imitation on Long-Horizon Agentic Tasks

Python 39 1 Updated Jan 30, 2026

Stateful runtime management for LLM agents—inject, manipulate, and retrieve Python objects across turns.

Python 191 14 Updated Aug 13, 2026
Python 266 42 Updated Nov 6, 2025

A live stream development of RL tunning for LLM agents

Python 4,146 590 Updated May 5, 2026

[ICLR 2026] End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning

Python 404 26 Updated Mar 30, 2026

SkyRL: A Modular Full-stack RL Library for LLMs

Python 2,149 403 Updated Aug 11, 2026

Official code for paper "TimeMaster: Training Time-Series Multimodal LLMs to Reason via Reinforcement Learning"

Python 68 7 Updated Oct 22, 2025

🦉 OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation

Python 20,080 2,299 Updated Aug 7, 2026

Code and implementations for the ACL 2025 paper "AgentGym: Evolving Large Language Model-based Agents across Diverse Environments" by Zhiheng Xi et al.

Python 829 115 Updated May 30, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,361 305 Updated Aug 13, 2026

Must-read Papers on LLM Agents.

3,101 185 Updated Jul 27, 2026

A Survey of Reinforcement Learning for Large Reasoning Models

TeX 2,478 132 Updated Aug 1, 2026

Awesome things about LLM-powered agents. Papers / Repos / Blogs / ...

2,254 229 Updated Apr 30, 2025

A repo lists papers related to LLM based agent

Python 2,334 155 Updated Jul 12, 2025

Official code for paper "Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement Learning"

Python 15 1 Updated Jun 12, 2025

🌍 AppWorld: A Controllable World of Apps and People for Benchmarking Function Calling and Interactive Coding Agent, ACL'24 Best Resource Paper.

Python 482 73 Updated Feb 17, 2026

verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

Python 2,218 212 Updated Jun 9, 2026

This is the implementation of MLF & spiking DS-ResNet

Python 16 2 Updated Dec 20, 2023

Qwen2.5-Omni is an end-to-end multimodal model by Qwen team at Alibaba Cloud, capable of understanding text, audio, vision, video, and performing real-time speech generation.

Jupyter Notebook 4,066 326 Updated Jun 12, 2025

Embodied Agent Interface (EAI): Benchmarking LLMs for Embodied Decision Making (NeurIPS D&B 2024 Oral)

Python 297 28 Updated Mar 6, 2025

EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

Python 5,113 383 Updated Jul 30, 2026
Next