Skip to content
View langfengQ's full-sized avatar

Highlights

  • Pro

Block or report langfengQ

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 220 8 Updated Jul 17, 2026
Python 108 2 Updated Jul 1, 2026

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

JavaScript 238,896 36,283 Updated Aug 8, 2026

SkillsBench evaluates how well skills work and how effective agents are at using them.

PDDL 1,650 357 Updated Jul 23, 2026

[ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting RL-driven self-compression

Python 43 3 Updated Mar 1, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 385,631 81,058 Updated Aug 9, 2026

[ICLR'26] Stronger-MAS: A RL Framework for multi LLM agent system; [arxiv] MetaAgent-X: End-to-End Reinforcement Learning Automatic Multi-Agent Systems

Python 208 28 Updated May 15, 2026

Dr. MAS is an end-to-end RL training framework for multi-agent LLM systems, supporting the co-training of multiple (heterogeneous) LLMs.

Python 151 10 Updated Jul 27, 2026

An RL Recipe for Building Agentic LLMs via Self-Imitation on Long-Horizon Agentic Tasks

Python 39 1 Updated Jan 30, 2026

Stateful runtime management for LLM agents—inject, manipulate, and retrieve Python objects across turns.

Python 181 13 Updated Aug 1, 2026
Python 266 42 Updated Nov 6, 2025

A live stream development of RL tunning for LLM agents

Python 4,145 589 Updated May 5, 2026

[ICLR 2026] End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning

Python 403 26 Updated Mar 30, 2026

SkyRL: A Modular Full-stack RL Library for LLMs

Python 2,138 401 Updated Aug 6, 2026

Official code for paper "TimeMaster: Training Time-Series Multimodal LLMs to Reason via Reinforcement Learning"

Python 68 7 Updated Oct 22, 2025

🦉 OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation

Python 20,076 2,302 Updated Aug 7, 2026

Code and implementations for the ACL 2025 paper "AgentGym: Evolving Large Language Model-based Agents across Diverse Environments" by Zhiheng Xi et al.

Python 828 115 Updated May 30, 2026

An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models

Python 3,356 304 Updated Aug 9, 2026

Must-read Papers on LLM Agents.

3,097 185 Updated Jul 27, 2026

A Survey of Reinforcement Learning for Large Reasoning Models

TeX 2,476 131 Updated Aug 1, 2026

Awesome things about LLM-powered agents. Papers / Repos / Blogs / ...

2,255 227 Updated Apr 30, 2025

A repo lists papers related to LLM based agent

Python 2,333 155 Updated Jul 12, 2025

Official code for paper "Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement Learning"

Python 15 1 Updated Jun 12, 2025

🌍 AppWorld: A Controllable World of Apps and People for Benchmarking Function Calling and Interactive Coding Agent, ACL'24 Best Resource Paper.

Python 478 73 Updated Feb 17, 2026

verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

Python 2,193 208 Updated Jun 9, 2026

This is the implementation of MLF & spiking DS-ResNet

Python 16 2 Updated Dec 20, 2023

Qwen2.5-Omni is an end-to-end multimodal model by Qwen team at Alibaba Cloud, capable of understanding text, audio, vision, video, and performing real-time speech generation.

Jupyter Notebook 4,063 326 Updated Jun 12, 2025

Embodied Agent Interface (EAI): Benchmarking LLMs for Embodied Decision Making (NeurIPS D&B 2024 Oral)

Python 297 28 Updated Mar 6, 2025

EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

Python 5,105 383 Updated Jul 30, 2026
Next