Skip to content
View LHRLAB's full-sized avatar

Highlights

  • Pro

Block or report LHRLAB

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

LLM agent for electricity-computing co-scheduling, with the ECBench benchmark

Python 5 Updated Jul 31, 2026

An agentic framework for omni-modal question-answer tasks.

Python 14 Updated Mar 30, 2026

ICML 2026 "On the Salience of Low-Probability Tokens for AI-Generated Text Detection: A Multiscale Uncertainty Perspective"

Python 3 Updated May 25, 2026

Official code and metadata release for ERGeoBench, a benchmark for embodied reasoning and geo-localization in multimodal large language models.

Python 10 Updated May 29, 2026
Python 2 2 Updated Jul 21, 2025

Official implementation of AutoBM: physically consistent and simulation-executable programmatic generation for scientific modeling.

Python 5 1 Updated Apr 10, 2026

The first standardized multi-task multimodal benchmark for lung cancer clinical decision support.

Python 10 1 Updated Apr 9, 2026

ONOTE: A comprehensive benchmark for evaluating omnimodal LLMs on Symbolic Music Processing across Staff, Jianpu, and Guitar Tablature with deterministic, zero-hallucination metrics

Python 2 1 Updated Apr 8, 2026

🔬🦞 A self-evolving AI research colleague for scientists. 285 skills, zero hallucination, persistent memory.

TypeScript 875 103 Updated Jun 8, 2026

NS-VLA: Towards Neuro-Symbolic Vision-Language-Action Models

Python 11 2 Updated Jul 28, 2026
Python 4 Updated Apr 14, 2026

FlowSteer: agents designing agentic workflows via reinforced progressive canvas editing.

Python 95 12 Updated May 21, 2026

[ICLR 2026] Official resources of "Token-Guard: Towards Token-Level Hallucination Control via Self-Checking Decoding"

Python 96 6 Updated May 9, 2026

LexGenius: An Expert-Level Benchmark for Large Language Models in Chinese Legal General Intelligence

JavaScript 10 1 Updated Dec 8, 2025

[NeurIPS 2025] Official Repo of Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Python 126 6 Updated Dec 3, 2025

The official implementation of "HYPER: A Foundation Model for Inductive Link Prediction with Knowledge Hypergraphs" (ICLR 2026)

Python 18 2 Updated Mar 25, 2026

Tongyi Deep Research, the Leading Open-source Deep Research Agent

Python 19,811 1,504 Updated Feb 27, 2026

Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning

Python 60 6 Updated Feb 24, 2026

[AAAI 2026 Oral] From Stimuli to Minds: Enhancing Psychological Reasoning in LLMs via Bilateral Reinforcement Learning

Python 5 1 Updated Aug 5, 2025
Python 11 1 Updated Aug 20, 2025

MiroRL is an MCP-first reinforcement learning framework for deep research agent.

Python 247 27 Updated Aug 27, 2025

verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

Python 2,208 212 Updated Jun 9, 2026
Python 63 Updated Sep 3, 2025
Python 5 1 Updated Jun 25, 2025

[ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodal search tools.

Python 477 26 Updated Apr 7, 2026
Python 10 Updated May 22, 2025

Compiler-R1: Towards Agentic Compiler Auto-tuning with Reinforcement Learning

Python 37 4 Updated Jul 14, 2025
Python 9 1 Updated Aug 16, 2025

Accepted by ACL 2025

Python 30 Updated Aug 13, 2025
Next