Skip to content
View wshi83's full-sized avatar
🐈
Pawsitive
🐈
Pawsitive

Highlights

  • Pro

Block or report wshi83

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Agents' Last Exam

Python 945 60 Updated Aug 16, 2026
Python 18 5 Updated Aug 7, 2026
Python 120 21 Updated Jul 16, 2026

The original nirholas/claude-code before DMCA and take down. Once everything is cleared, it will return. Working with Anthropic and Github to get everything back.

6,249 13 Updated Aug 17, 2026
Jupyter Notebook 27 2 Updated Mar 17, 2026

This is the code repo for the paper AceSearcher: Bootstrapping Reasoning and Search for LLMs via Reinforced Self-Play (NeurIPS 2025 Spotlight).

Python 25 Updated Sep 29, 2025

Large model-assisted paper review

595 50 Updated Mar 19, 2026
Python 1 Updated Oct 7, 2025

gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI

Python 20,313 2,138 Updated Jul 24, 2026

Agentic design of virtual cell models

Python 161 18 Updated Aug 1, 2026

Collection of latest papers and materials in the area of RLVR!

Python 138 7 Updated Aug 17, 2026

Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving

Python 449 32 Updated Aug 19, 2025

A Survey of Reinforcement Learning for Large Reasoning Models

TeX 2,479 133 Updated Aug 16, 2026

Official Code Repository for paper "Towards Better Instruction Following Retrieval Models"

Python 8 Updated May 16, 2025

[ICLR'26] MedAgentGYM: Training LLM Agents for Code-Based Medical Reasoning at Scale

Python 127 4 Updated Apr 12, 2026

Official Code Repository for WorkForceAgent-R1

Python 7 Updated Jun 1, 2025

[Patterns] MedAgentsBench: Benchmarking Thinking Models and Agent Frameworks for Complex Medical Reasoning

Jupyter Notebook 82 9 Updated Mar 10, 2026
Python 30 4 Updated Apr 8, 2025

GeoAI: Artificial Intelligence for Geospatial Data

Python 3,284 466 Updated Aug 17, 2026

Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning

Python 1,613 113 Updated Aug 10, 2026

[EMNLP 2024] MIMIR: A Streamlined Platform for Personalized Agent Tuning in Domain Expertise https://arxiv.org/abs/2404.04285

Python 3 Updated Nov 10, 2024

Code and data for TrialGPT.

Python 167 74 Updated Jan 24, 2025
Python 13 6 Updated May 15, 2024

QBRC Somatic Mutation Calling Pipeline

C 16 7 Updated Feb 8, 2022

[EMNLP2025] LightRAG: Simple and Fast Retrieval-Augmented Generation

Python 38,930 5,474 Updated Aug 17, 2026
Python 17 Updated Jan 26, 2024

[EMNLP'24] MedAdapter: Efficient Test-Time Adaptation of Large Language Models Towards Medical Reasoning

Python 36 3 Updated Dec 26, 2024

Code for paper Chain-of-Table: Evolving Tables in the Reasoning Chain for Table Understanding

Python 94 19 Updated Jun 18, 2024
Next