Skip to content
View Zijun9's full-sized avatar

Highlights

  • Pro

Block or report Zijun9

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

DeepSeek Harness: Everything is a Plugin.

TypeScript 92,485 8,435 Updated Aug 13, 2026

Diving into Reliable Self-Evolving Agents: A Survey

HTML 67 1 Updated Aug 12, 2026

Official AHE code — Agentic Harness Engineering: observability-driven automatic evolution of coding-agent harnesses (concurrent w/ meta-harness). NexAU-AHE reaches 84.7% ± 2.1 pass@1 on Terminal-Be…

Python 825 93 Updated Aug 3, 2026

AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI

TypeScript 90,326 11,206 Updated Aug 14, 2026

A benchmark of real-world DL kernel problems

Python 277 33 Updated Jul 15, 2026

A self-evolving multi-agent system for autonomous research, operating 24/7 to explore, learn, and improve.

Python 256 20 Updated Aug 14, 2026

Agentic RL on Any Harness at Scale

Python 780 83 Updated Aug 13, 2026

SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?

Python 500 95 Updated May 18, 2026

open source SWE-Atlas

Shell 67 5 Updated Jul 20, 2026

The roadmap of long-horizon agents

954 36 Updated Aug 11, 2026

Synthetic data generation, post-training, and E2B benchmark evaluation infrastructure.

Shell 111 5 Updated Aug 13, 2026

🗓️ The hardest life-admin benchmark for agents — lawsuits, escrow shortfalls, apartment hunts, exams. 20 long-horizon tasks × 20–30 stages across 10 domains and 21 services, scored by 1247 atomic c…

PLpgSQL 6 1 Updated Jul 31, 2026

An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

Python 571 139 Updated Aug 13, 2026

[ICLR 2026] VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications

Python 164 17 Updated Feb 22, 2026

Ď„-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Python 1,796 452 Updated Aug 12, 2026

Official implementation of CORE (ICLR 2026) — Concept-Oriented Reinforcement for bridging the definition–application gap in mathematical reasoning. Concept-guided GRPO (CORE-Base/CR/KL), SC@21 eval…

Python 1 Updated Jul 7, 2026

Post-training with Tinker

Python 4,019 509 Updated Aug 14, 2026

The agent that grows with you

Python 230,519 45,675 Updated Aug 14, 2026

AI turns documents or topics into real, native PowerPoint decks—with native shapes, transitions and animations, data-backed charts and tables on demand, audio narration from speaker notes, and supp…

Python 46,804 3,797 Updated Aug 14, 2026

RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios

Python 616 59 Updated Jun 12, 2026
Python 585 79 Updated Jul 27, 2026

AgentCPM-GUI: An on-device GUI agent for operating Android apps, enhancing reasoning ability with reinforcement fine-tuning for efficient task execution.

Python 1,404 132 Updated Jan 11, 2026

My learning notes for ML SYS.

HTML 6,865 478 Updated Aug 12, 2026

ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works…

Python 14,683 1,293 Updated Aug 11, 2026

from vibe coding to agentic engineering - practice makes claude perfect

HTML 64,469 6,404 Updated Aug 14, 2026

Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepowe…

TeX 11,699 855 Updated Jun 16, 2026

Code and implementations for the paper "AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning" by Zhiheng Xi et al.

Python 844 85 Updated Feb 15, 2026
Next