Skip to content
View Zijun9's full-sized avatar

Highlights

  • Pro

Block or report Zijun9

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

DeepSeek Harness: Everything is a Plugin.

TypeScript 39,467 3,086 Updated Aug 13, 2026

Diving into Reliable Self-Evolving Agents: A Survey

HTML 53 Updated Aug 12, 2026

Official AHE code — Agentic Harness Engineering: observability-driven automatic evolution of coding-agent harnesses (concurrent w/ meta-harness). NexAU-AHE reaches 84.7% ± 2.1 pass@1 on Terminal-Be…

Python 823 93 Updated Aug 3, 2026

AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI

TypeScript 89,566 11,115 Updated Aug 13, 2026

A benchmark of real-world DL kernel problems

Python 277 33 Updated Jul 15, 2026

A self-evolving multi-agent system for autonomous research, operating 24/7 to explore, learn, and improve.

Python 218 14 Updated Aug 13, 2026

Agentic RL on Any Harness at Scale

Python 776 82 Updated Aug 13, 2026

SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?

Python 500 95 Updated May 18, 2026

open source SWE-Atlas

Shell 67 5 Updated Jul 20, 2026

The roadmap of long-horizon agents

945 35 Updated Aug 11, 2026

Synthetic data generation, post-training, and E2B benchmark evaluation infrastructure.

Shell 110 5 Updated Aug 13, 2026

🗓️ The hardest life-admin benchmark for agents — lawsuits, escrow shortfalls, apartment hunts, exams. 20 long-horizon tasks × 20–30 stages across 10 domains and 21 services, scored by 1247 atomic c…

PLpgSQL 6 1 Updated Jul 31, 2026

An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

Python 570 138 Updated Aug 13, 2026

[ICLR 2026] VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications

Python 164 17 Updated Feb 22, 2026

Ď„-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Python 1,792 449 Updated Aug 12, 2026

Official implementation of CORE (ICLR 2026) — Concept-Oriented Reinforcement for bridging the definition–application gap in mathematical reasoning. Concept-guided GRPO (CORE-Base/CR/KL), SC@21 eval…

Python 1 Updated Jul 7, 2026

Post-training with Tinker

Python 4,018 509 Updated Aug 13, 2026

The agent that grows with you

Python 230,104 45,519 Updated Aug 13, 2026

AI turns documents or topics into real, native PowerPoint decks—with native shapes, transitions and animations, data-backed charts and tables on demand, audio narration from speaker notes, and supp…

Python 46,493 3,776 Updated Aug 13, 2026

RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios

Python 615 58 Updated Jun 12, 2026
Python 585 79 Updated Jul 27, 2026

AgentCPM-GUI: An on-device GUI agent for operating Android apps, enhancing reasoning ability with reinforcement fine-tuning for efficient task execution.

Python 1,406 133 Updated Jan 11, 2026

My learning notes for ML SYS.

HTML 6,865 477 Updated Aug 12, 2026

ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works…

Python 14,645 1,289 Updated Aug 11, 2026

from vibe coding to agentic engineering - practice makes claude perfect

HTML 64,451 6,403 Updated Aug 13, 2026

Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepowe…

TeX 11,673 852 Updated Jun 16, 2026

Code and implementations for the paper "AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning" by Zhiheng Xi et al.

Python 843 85 Updated Feb 15, 2026
Next