Skip to content
View Necolizer's full-sized avatar

Highlights

  • Pro

Block or report Necolizer

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

RL environments + evals for AI agents. Define once, train anything.

Python 290 67 Updated Aug 9, 2026

Multilingual Document Layout Parsing in a Single Vision-Language Model

Python 9,065 801 Updated Mar 24, 2026

A complete list involved in the emotion-controllable face generation survey

5 Updated May 7, 2026

An agentic skills framework & software development methodology that works.

Shell 269,649 24,098 Updated Aug 8, 2026

🐹 Clean, uninstall, analyze, optimize, and monitor your Mac from the terminal.

Shell 62,776 2,215 Updated Aug 9, 2026

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

JavaScript 238,981 36,298 Updated Aug 9, 2026

Democratizing Reinforcement Learning for LLMs

Python 5,772 599 Updated Aug 7, 2026

[ICML'26 & COLM'26] Agent0 Series: Self-Evolving Agents from Zero Data

Python 1,245 146 Updated Jul 10, 2026

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,898 996 Updated Jul 14, 2026

Tongyi Deep Research, the Leading Open-source Deep Research Agent

Python 19,804 1,503 Updated Feb 27, 2026

Search Self-Play: Pushing the Frontier of Agent Capability without Supervision

Python 105 8 Updated Jul 23, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,821 1,130 Updated Aug 7, 2026

MedSoft-Diffusion was early accepted to MICCAI 2025 (top 9%, scores: 5/4/4).

Python 43 Updated Mar 1, 2025

🥨 Lobe Icons - Brings AI/LLM brand logos to your React & React Native apps — static SVG/PNG/WebP, no dependencies.

TypeScript 2,352 224 Updated Jul 24, 2026
TypeScript 1 Updated May 29, 2025

This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."

MATLAB 17,413 1,657 Updated Aug 5, 2026

Scaling Deep Research via Reinforcement Learning in Real-world Environments.

Python 792 54 Updated May 10, 2026
Python 4,596 503 Updated Apr 22, 2026

Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL

Python 5,278 477 Updated Nov 13, 2025

ZeroSearch: Incentivize the Search Capability of LLMs without Searching

Python 1,307 120 Updated Aug 16, 2025

SkyRL: A Modular Full-stack RL Library for LLMs

Python 2,138 400 Updated Aug 6, 2026

Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning

Python 1,599 110 Updated Aug 3, 2026

World model reasoning RL for multi-turn VLM agents

Python 493 60 Updated Jul 23, 2026

High-velocity, monorepo-scale workflow for Git

Rust 4,109 111 Updated Aug 1, 2026

Megvii FILE Library - Working with Files in Python same as the standard library

Python 176 20 Updated Jul 27, 2026

A Python package with CLI designed to accelerate the calculation and analysis of materials’︁ transport and thermoelectric properties

Python 3 Updated Apr 17, 2026

Kimi-VL: Mixture-of-Experts Vision-Language Model for Multimodal Reasoning, Long-Context Understanding, and Strong Agent Capabilities

1,220 97 Updated Jul 15, 2025

A curated list of reinforcement learning (RL) for agents.

111 5 Updated Jul 27, 2026

[ICLR 2026] Computer Agent Arena: Toward Human-Centric Evaluation and Analysis of Computer-Use Agents

HTML 67 4 Updated Feb 26, 2026
Next