-
RedNote (Xiaohongshu)
- China
-
02:10
(UTC +08:00) - https://necolizer.github.io/
- https://orcid.org/0000-0001-6644-4075
- https://scholar.google.com/citations?user=fxBaCW8AAAAJ
Highlights
- Pro
Lists (3)
Sort Name ascending (A-Z)
Stars
RL environments + evals for AI agents. Define once, train anything.
Multilingual Document Layout Parsing in a Single Vision-Language Model
A complete list involved in the emotion-controllable face generation survey
An agentic skills framework & software development methodology that works.
🐹 Clean, uninstall, analyze, optimize, and monitor your Mac from the terminal.
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
[ICML'26 & COLM'26] Agent0 Series: Self-Evolving Agents from Zero Data
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
Tongyi Deep Research, the Leading Open-source Deep Research Agent
Search Self-Play: Pushing the Frontier of Agent Capability without Supervision
slime is an LLM post-training framework for RL Scaling.
MedSoft-Diffusion was early accepted to MICCAI 2025 (top 9%, scores: 5/4/4).
🥨 Lobe Icons - Brings AI/LLM brand logos to your React & React Native apps — static SVG/PNG/WebP, no dependencies.
This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."
Scaling Deep Research via Reinforcement Learning in Real-world Environments.
Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL
ZeroSearch: Incentivize the Search Capability of LLMs without Searching
SkyRL: A Modular Full-stack RL Library for LLMs
Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning
World model reasoning RL for multi-turn VLM agents
High-velocity, monorepo-scale workflow for Git
Megvii FILE Library - Working with Files in Python same as the standard library
A Python package with CLI designed to accelerate the calculation and analysis of materials’︁ transport and thermoelectric properties
Kimi-VL: Mixture-of-Experts Vision-Language Model for Multimodal Reasoning, Long-Context Understanding, and Strong Agent Capabilities
A curated list of reinforcement learning (RL) for agents.
[ICLR 2026] Computer Agent Arena: Toward Human-Centric Evaluation and Analysis of Computer-Use Agents