-
Zhejiang University
- Hangzhou, China
- https://jianbiaomei.github.io
Stars
slime is an LLM post-training framework for RL Scaling.
Framework for evaluating and improving agents
Synthetic data generation, post-training, and E2B benchmark evaluation infrastructure.
Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills
Open-source implementation of AlphaEvolve
Research and development (R&D) is crucial for the enhancement of industrial productivity, especially in the AI era, where the core aspects of R&D are mainly focused on data and models. We are commi…
A Survey of Self-Evolving Agents | A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Self-Evolving Agents.
[ACM TIST 2023] Fast Real-time Video Object Segmentation with Tangled Memory Network
[IEEE TIP 2022] Delving Deeper into Mask Utilization in Video Object Segmentation
[ACM MM 2023] CenterLPS: Segment Instances by Centers for LiDAR Panoptic Segmentation
An adaptive dual controller framework for cost-efficient long-horizon game control.
NEO Series: Native Vision-Language Models from First Principles
Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond
SenseNova-U series: Native Unified Paradigm with NEO-unify from the First Principles
AI agents running research on single-GPU nanochat training automatically
Hy3 preview (295B A21B), a leading reasoning and agent model in its size, with great cost efficiency
The agent that grows with you
Official repository for the paper "Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation"
JoyAI-Image is the unified multimodal foundation model for image understanding, text-to-image generation, and instruction-guided image editing.
An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
Supercharge your AI agents by versioning, tracking, and merging overlapping skills.
OpenClaw-RL: Train any agent simply by talking
[RSS 2026] Causal video-action world model for generalist robot control
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
[ICML-2026] Official implementation of "SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience"