Skip to content
View yzoaim's full-sized avatar
  • Beijing
  • 02:32 (UTC +08:00)

Block or report yzoaim

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

🐧 Harness for RSI. AI Build Better AI

TypeScript 1,268 120 Updated Aug 12, 2026

Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning

Python 426 49 Updated May 28, 2026

Wrap Antigravity, ChatGPT Codex, Claude Code, Grok Build as an OpenAI/Gemini/Claude/Codex compatible API service, allowing you to enjoy the free Gemini 3.1 Pro, GPT 5.6 Series, Grok 4.5, Claude mod…

Go 47,189 7,300 Updated Aug 13, 2026

Unofficial implementation of the Dreamer 4 world model in PyTorch.

Python 382 40 Updated Jul 9, 2026

Lets make video diffusion practical!

Python 17,203 1,736 Updated Oct 16, 2025

Environment-Native Verified Search (ENVS): GUI agent post training pipeline that reaches higher accuracy at lower compute than online RL.

Python 5 Updated Jun 23, 2026

Research and development (R&D) is crucial for the enhancement of industrial productivity, especially in the AI era, where the core aspects of R&D are mainly focused on data and models. We are commi…

Python 14,223 1,824 Updated Aug 4, 2026

[ICML 2026] What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom

Python 23 Updated May 15, 2026

Code for paper OpenWebRL: Online Multi-Turn Reinforcement Learning for Visual Web Agents

Python 38 3 Updated Aug 8, 2026

InfoSFT is a modified supervised fine-tuning algorithm that generalizes better and forgets less.

Python 5 Updated May 19, 2026

CUA-Gym-Hub: mock web apps as reproducible RL training environments for computer-use agents

JavaScript 71 10 Updated Jul 27, 2026

Scalable pipeline for synthesizing verifiable RLVR training data for computer-use agents

Python 186 16 Updated Aug 13, 2026
Python 289 21 Updated Jul 31, 2026

Symphony turns project work into isolated, autonomous implementation runs, allowing teams to manage work instead of supervising coding agents.

Elixir 26,560 2,708 Updated Aug 12, 2026

OpenSeeker: A search agent with open-source data and models

Python 767 60 Updated Jun 25, 2026

[NeurIPS 2025 Spotlight] OpenCUA: Open Foundations for Computer-Use Agents

Python 818 106 Updated May 25, 2026

Code for "WebVoyager: WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models"

Python 1,117 122 Updated Mar 4, 2024

An Illusion of Progress? Assessing the Current State of Web Agents

Python 193 14 Updated Jun 25, 2026

Windows Agent Arena (WAA) 🪟 is a scalable OS platform for testing and benchmarking of multi-modal AI agents.

Python 886 99 Updated Apr 13, 2026

This is the official code base of AgentNetTool in OpenCUA. Website: https://opencua.xlang.ai/

TypeScript 53 10 Updated Sep 3, 2025

SWE-bench: Can Language Models Resolve Real-world Github Issues?

Python 5,632 938 Updated Aug 13, 2026

[NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments

Python 3,080 516 Updated Aug 12, 2026

Implementation of paper: Scaling the Scaling Logic

Python 3 Updated Mar 1, 2026
Python 160 6 Updated May 14, 2025

The official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.

Python 9,048 777 Updated Mar 25, 2026

[ICLR 2026] Official code for TraceRL: Revolutionizing post-training for Diffusion LLMs, powering the SOTA TraDo series.

Python 519 44 Updated Jan 28, 2026

Easy and Efficient dLLM Fine-Tuning

Python 264 18 Updated Aug 4, 2026

DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation

Python 834 59 Updated Jul 9, 2025

MMaDA - Open-Sourced Multimodal Large Diffusion Language Models (dLLMs with block diffusion, mixed-CoT, unified RL)

Python 1,663 91 Updated Feb 14, 2026
Next