Skip to content
View xxlllz's full-sized avatar

Block or report xxlllz

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
TypeScript 16 Updated Jun 15, 2026
Python 63 7 Updated Jun 15, 2026

[CVPR 2026] Fine-Grained Post-Training Quantization for Large Vision Language Models with Quantization-Aware Integrated Gradients

Python 23 1 Updated Jun 21, 2026

[ACL'26 Findings] OmniDiagram: Advancing Unified Diagram Code Generation via Visual Interrogation Reward

13 1 Updated Apr 14, 2026
Jupyter Notebook 7 Updated Apr 21, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,620 1,094 Updated Jul 24, 2026

World Models for Policy Refinement in StarCraft II

Python 10 1 Updated Feb 18, 2026
Python 7 1 Updated Jun 30, 2026
Python 23 1 Updated May 28, 2026
Python 26 1 Updated Jan 23, 2026

Code for 🌍 UI-Simulator: LLMs as Scalable, General-Purpose Simulators For Evolving Digital Agent Training

Python 21 1 Updated Oct 17, 2025

Resources and paper list for "Thinking with Images for LVLMs". This repository accompanies our survey on how LVLMs can leverage visual information for complex reasoning, planning, and generation.

1,493 47 Updated Mar 9, 2026

🌎💪 BrowserGym, a Gym environment for web task automation

Python 1,288 183 Updated Jul 17, 2026

[AAAI'26] Advancing Chemical Vision-Language Models via Efficient Visual Token Reduction and Complex Reaction Tasks

Python 12 Updated Feb 28, 2026

Official Implementation for the paper "VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models"

23 1 Updated Aug 14, 2025

[CVPR 2026] Reading or Reasoning? Format Decoupled Reinforcement Learning for Document OCR

18 1 Updated Mar 23, 2026

SynthAgent: Adapting Web Agents with Synthetic Supervision, ACL 2026

Python 33 1 Updated Apr 26, 2026

[ICML'24] SeeAct is a system for generalist web agents that autonomously carry out tasks on any given website, with a focus on large multimodal models (LMMs) such as GPT-4V(ision).

Python 851 110 Updated Feb 3, 2025

EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

Python 5,081 383 Updated Jul 23, 2026

OCRVerse: Towards Holistic OCR in End-to-End Vision-Language Models

30 1 Updated Feb 4, 2026
Python 42 1 Updated Jan 9, 2026

The official implementation of BFSM.

18 2 Updated Sep 30, 2025

[NeurIPS'25 Spotlight🔥] Official Implementation of RobustMerge: Parameter-Efficient Model Merging for MLLMs with Direction Robustness

Python 67 4 Updated Jun 24, 2026

Tongyi Deep Research, the Leading Open-source Deep Research Agent

Python 19,713 1,507 Updated Feb 27, 2026
Python 13 Updated Aug 7, 2025
67 Updated Sep 6, 2025

GRPO training code which scales to 32xH100s for long horizon terminal/coding tasks. Base agent is now the top Qwen3 agent on Stanford's TerminalBench leaderboard.

Python 399 26 Updated Aug 24, 2025

[ACL 2025 Oral] The official repository of our paper: CADReview: Automatically Reviewing CAD Programs with Error Detection and Correction

Python 23 3 Updated Aug 8, 2025

[ICLR 2026] Breaking the SFT Plateau: Multimodal Structured Reinforcement Learning for Chart-to-Code Generation

12 Updated Jan 27, 2026

Chart-R1: Chain-of-Thought Supervision and Reinforcement for Advanced Chart Reasoner

24 Updated Aug 7, 2025
Next