Skip to content
View Artanic30's full-sized avatar
🏠
Working from home
🏠
Working from home

Organizations

@JeekITClub @ShanghaitechGeekPie

Block or report Artanic30

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[ICML 2026 Oral] Agent-native Mid-training for Software Engineering

Go 73 7 Updated Jun 7, 2026

CVPR 2026 Accepted Paper WikiCLIP: An Efficient Contrastive Baseline for Open-domain Visual Entity Recognition

Python 2 Updated May 13, 2026

[ICLR2026] The official repository for the CodeGym project: "Generalizable End-to-End Tool-Use RL with Synthetic CodeGym"

Python 36 4 Updated Oct 14, 2025

ICLR 2026 Accepted Paper Wiki-R1: Incentivizing Multimodal Reasoning for Knowledge-based VQA via Data and Sampling Curriculum

Python 2 1 Updated Apr 7, 2026
Python 207 18 Updated Mar 16, 2026

An Open-Source Asynchronous Coding Agent

Python 10,519 1,223 Updated Aug 9, 2026

Multimodal Retrieval-augmented Generation Framework Built by Tongyi Lab, Alibaba Group.

Python 975 92 Updated Apr 29, 2026

Official Implementation for TMLR25 DA-DPO: Cost-efficient Difficulty-aware Preference Optimization for Reducing MLLM Hallucinations

Python 4 Updated Jan 25, 2026

Youtu-Tip: Tap for Intelligence, Keep on Device.

Python 593 67 Updated Feb 27, 2026

[ICML 2026] Official resources of "Graph-R1: Towards Agentic GraphRAG Framework via End-to-end Reinforcement Learning".

Python 587 76 Updated Apr 30, 2026

[EMNLP 2025] Official implementation for paper "MoLoRAG: Bootstrapping Document Understanding via Multi-modal Logic-aware Retrieval"

Python 27 1 Updated Mar 17, 2026

Official Repository of MMLONGBENCH-DOC: Benchmarking Long-context Document Understanding with Visualizations

Python 154 6 Updated Sep 28, 2025

NeurIPS 2025 Accepted Paper NoisyGRPO: Incentivizing Multimodal CoT Reasoning via Noise Injection and Bayesian Estimation

Python 4 1 Updated Nov 28, 2025

Official Code for "Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search"

Python 424 18 Updated Jan 29, 2026

Parsing-free RAG supported by VLMs

Python 977 76 Updated Aug 9, 2026

SophiaVL-R1: Reinforcing MLLMs Reasoning with Thinking Reward

Python 94 3 Updated Aug 8, 2025

[NeurIPS 2025] NoisyRollout: Reinforcing Visual Reasoning with Data Augmentation

Python 112 3 Updated Sep 18, 2025

✨✨ [ICLR 2026] R1-Reward: Training Multimodal Reward Model Through Stable Reinforcement Learning

Python 292 22 Updated May 9, 2025
Python 30 Updated Jul 23, 2025

Multimodal RewardBench

Python 68 1 Updated Feb 21, 2025

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,877 4,362 Updated Aug 8, 2026

One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks

Python 4,353 636 Updated Aug 6, 2026

[CVPR 2025 (Oral)] Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Key

Python 112 4 Updated Jan 9, 2026

The Next Step Forward in Multimodal LLM Alignment

Python 199 8 Updated May 1, 2025

[CVPR 2025] Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering

Python 56 1 Updated Jul 14, 2025

MME-CoT: Benchmarking Chain-of-Thought in LMMs for Reasoning Quality, Robustness, and Efficiency

Python 136 7 Updated Aug 5, 2025

[ICLR2026] This is the first paper to explore how to effectively use R1-like RL for MLLMs and introduce Vision-R1, a reasoning MLLM that leverages cold-start initialization and RL training to incen…

Python 1,584 27 Updated Mar 20, 2026

Repo for Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent

Python 430 34 Updated Apr 22, 2025
Python 7 2 Updated Feb 21, 2025
Next