Skip to content
View BRZ911's full-sized avatar
  • China
  • 14:50 (UTC +08:00)

Highlights

  • Pro

Organizations

@MLNLP-World

Block or report BRZ911

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[ACL 2025] The code repository for "Mitigating Visual Forgetting via Take-along Visual Conditioning for Multi-modal Long CoT Reasoning" in PyTorch.

Python 140 Updated May 16, 2025

ACL'2025: SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs. and preprint: SoftCoT++: Test-Time Scaling with Soft Chain-of-Thought Reasoning

Python 94 16 Updated May 30, 2025

[CVPR 2026] Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens

Python 295 17 Updated Aug 2, 2025

Claude Code skill: translate English PDFs to Chinese via pdf2zh, preserving layout/formulas/figures.

Shell 28 6 Updated Jul 6, 2026
Python 1 Updated Jul 21, 2026

Awesome-Long2short-on-LRMs is a collection of state-of-the-art, novel, exciting long2short methods on large reasoning models. It contains papers, codes, datasets, evaluations, and analyses.

262 10 Updated Mar 7, 2026

Latent Video Cache for Video Reasoning

Python 5 Updated Jul 7, 2026

GitHub Pages site for the Next-Generation AI Systems paper

HTML 2 Updated Jun 15, 2026

Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

Python 847 58 Updated Jun 29, 2026

Official codebase for the paper Latent Visual Reasoning

Python 171 9 Updated Oct 22, 2025

Fully open reproduction of DeepSeek-R1

Python 26,412 2,446 Updated Apr 2, 2026

An Elegant Academic Homepage Builder

TypeScript 663 198 Updated May 13, 2026

[ECCV 2026] Demystifying Video Reasoning

Python 46 3 Updated Jul 14, 2026

[ACL 2026] Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning

Python 93 2 Updated Jan 22, 2026

[ICML 2026] Multimodal deep-research MLLM and benchmark. The first long-horizon multimodal deep-research MLLM, extending the number of reasoning turns to dozens and the number of search-engine inte…

Python 658 56 Updated Jun 8, 2026

[AAAI'2026] Let's Think with Images Efficiently! An Interleaved-Modal Chain-of-Thought Reasoning Framework with Dynamic and Precise Visual Thoughts

Python 12 Updated Jun 10, 2026

OMNIFLOW

Python 3 1 Updated Feb 21, 2026

[ICML 2026] Official implementation of "Open-o3 Video: Grounded Video Reasoning with Explicit Spatio-Temporal Evidence"

Python 157 7 Updated May 1, 2026

一个用于在 macOS 上平滑你的鼠标滚动效果或单独设置滚动方向的小工具, 让你的滚轮爽如触控板 | A lightweight tool used to smooth scrolling and set scroll direction independently for your mouse on macOS

Swift 20,973 663 Updated Jul 16, 2026

Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks

Python 4,305 737 Updated Jul 22, 2026

OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.

Python 7,237 817 Updated Jul 27, 2026

A Systematic Survey of Deep Research

320 15 Updated Jan 1, 2026

🔥🔥🔥 Latest Papers, Codes and Datasets on Video-LMM Post-Training

Python 296 13 Updated Mar 3, 2026

We introduce Reasoning via Video, a new paradigm that uses maze-solving video generation to probe multimodal reasoning; our VR-Bench shows that fine-tuned video models consistently outperform stron…

Python 66 6 Updated Feb 4, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,680 4,281 Updated Jul 27, 2026

[ICML 2026 Spotlight] Latent Collaboration in Multi-Agent Systems

Python 1,069 163 Updated Jun 18, 2026

[ECCV 2026] Official repo of "Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens"

Python 379 22 Updated Apr 17, 2026

Lightweight Image Video Action Generation Inference Framework

Python 2,534 236 Updated Jul 27, 2026

This is a collection of recent papers on reasoning in video generation models.

165 6 Updated Jul 21, 2026

Thinking with Videos from Open-Source Priors. We reproduce chain-of-frames visual reasoning by fine-tuning open-source video models. Give it a star 🌟 if you find it useful.

Python 230 9 Updated Apr 13, 2026
Next