Skip to content
View BRZ911's full-sized avatar
  • China
  • 21:00 (UTC +08:00)

Highlights

  • Pro

Organizations

@MLNLP-World

Block or report BRZ911

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[ACL 2025] The code repository for "Mitigating Visual Forgetting via Take-along Visual Conditioning for Multi-modal Long CoT Reasoning" in PyTorch.

Python 139 Updated May 16, 2025

ACL'2025: SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs. and preprint: SoftCoT++: Test-Time Scaling with Soft Chain-of-Thought Reasoning

Python 94 16 Updated May 30, 2025

[CVPR 2026] Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens

Python 294 17 Updated Aug 2, 2025

Claude Code skill: translate English PDFs to Chinese via pdf2zh, preserving layout/formulas/figures.

Shell 40 6 Updated Jul 6, 2026
Python 2 Updated Jul 21, 2026

Awesome-Long2short-on-LRMs is a collection of state-of-the-art, novel, exciting long2short methods on large reasoning models. It contains papers, codes, datasets, evaluations, and analyses.

262 10 Updated Mar 7, 2026

Latent Video Cache for Video Reasoning

Python 6 Updated Jul 7, 2026

GitHub Pages site for the Next-Generation AI Systems paper

HTML 2 Updated Jun 15, 2026

Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

Python 902 61 Updated Jun 29, 2026

Official codebase for the paper Latent Visual Reasoning

Python 172 10 Updated Oct 22, 2025

Fully open reproduction of DeepSeek-R1

Python 26,428 2,447 Updated Apr 2, 2026

An Elegant Academic Homepage Builder

TypeScript 680 203 Updated May 13, 2026

[ECCV 2026] Demystifying Video Reasoning

Python 47 3 Updated Jul 14, 2026

[ACL 2026] Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning

Python 93 2 Updated Jan 22, 2026

[ICML 2026] Multimodal deep-research MLLM and benchmark. The first long-horizon multimodal deep-research MLLM, extending the number of reasoning turns to dozens and the number of search-engine inte…

Python 670 56 Updated Aug 7, 2026

[AAAI'2026] Let's Think with Images Efficiently! An Interleaved-Modal Chain-of-Thought Reasoning Framework with Dynamic and Precise Visual Thoughts

Python 12 Updated Jun 10, 2026

OMNIFLOW

Python 3 1 Updated Feb 21, 2026

[ICML 2026] Official implementation of "Open-o3 Video: Grounded Video Reasoning with Explicit Spatio-Temporal Evidence"

Python 159 7 Updated May 1, 2026

一个用于在 macOS 上平滑你的鼠标滚动效果或单独设置滚动方向的小工具, 让你的滚轮爽如触控板 | A lightweight tool used to smooth scrolling and set scroll direction independently for your mouse on macOS

Swift 21,074 668 Updated Aug 4, 2026

Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks

Python 4,333 743 Updated Aug 7, 2026

OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.

Python 7,286 828 Updated Aug 4, 2026

A Systematic Survey of Deep Research

322 17 Updated Jan 1, 2026

🔥🔥🔥 Latest Papers, Codes and Datasets on Video-LMM Post-Training

Python 298 14 Updated Mar 3, 2026

We introduce Reasoning via Video, a new paradigm that uses maze-solving video generation to probe multimodal reasoning; our VR-Bench shows that fine-tuned video models consistently outperform stron…

Python 66 6 Updated Feb 4, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,867 4,356 Updated Aug 8, 2026

[ICML 2026 Spotlight] Latent Collaboration in Multi-Agent Systems

Python 1,079 165 Updated Jun 18, 2026

[ECCV 2026] Official repo of "Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens"

Python 390 22 Updated Jul 30, 2026

Lightweight Image Video Action Generation Inference Framework

Python 2,599 249 Updated Aug 8, 2026

This is a collection of recent papers on reasoning in video generation models.

165 6 Updated Jul 30, 2026

Thinking with Videos from Open-Source Priors. We reproduce chain-of-frames visual reasoning by fine-tuning open-source video models. Give it a star 🌟 if you find it useful.

Python 231 9 Updated Apr 13, 2026
Next