Skip to content
View ncTimTang's full-sized avatar

Highlights

  • Pro

Block or report ncTimTang

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

The official repository for ICML 2026 paper "Reasoning LLM Improves Speaker Recognition in Long-form TV Dramas"

Python 11 1 Updated Jul 3, 2026

About This repository is a curated collection of the most exciting and influential CVPR 2026 papers. 🔥 [Paper + Code + Demo]

Python 568 31 Updated Jun 6, 2026
JavaScript 1 Updated Aug 5, 2026

🔥🔥🔥 [Awesome] Latest Papers, Codes & Datasets on Streaming / Online Video Understanding — Building Always-on, Real-time Video AI 🤖

431 27 Updated Aug 5, 2026

A simple video streaming baseline that outperforms SOTAs.

Python 157 8 Updated May 1, 2026

StreamingVLM: Real-Time Understanding for Infinite Video Streams

Python 1,063 65 Updated Oct 15, 2025

Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding

Python 369 4 Updated Aug 5, 2026

[CVPR 2026] LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding

Python 51 2 Updated Jul 7, 2026
JavaScript 5 3 Updated May 4, 2026

vHeat: Building Vision Models upon Heat Conduction

Python 285 11 Updated Jun 12, 2025

[ICML 2026 Spotlight] Code for miXed Discrete Diffusion Language Model

Python 29 1 Updated Mar 16, 2026

LLaDA Continue Pretraining with XDLM

Python 6 Updated Feb 11, 2026

[ICLR 2026] VideoAnchor: Reinforcing Subspace-Structured Visual Cues for Coherent Visual-Spatial Reasoning

Python 5 Updated Feb 28, 2026

Official repo of From Indoor to Open World: Revealing the Spatial Reasoning Gap in MLLMs

Python 24 1 Updated Jun 23, 2026

[ICLR 2026] An official implementation of "CapRL: Stimulating Dense Image Caption Capabilities via Reinforcement Learning"

Python 227 8 Updated Jun 23, 2026

verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

Python 2,192 208 Updated Jun 9, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 73,920 9,042 Updated Aug 6, 2026

A comprehensive and up-to-date compilation of datasets, tools, methods, review papers, and competitions for remote sensing change detection.

2,295 404 Updated Apr 16, 2026

Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.

Jupyter Notebook 19,747 1,831 Updated Jan 30, 2026

Reinforcement Learning of Vision Language Models with Self Visual Perception Reward

Python 180 18 Updated Mar 14, 2026

REverse-Engineered Reasoning for Open-Ended Generation

Python 98 7 Updated Sep 10, 2025

personal homepage of tangxi

HTML 1 Updated Apr 18, 2025

Fast and memory-efficient exact attention

Python 24,655 2,970 Updated Aug 7, 2026

EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

Python 5,105 383 Updated Jul 30, 2026

Official Repo for Open-Reasoner-Zero

Python 2,098 121 Updated Jun 2, 2025

Solve Visual Understanding with Reinforced VLMs

Python 6,017 385 Updated Jul 7, 2026

This repository provides valuable reference for researchers in the field of multimodality, please start your exploratory travel in RL-based Reasoning MLLMs!

1,439 68 Updated Aug 2, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,870 4,359 Updated Aug 8, 2026

The official implementation of the **ReDDiT: Rehashing Noise for Discrete Visual Generation** paper.

Python 12 Updated Sep 27, 2025

[ICLR 2026] Geometric-Mean Policy Optimization

Python 104 11 Updated Jan 26, 2026
Next