Skip to content
View whwu95's full-sized avatar
♥️
I may be slow to respond.
♥️
I may be slow to respond.

Highlights

  • Pro

Block or report whwu95

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Create beautiful slides on the web using a coding agent's frontend skills

JavaScript 26,704 2,165 Updated Jun 23, 2026

An open-source implementaion for fine-tuning Qwen-VL series by Alibaba Cloud.

Python 1,947 220 Updated Jul 25, 2026

Wan: Open and Advanced Large-Scale Video Generative Models

Python 16,709 3,093 Updated Mar 5, 2026

EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

Python 5,095 384 Updated Jul 30, 2026

A Scientific Multimodal Foundation Model

842 48 Updated Jul 17, 2026

[CVPR 2025 Oral] VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection

Python 140 4 Updated Jul 28, 2025

Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL

Python 5,206 466 Updated Nov 13, 2025

Awesome Reasoning in MLLMs: Papers and Projects about learning to reason with MLLMs, including Chain-of-Thought (CoT), OpenAl o1, and DeepSeek-R1

63 4 Updated Mar 18, 2025
TeX 124 54 Updated Jan 29, 2025

[NIPS'25 Spotlight] Mulberry, an o1-like Reasoning and Reflection MLLM Implemented via Collective MCTS

Python 1,244 113 Updated Jan 16, 2026

Efficient Multimodal Large Language Models: A Survey

387 21 Updated Apr 29, 2025

Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.

Jupyter Notebook 19,705 1,838 Updated Jan 30, 2026

A series of math-specific large language models of our Qwen2 series.

Python 1,082 161 Updated Jan 11, 2025

LMDeploy is a toolkit for compressing, deploying, and serving LLMs.

Python 7,982 720 Updated Aug 1, 2026

The official repo of Qwen2-Audio chat & pretrained large audio language model proposed by Alibaba Cloud.

Python 2,097 168 Updated Apr 21, 2025

Retrieval-Augmented Generation in 3 Lines of Code!

Python 54 10 Updated Jul 30, 2026

AudioBench: A Universal Benchmark for Audio Large Language Models

Python 319 15 Updated May 29, 2026

【NeurIPS 2024】The official code of paper "Automated Multi-level Preference for MLLMs"

Python 22 1 Updated Sep 26, 2024

【NeurIPS 2024】Dense Connector for MLLMs

Python 182 8 Updated Oct 14, 2024

FreeVA: Offline MLLM as Training-Free Video Assistant

Python 69 1 Updated Jun 9, 2024

AcadHomepage: A Modern and Responsive Academic Personal Homepage

SCSS 2,877 5,793 Updated Jul 19, 2026

Awesome-LLM-Tabular: a curated list of Large Language Model applied to Tabular Data

428 35 Updated Apr 6, 2026

GPT4Vis: What Can GPT-4 Do for Zero-shot Visual Recognition?

Python 184 18 Updated May 22, 2024

【ICCV'2023】What Can Simple Arithmetic Operations Do for Temporal Modeling?

Python 74 6 Updated Jan 26, 2024

Demonstrate all the questions on LeetCode in the form of animation.(用动画的形式呈现解LeetCode题目的思路,完整单步/回看/变速/语音讲解在 algomooc.com)

Java 76,642 13,891 Updated Jun 12, 2026

Enjoy https://shields.io

Go 461 245 Updated Jul 20, 2026
JavaScript 4,307 1,939 Updated Jun 21, 2024

[ICCV 2023] Official Implementation of "Generalized Lightness Adaptation with Channel Selective Normalization"

Python 86 6 Updated Jan 22, 2024

A curated list of papers and open-source resources focused on 3D AIGC.

348 23 Updated Jul 2, 2026

The largest curated collection of markdown badges for your personal developer branding, profile, and projects.

SCSS 16,874 1,795 Updated Jul 30, 2026
Next