Skip to content
View ziwei-cui's full-sized avatar
✌️
✌️
  • HuaZhong University of Science and Technology
  • WuHan

Block or report ziwei-cui

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

MAGI-1: Autoregressive Video Generation at Scale

Python 3,760 238 Updated Jun 17, 2026

[CVPR 2025] GaussTR: Foundation Model-Aligned Gaussian Transformer for Self-Supervised 3D Spatial Understanding

Python 218 13 Updated Jan 5, 2026

The first decoder-only multimodal state space model

Python 104 6 Updated May 19, 2025

Pathology Foundation Model - Nature Medicine

Jupyter Notebook 764 91 Updated Mar 26, 2025
Python 277 34 Updated Apr 17, 2025

finetuning SAM with non-promptable decoder on medical images

Python 140 12 Updated Jul 18, 2023

From a video, automatically create an Instance Segmentation dataset using Detectors like YoloX and Segment Anything

Python 1 Updated Apr 7, 2024

[MedIA'25] UN-SAM: Domain-Adaptive Self-Prompt Segmentation for Universal Nuclei Images

Python 93 8 Updated Jun 25, 2025
Python 225 10 Updated Jun 24, 2024

[AAAI 2025] Linear-complexity Visual Sequence Learning with Gated Linear Attention

Python 117 1 Updated Jun 17, 2024

[ICLR 2025] This is the official repository of our paper "MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine“

Python 411 31 Updated Jul 11, 2025

Segment Anything in Medical Images

Jupyter Notebook 4,371 596 Updated May 7, 2025

The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.

Jupyter Notebook 54,660 6,354 Updated Sep 18, 2024

[KBS'25] NuSegDG: Integration of Heterogeneous Space and Gaussian Kernel for Domain-Generalized Nuclei Segmentation

Python 31 5 Updated Aug 29, 2024

[CVPR 2025 Highlight] Truncated Diffusion Model for Real-Time End-to-End Autonomous Driving

Python 1,470 145 Updated Dec 8, 2025

[ECCV 2024] Code for "Unleashing the Power of Prompt-driven Nucleus Instance Segmentation"

Python 60 7 Updated Jan 9, 2025

Bridging Large Vision-Language Models and End-to-End Autonomous Driving

Python 552 49 Updated Mar 15, 2026

Swin-LiteMedSAM: A Lightweight Box-Based Segment Anything Model for Large-Scale Medical Image Datasets

Python 19 Updated Mar 31, 2025

PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.

Python 7,401 1,295 Updated Mar 16, 2026

The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use th…

Jupyter Notebook 19,668 2,524 Updated May 30, 2026

[arXiv '24] Efficient Cell Nuclei Instance Segmentation with Large Convolution Kernels

Python 48 7 Updated Aug 28, 2024
Jupyter Notebook 256 15 Updated Jun 19, 2026

Multi-modality pre-training

Python 511 36 Updated Mar 27, 2026

Code for CVPR25 paper "VideoTree: Adaptive Tree-based Video Representation for LLM Reasoning on Long Videos"

Python 167 15 Updated Jun 23, 2025

Video Highlight generation using short time analysis and keyframe algorithm

Python 3 Updated Jul 14, 2024

[ECCV 2024🔥] Official implementation of the paper "ST-LLM: Large Language Models Are Effective Temporal Learners"

Python 155 7 Updated Sep 10, 2024

Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.

Python 27,485 2,030 Updated Jan 9, 2026

mllm-npu: training multimodal large language models on Ascend NPUs

Python 95 2 Updated Aug 29, 2024

Official Repository of paper VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding

Python 293 19 Updated Aug 5, 2025
Next