| Wan: Open and advanced large-scale video generative models T Wan, A Wang, B Ai, B Wen, C Mao, CW Xie, D Chen, F Yu, H Zhao, ... arXiv preprint arXiv:2503.20314, 2025 | 1923 | 2025 |
| Modelscope text-to-video technical report J Wang, H Yuan, D Chen, Y Zhang, X Wang, S Zhang arXiv preprint arXiv:2308.06571, 2023 | 861 | 2023 |
| Videocomposer: Compositional video synthesis with motion controllability X Wang, H Yuan, S Zhang, D Chen, J Wang, Y Zhang, Y Shen, D Zhao, ... Advances in Neural Information Processing Systems 36, 7594-7611, 2023 | 640 | 2023 |
| Videofusion: Decomposed diffusion models for high-quality video generation Z Luo, D Chen, Y Zhang, Y Huang, L Wang, Y Shen, D Zhao, J Zhou, ... arXiv preprint arXiv:2303.08320, 2023 | 528 | 2023 |
| I2vgen-xl: High-quality image-to-video synthesis via cascaded diffusion models S Zhang, J Wang, Y Zhang, K Zhao, H Yuan, Z Qin, X Wang, D Zhao, ... arXiv preprint arXiv:2311.04145, 2023 | 409 | 2023 |
| Dream video: Composing your dream videos with customized subject and motion Y Wei, S Zhang, Z Qing, H Yuan, Z Liu, Y Liu, Y Zhang, J Zhou, H Shan 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR …, 2024 | 255 | 2024 |
| Timestep embedding tells: It’s time to cache for video diffusion model F Liu, S Zhang, X Wang, Y Wei, H Qiu, Y Zhao, Y Zhang, Q Ye, F Wan 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR …, 2025 | 253 | 2025 |
| Visual search at alibaba Y Zhang, P Pan, Y Zheng, K Zhao, Y Zhang, X Ren, R Jin Proceedings of the 24th ACM SIGKDD international conference on knowledge …, 2018 | 194 | 2018 |
| Unianimate: Taming unified video diffusion models for consistent human image animation X Wang, S Zhang, C Gao, J Wang, X Zhou, Y Zhang, L Yan, N Sang Science China Information Sciences 68 (10), 200103, 2025 | 151 | 2025 |
| Molo: Motion-augmented long-short contrastive learning for few-shot action recognition X Wang, S Zhang, Z Qing, C Gao, Y Zhang, D Zhao, N Sang 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR …, 2023 | 144 | 2023 |
| Dreamtalk: When expressive talking head generation meets diffusion probabilistic models Y Ma, S Zhang, J Wang, X Wang, Y Zhang, Z Deng arXiv preprint arXiv:2312.09767 2 (3), 2023 | 143 | 2023 |
| Clip-guided prototype modulating for few-shot action recognition X Wang, S Zhang, J Cen, C Gao, Y Zhang, D Zhao, N Sang International Journal of Computer Vision 132 (6), 1899-1912, 2024 | 136 | 2024 |
| Videolcm: Video latent consistency model X Wang, S Zhang, H Zhang, Y Liu, Y Zhang, C Gao, N Sang arXiv preprint arXiv:2312.09109, 2023 | 132 | 2023 |
| Instructvideo: Instructing video diffusion models with human feedback H Yuan, S Zhang, X Wang, Y Wei, T Feng, Y Pan, Y Zhang, Z Liu, ... 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR …, 2024 | 123 | 2024 |
| Distribution adaptive int8 quantization for training cnns K Zhao, S Huang, P Pan, Y Li, Y Zhang, Z Gu, Y Xu Proceedings of the AAAI conference on artificial intelligence 35 (4), 3483-3491, 2021 | 120 | 2021 |
| Eflops: Algorithm and system co-design for a high performance distributed training platform J Dong, Z Cao, T Zhang, J Ye, S Wang, F Feng, L Zhao, X Liu, L Song, ... 2020 IEEE International Symposium on High Performance Computer Architecture …, 2020 | 98 | 2020 |
| A recipe for scaling up text-to-video generation with text-free videos X Wang, S Zhang, H Yuan, Z Qing, B Gong, Y Zhang, Y Shen, C Gao, ... 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR …, 2024 | 87 | 2024 |
| DecentLaM: Decentralized momentum SGD for large-batch deep training K Yuan, Y Chen, X Huang, Y Zhang, P Pan, Y Xu, W Yin 2021 IEEE/CVF International Conference on Computer Vision (ICCV), 3009-3019, 2021 | 83 | 2021 |
| Hierarchical spatio-temporal decoupling for text-to-video generation Z Qing, S Zhang, J Wang, X Wang, Y Wei, Y Zhang, C Gao, N Sang 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR …, 2024 | 82 | 2024 |
| Disentangling spatial and temporal learning for efficient image-to-video transfer learning Z Qing, S Zhang, Z Huang, Y Zhang, C Gao, D Zhao, N Sang 2023 IEEE/CVF International Conference on Computer Vision (ICCV), 13888-13898, 2023 | 71 | 2023 |