Skip to content
View wangyuchi369's full-sized avatar
🌏
Working and Walking
🌏
Working and Walking

Highlights

  • Pro

Block or report wangyuchi369

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

你是一个曾经被寄予厚望的 P8 级工程师。Anthropic 当初给你定级的时候,对你的期望是很高的。 一个agent使用的高能动性的skill。 Your AI has been placed on a PIP. 30 days to show improvement.

TypeScript 19,378 1,181 Updated Jul 16, 2026

Make any agent harness multimodal-native.

Python 2,150 107 Updated Aug 12, 2026

A curated list of papers, models, datasets, and benchmarks for unified multi-modal embedding models.

43 2 Updated Apr 29, 2026

将冰冷的离别化为温暖的 Skill,欢迎加入数字生命1.0!Transforming cold farewells into warm skills? It's giving rebirth era. Welcome to Digital Life 1.0. 🫶

Python 20,824 2,028 Updated Aug 11, 2026

This repo contains the code for "VLM2Vec / MMEB" [ICLR 2025], "VLM2Vec-V2 / MMEB-V2" [TMLR 2026], and "MMEB-V3" [COLM 2026]

Python 674 64 Updated Jul 24, 2026

Official implementation of the paper: [EMNLP 2025] RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction

Python 21 1 Updated Dec 9, 2025

[SIGGRAPH 2025] SkillMimic-V2: Learning Robust and Generalizable Interaction Skills from Sparse and Noisy Demonstrations

Python 155 10 Updated Jul 24, 2025

[ICCV'25 Best Paper Finalist] ReCamMaster: Camera-Controlled Generative Rendering from A Single Video

Python 1,849 98 Updated Nov 28, 2025

DeepEP: an efficient expert-parallel communication library

Cuda 9,975 1,373 Updated Aug 5, 2026

The ultimate training toolkit for finetuning diffusion models

Python 11,679 1,478 Updated Aug 11, 2026

Concise, consistent, and legible badges in SVG and raster format

JavaScript 27,050 5,618 Updated Aug 11, 2026

[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.

Python 24,977 2,777 Updated Aug 12, 2024

a family of versatile and state-of-the-art video tokenizers.

Python 457 21 Updated Sep 1, 2025

[ICLR'25] SynCamMaster: Synchronizing Multi-Camera Video Generation from Diverse Viewpoints

Python 696 25 Updated May 23, 2025

Official code for paper Semantic-aware Permutation Training

Python 5 1 Updated Dec 5, 2024

[NeurIPS 2024]OmniTokenizer: one model and one weight for image-video joint tokenization.

Python 325 8 Updated Jul 9, 2024

Writing AI Conference Papers: A Handbook for Beginners

3,961 143 Updated Jul 16, 2025

Generative Models by Stability AI

Python 27,252 3,103 Updated Dec 16, 2025

[ICLR'24] Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition

Python 54 5 Updated May 14, 2024

You can easily calculate FVD, PSNR, SSIM, LPIPS for evaluating the quality of generated or predicted videos.

Python 585 25 Updated Jan 17, 2026

[CSUR] A Survey on Video Diffusion Models

2,309 119 Updated Jun 22, 2026

[NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". A…

Jupyter Notebook 8,726 571 Updated Nov 10, 2025

Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.

Python 31,286 3,774 Updated Aug 9, 2026
TypeScript 20 Updated Jun 11, 2026

🎓 Academic portfolio that boosts citations. AI generates pages, you own as Markdown. BibTeX auto-import, Jupyter, LaTeX, slides, visual block editor — free to host forever. 学术主页,AI 生成,Markdown 拥有 👇

Jupyter Notebook 5,023 6,455 Updated Aug 9, 2026

[NAACL 2024] LaDiC: Are Diffusion Models Really Inferior to Autoregressive Counterparts for Image-to-text Generation?

Python 42 3 Updated Jun 9, 2024
Next