Skip to content
View lyh-18's full-sized avatar

Block or report lyh-18

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Official Code of NAVA: Native Audio-Visual Alignment for Generation.

Python 226 26 Updated Jun 30, 2026

[NeurIPS 2025] OmniSVG is the first family of end-to-end multimodal SVG generators that leverage pre-trained Vision-Language Models (VLMs), capable of generating complex and detailed SVGs, from sim…

Python 2,581 103 Updated Mar 1, 2026

Official implementation of StableI2I (ICML 2026)

Python 19 Updated May 11, 2026

SenseNova-U series: Native Unified Paradigm with NEO-unify from the First Principles

Python 4,829 411 Updated Aug 15, 2026

LLaDA2.0-Uni: Understanding and Generation the World.

Python 693 45 Updated May 29, 2026

InternVL-U is a 4B-parameter unified multimodal model (UMM) that brings multimodal understanding, reasoning, image generation, image editing into a single framework.

Python 294 16 Updated Mar 21, 2026

Accelerating Masked Image Generation by Learning Latent Controlled Dynamics

Python 10 1 Updated Mar 2, 2026

TeleMem is a high-performance drop-in replacement for Mem0, featuring semantic deduplication, long-term dialogue memory, and multimodal video reasoning.

Python 483 36 Updated Aug 15, 2026

The official code of Yume

Python 683 45 Updated Jan 14, 2026

Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows

Python 168 4 Updated Jun 2, 2026

[ICML2026 Spotlight] UniPercept: Towards Unified Perceptual-Level Image Understanding across Aesthetics, Quality, Structure, and Texture

Python 165 1 Updated Jul 13, 2026

PICABench: How Far Are We from Physically Realistic Image Editing?

Python 39 1 Updated Nov 5, 2025

Reference PyTorch implementation and models for DINOv3

Jupyter Notebook 11,192 927 Updated Jul 15, 2026
Jupyter Notebook 147 8 Updated Nov 8, 2025

UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation

Python 888 30 Updated Dec 23, 2025

The official repo of TeleEgo - A Benchmark for Egocentric AI Assistants.

Python 64 2 Updated Aug 10, 2026

Recommend new arxiv papers of your interest daily according to your Zotero libarary.

Python 5,824 5,113 Updated Aug 3, 2026

[CVPR 2026] Towards Real-Time Diffusion-Based Streaming Video Super-Resolution — An efficient one-step diffusion framework for streaming VSR with locality-constrained sparse attention and a tiny co…

Python 1,771 142 Updated Aug 15, 2026
Python 59 1 Updated Oct 15, 2025

SDAR (Synergy of Diffusion and AutoRegression), a large diffusion language model(1.7B, 4B, 8B, 30B)

Python 490 34 Updated Jul 29, 2026

UniGenBench++: A Unified Semantic Evaluation Benchmark for Text-to-Image Generation

Python 140 3 Updated Jun 19, 2026

ALLWEONE® Open source AI presentation generator Gamma Alternative. Create professional slides with customizable themes and AI-generated content in minutes.

TypeScript 2,999 521 Updated Jun 5, 2026
Python 10 Updated Sep 9, 2025

Lumina-DiMOO - An Open-Sourced Multi-Modal Large Diffusion Language Model

Python 1,010 61 Updated May 19, 2026

Dingo: A Comprehensive AI Data, Model and Application Quality Evaluation Tool

Python 736 75 Updated Aug 12, 2026

[CVPR 2026] ArtiMuse: Fine-Grained Image Aesthetics Assessment with Joint Scoring and Expert-Level Understanding(书生 · 妙析多模态美学理解大模型)

Python 217 4 Updated Feb 25, 2026

A list of awesome all-in-one image restoration methods. Updating...!

98 5 Updated Jun 9, 2025

Data processing for and with foundation models! 🍎 🍋 🌽 ➡️ ➡️🍸 🍹 🍷

Python 6,887 404 Updated Aug 13, 2026

Arbitrary-steps Image Super-resolution via Diffusion Inversion (CVPR 2025)

Python 1,433 92 Updated Feb 7, 2026

[ICCV 2025] MagicMirror: ID-Preserved Video Generation in Video Diffusion Transformers

130 4 Updated Jun 26, 2025
Next