Stars
调用Seedream 系列模型的api服务实现本地生图,包含最新的seedream5.0lite模型。A custom node for ComfyUI to generate images using Volcano Engine's Seedream API.
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Illumination Drawing Tools for Text-to-Image Diffusion Models
Generate detailed image descriptions and analysis using Molmo models in ComfyUI.
[NeurIPS 2024] Depth Anything V2. A More Capable Foundation Model for Monocular Depth Estimation
Hunyuan-DiT : A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
MiniCPM5-1B: A SOTA 1B on-device LLM, small yet powerful.
Drop in a screenshot and convert it to clean code (HTML/Tailwind/React/Vue)
A visual no-code/code-free web crawler/spider易采集:一个可视化浏览器自动化测试/数据采集/网页爬虫软件,可以无代码图形化的设计和执行爬虫任务。别名:ServiceWrapper面向Web应用的智能化服务封装系统。
[SIGGRAPH Asia 2024, Journal Track] ToonCrafter: Generative Cartoon Interpolation
MusePose: a Pose-Driven Image-to-Video Framework for Virtual Human Generation
SD-Trainer. LoRA & Dreambooth training scripts & GUI use kohya-ss's trainer, for diffusion model.
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
The code releasing for https://image-dream.github.io/
[ECCV 2024] Single Image to 3D Textured Mesh in 10 seconds with Convolutional Reconstruction Model.
Industry leading face manipulation platform
Stable Diffusion web UI
Emote Portrait Alive: Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
Outfit Anyone: Ultra-high quality virtual try-on for Any Clothing and Any Person