Skip to content
View BackWorld's full-sized avatar

Block or report BackWorld

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Project Lyra: Open Generative 3D World Models

Python 2,232 227 Updated Jul 20, 2026

FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity generation, superior identity co…

Python 1,339 79 Updated Apr 3, 2026

MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and complex real‑world scenario…

Python 3,986 353 Updated Jul 26, 2026

This is the official implementation of our paper: "MiniMax-Remover: Taming Bad Noise Helps Video Object Removal"

Python 593 54 Updated Jul 27, 2025

ComfyUI wrapper for segment anything 3

Python 563 84 Updated Jul 9, 2026

A ComfyUI custom node designed for advanced image background removal and object, face, clothes, and fashion segmentation, utilizing multiple models including RMBG-2.0, INSPYRENET, BEN, BEN2, BiRefN…

Python 2,064 130 Updated Jul 28, 2026

HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation.

Python 1,061 103 Updated Sep 28, 2025

免费AI去水印在线工具汇总:一键去除图片和视频水印

732 41 Updated Apr 15, 2025

基于AI的图片/视频硬字幕去除、文本水印去除,无损分辨率生成去字幕、去水印后的图片/视频文件。无需申请第三方API,本地实现。AI-based tool for removing hard-coded subtitles and text-like watermarks from videos or Pictures.

Python 12,313 1,577 Updated Jun 30, 2026

Open Source Speech Language Model

Jupyter Notebook 1,007 111 Updated May 11, 2026

[CVPR 2026 Highlight] MatAnyone 2: Scaling Video Matting via a Learned Quality Evaluator

Python 814 53 Updated Jul 28, 2026

[CVPR 2025] MatAnyone: Stable Video Matting with Consistent Memory Propagation

Python 1,599 113 Updated Mar 4, 2026

SoTA open-source TTS

Python 25,974 3,474 Updated Jul 21, 2026

A voice conversion extension node for ComfyUI based on FreeVC, enabling high-quality voice conversion capabilities within the ComfyUI framework.

Python 66 2 Updated Apr 3, 2025

FreeVC: Towards High-Quality Text-Free One-Shot Voice Conversion

Python 716 128 Updated Jan 19, 2025

Easily train a good VC model with voice data <= 10 mins!

Python 37,305 5,193 Updated Aug 4, 2026

[TMLR] Memory-Guided Diffusion for Expressive Talking Video Generation

Python 1,070 105 Updated Aug 6, 2025

Official implementation of "Sonic: Shifting Focus to Global Audio Perception in Portrait Animation"

Python 3,269 291 Updated Jan 8, 2026

real time face swap and one-click video deepfake with only a single image

Python 95,894 13,996 Updated Aug 7, 2026

Hackable and optimized Transformers building blocks, supporting a composable construction.

Python 10,542 781 Updated Aug 7, 2026

CUDA integration for Python, plus shiny features

Python 2,049 298 Updated Jul 16, 2026

[CVPR 2026] PersonaLive! : Expressive Portrait Image Animation for Live Streaming

Python 3,519 496 Updated May 15, 2026

SoftVC VITS Singing Voice Conversion

Python 28,138 5,042 Updated Nov 11, 2023

1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

Python 60,792 6,611 Updated Jul 22, 2026

Clone a voice in 5 seconds to generate arbitrary speech in real-time

Python 60,085 9,397 Updated Mar 9, 2026

AIGCPanel 是一个简单易用的一站式AI数字人系统,支持视频合成、声音合成、声音克隆,简化本地模型管理、一键导入和使用AI模型。

TypeScript 5,440 815 Updated Jul 16, 2026

Enhanced the comfyui savevideo node to support previewing and saving videos containing alpha channels.

Python 46 3 Updated Mar 11, 2026
Next