Stars
Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation
JamePeng / llama-cpp-python
Forked from abetlen/llama-cpp-pythonPython bindings for llama.cpp
[ICML 2026] ByteDance's All-in-One Video Generation Model for Human-Object Interaction Video Generation
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
Fast LLM speculative inference server for consumer hardware.
A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.
首家工业级全流程 AI 影视生产平台。Industry-first professional AI Agent platform for controllable film & video production. From shorts to live-action with Hollywood-standard workflows.
An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of…
Comprehensive production pipeline for quad-modal AI filmmaking with Seedance 2.0
Convert 2D videos to 3D VR format using AI depth estimation.
PixelFlow is a professional-grade AI visual creation engine that reimagines image generation through a node-based workflow. Designed for concept artists, designers, and prompt engineers, it leverag…
Professional Antigravity Account Manager & Switcher. One-click seamless account switching for Antigravity Tools. Built with Tauri v2 + React (Rust).专业的 Antigravity 账号管理与切换工具。为 Antigravity 提供一键无缝账号切…
Self-hosted multi-protocol AI API proxy for Antigravity, Codex, Grok, Kiro, OpenAI, Claude, and custom providers. Supports OpenAI-compatible API, Claude API, Gemini protocol conversion, GPT, Grok B…
All available LTX-2 models, encoders, workflows, LoRAs for ComfyUI
Emote Portrait Alive: Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
OpenGPT 4o is a free alternative to OpenAI GPT 4o
Archived — A list of games, add-ons, maps, etc. hosted on GitHub. Any genre. Any platform. Any engine.
Comflowyspace is an intuitive, user-friendly, open-source AI tool for generating images and videos, democratizing access to AI technology.
Controllable video and image Generation, SVD, Animate Anyone, ControlNet, ControlNeXt, LoRA
real time face swap and one-click video deepfake with only a single image
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
[ICLR 2025] CatVTON is a simple and efficient virtual try-on diffusion model with 1) Lightweight Network (899.06M parameters totally), 2) Parameter-Efficient Training (49.57M parameters trainable) …
DeepFuze is a state-of-the-art deep learning tool that seamlessly integrates with ComfyUI to revolutionize facial transformations, lipsyncing, Face Swapping, Lipsync Translation, video generation, …
Open source real-time translation app for Android that runs locally
Curated list of chatgpt prompts from the top-rated GPTs in the GPTs Store. Prompt Engineering, prompt attack & prompt protect. Advanced Prompt Engineering papers.