Introduction
Welcome to Flaq.ai - your gateway to cutting-edge AI image, video, and LLM models.
What is Flaq.ai
Flaq.ai is a powerful AI generation platform that provides unified access to state-of-the-art image, video, and LLM models from leading AI research labs including Google, ByteDance, Alibaba, and DeepSeek.
Our platform simplifies the integration of advanced AI models into your applications through a consistent, easy-to-use API interface. Whether you're building creative tools, automating content generation, or exploring AI capabilities, Flaq.ai provides the infrastructure you need.
Key Features
- Unified API: Consistent API interface across all models for seamless integration
- High-Quality Output: Generate professional-grade images and videos with advanced AI technology
- Flexible Parameters: Fine-tune generation with aspect ratios, styles, and model-specific controls
- Fast Processing: Optimized infrastructure for quick generation and delivery
- Developer-Friendly: Comprehensive documentation, code examples, and SDKs
Supported Models
Image Generation Models
-
Nano Banana Pro (Google Gemini 3.0 Pro Image)
Premier AI-powered visual generation with native 4K output, context-aware understanding, and multilingual typography. -
Nano Banana Pro Edit (Google Gemini 3.0 Pro Image Edit)
Advanced image editing capabilities with precise control over modifications and enhancements. -
Nano Banana 2 (Google Gemini 3.1 Flash Image)
Lightning-fast AI image generation optimized for cost-effectiveness and high-performance. Benchmarked against Nano Banana Pro with exceptional speed while maintaining professional quality. -
Nano Banana 2 Edit (Google Gemini 3.1 Flash Image Edit)
Fast and efficient image editing with natural language commands, offering rapid iteration and precise adjustments at optimized cost. -
Seedream 4.5 (ByteDance)
Fast, stylized generation with strong anime and illustration aesthetics. -
Seedream 5.0 (ByteDance)
Next-generation image generation with enhanced prompt adherence, superior detail rendering, and improved aesthetic quality. -
Seedream 5.0 Edit (ByteDance)
Advanced image editing powered by Seedream 5.0, supporting precise instruction-based modifications. -
Seedream 5.0 Pro (ByteDance)
Premium text-to-image generation with enhanced prompt adherence, superior detail rendering, and support for eight aspect ratios including 21:9. -
Seedream 5.0 Pro Edit (ByteDance)
High-fidelity image editing powered by Seedream 5.0 Pro, supporting instruction-based modifications with 1–10 reference images. -
Qwen Image 2.0 (Alibaba)
High-performance image generation with sharp text rendering and realistic visuals. Optimized for stable, high-quality output at competitive pricing. -
Qwen Image 3.0 (Alibaba)
High-quality text-to-image generation with strong prompt adherence and flexible composition control. -
Qwen Image 3.0 Edit (Alibaba)
Instruction-based image editing with multi-image input support and precise visual refinement. -
Qwen Image 3.0 Pro (Alibaba)
Professional text-to-image generation with knowledge-rich composition and multilingual text rendering. -
Qwen Image 3.0 Pro Edit (Alibaba)
High-fidelity image editing with multi-image composition, instruction-based modifications, and consistent visual refinement. -
Qwen Image Lora Edit (Alibaba)
Prompt-guided image editing for background changes, style updates, object edits, and polished production-ready visual refinements. -
Wan 2.7 Image Pro (Alibaba)
Premium text-to-image generation from Alibaba's Wan 2.7 series, delivering enhanced visual quality with support for seven aspect ratios including 1:1, 16:9, 9:16, 4:3, 3:4, 3:2, and 2:3. -
Wan 2.7 Image Pro Edit (Alibaba)
High-fidelity image-to-image editing powered by Wan 2.7 Image Pro Edit, supporting instruction-based modifications with 1–9 reference images for precise visual refinement. -
Wan 2.7 Image (Alibaba)
Cost-effective text-to-image generation from the Wan 2.7 series, ideal for high-volume creative workflows with flexible aspect ratio control. -
Wan 2.7 Image Edit (Alibaba)
Instruction-based image editing with multi-image input support, enabling efficient visual modifications at competitive pricing. -
Z Image (Alibaba)
Affordable text-to-image generation for marketing visuals, product imagery, social assets, and creative automation workflows. -
Grok Imagine (xAI)
Image generation powered by xAI's Grok, delivering creative and diverse visual outputs. -
Grok Imagine Edit (xAI)
Instruction-based image editing with Grok's understanding of complex natural language commands. -
GPT Image 2 (OpenAI)
High-quality text-to-image generation with strong prompt adherence, flexible quality controls, and reliable typography rendering for production workflows. -
GPT Image 2 Edit (OpenAI)
Prompt-driven image editing with support for multi-image inputs, context-aware modifications, and consistent visual preservation across complex edits. -
GPT Image 2 Client (OpenAI)
OpenAI GPT Image 2 generation access through the client variant, offering the same core text-to-image workflow for teams standardizing on this route. -
GPT Image 2 Edit Client (OpenAI)
Client-route access to GPT Image 2 editing, supporting instruction-based edits, multi-image composition, and flexible creative refinement workflows.
Video Generation Models
-
Veo 3.1 Text to Video (Google)
High-quality video generation from text prompts with advanced motion synthesis. -
Veo 3.1 Image to Video (Google)
Transform static images into dynamic videos with sophisticated motion effects. -
Veo 3.1 Fast Text to Video (Google)
Rapid video generation from text prompts, optimized for processing speed. -
Veo 3.1 Fast Image to Video (Google)
Faster image-to-video conversion with reduced processing time. -
Wan 2.6 Text to Video (Alibaba)
Advanced text-to-video synthesis delivering high-quality outputs. -
Wan 2.6 Image to Video (Alibaba)
Sophisticated image-to-video transformation featuring complex motion and scene composition. -
Wan 2.7 Text to Video (Alibaba)
Next-generation text-to-video synthesis with improved motion quality and prompt adherence over Wan 2.6. -
Wan 2.7 Image to Video (Alibaba)
Advanced image-to-video transformation with enhanced scene dynamics and temporal consistency. -
Wan 2.7 Lora (Alibaba)
Advanced Wan 2.7 animation workflows for turning still visuals into cinematic motion with prompt-guided camera movement and stable subject preservation. -
Wan 2.7 Video Edit (Alibaba)
AI-powered video editing that applies text-guided modifications to existing video content. -
MiniMax H3 (MiniMax)
Video generation family supporting text-to-video, start-and-end-frame image-to-video, and multimodal reference-to-video workflows with image, video, and audio inputs. -
Happy Horse 1.0 (Alibaba)
High-quality video generation model supporting both text-to-video and image-to-video workflows, with flexible clip creation for creative and production use cases. -
Seedance 1.5 Pro (ByteDance)
Native audio-visual sync video model with multilingual lip-sync capabilities and cinematic quality for professional video production. -
Seedance 2.0 (ByteDance)
Next-generation audio-visual video model with standard and fast variants, supporting text-to-video, image-to-video, and reference-to-video generation modes. -
Seedance 2.5 (ByteDance)
Video generation family supporting text-to-video, first-frame image-to-video with optional end-frame guidance, and reference-to-video workflows with video, image, and audio inputs. -
Kling 3.0 (Kuaishou)
Efficient video generation with solid physical coherency and smooth scene transitions, ideal for dynamic content creation. -
Kling 3.0 Turbo (Kuaishou)
Fast Kling video generation for text-to-video and image-to-video workflows with optional sound, designed for quick creative iteration and production use. -
Kling Video O3 (Kuaishou)
Advanced reasoning-enhanced video model supporting text-to-video, image-to-video, reference-to-video, and video editing in both standard and pro tiers. -
Vidu Q3 (Vidu)
High-quality video generation with turbo and pro variants, supporting text-to-video, image-to-video, and start-end frame control. -
Pixverse C1 (PixVerse)
Text-to-video, first-frame image-to-video, and first-and-last-frame transition generation with flexible resolution and duration controls. -
Pixverse V6 (PixVerse)
Text-to-video, image-to-video, transition, and video extend workflows with optional negative prompts and seeding on transition and extend variants. -
Grok Imagine Video (xAI)
AI video generation by xAI, supporting both text-to-video and image-to-video creation. -
Grok Imagine Video 1.5 (xAI)
Next-generation image-to-video animation by xAI with expanded aspect ratio support and flexible duration options. -
Video Upscaler (Flaq AI)
Accessible video enhancement for improving the resolution and visual quality of existing footage through a simple video-to-video workflow. -
Video Upscaler Pro (Flaq AI)
Quality-focused video enhancement for refining existing footage across professional media and production workflows. -
Happy Horse 1.1 (Alibaba)
Versatile video generation for text-to-video, image-to-video, and reference-guided creative workflows. -
Seedance 2.0 Mini (ByteDance)
Cost-effective video generation for efficient text-to-video, image-to-video, and reference-to-video workflows. -
FLUX 3 (Black Forest Labs)
Video generation family supporting text-to-video, image-to-video, start-and-end-frame transitions, and video extension with native audio.
LLM Models
-
GPT 5.4 (OpenAI)
Fast and affordable OpenAI LLM API access for text-to-text, image-to-text, web search, and file analysis workflows. -
GPT 5.5 (OpenAI)
Advanced OpenAI LLM access for reasoning, coding help, multimodal understanding, web-aware answers, and deep document analysis. -
GPT 5.6 Sol (OpenAI)
Flagship OpenAI LLM access for advanced reasoning, coding, writing, multimodal understanding, web-aware answers, and professional file analysis workflows. -
GPT 5.6 Terra (OpenAI)
Balanced OpenAI LLM access for dependable reasoning, writing, coding help, visual understanding, web research, and cost-effective document analysis. -
GPT 5.6 Luna (OpenAI)
Fast and affordable OpenAI LLM access for high-throughput chat, writing, coding, image understanding, web search, and document analysis workflows. -
Claude Sonnet 4.6 (Anthropic)
Balanced Claude LLM API for fast reasoning, writing, coding help, and file analysis with stable production access. -
Claude Sonnet 5 (Anthropic)
Latest balanced Claude LLM API for fast reasoning, writing, coding help, and file analysis with stable production access. -
Claude Opus 4.6 (Anthropic)
High-capability Claude Opus model for careful reasoning, complex writing, coding support, and document review workflows. -
Claude Opus 4.7 (Anthropic)
Latest Claude Opus LLM for advanced reasoning, technical analysis, file understanding, and professional AI applications. -
Claude Opus 4.8 (Anthropic)
Most capable Claude Opus model for complex reasoning, deep technical analysis, file and image understanding, and advanced AI workflows. -
Claude Opus 5 (Anthropic)
Advanced Claude Opus LLM for demanding reasoning, writing, coding, and file or image analysis through text-to-text and file-analysis variants. -
Claude Fable 5 (Anthropic)
Mythos-class Claude LLM for long-horizon reasoning, agentic coding, web-aware research, and file or image analysis through text-to-text, web search, and file analysis variants. -
Gemini 3.5 Flash (Google)
Fast and efficient Google LLM for text generation, image understanding, and file analysis with low-latency responses. -
Qwen 3.7 (Alibaba)
Alibaba's reasoning LLM with Max and Plus tiers, supporting text generation and web-search-augmented answers. -
Qwen 3.8 Max (Alibaba)
Max-tier Alibaba LLM for advanced text reasoning and web-search-augmented answers, with streaming and flexible generation controls. -
Qwen Character (Alibaba)
Alibaba character roleplay LLMs with Plus and Flash tiers, supporting persona-driven chat, reusable profiles, memory-aware dialogue, and scalable storytelling workflows. -
DeepSeek v4 (DeepSeek)
DeepSeek LLM access with Pro and Flash tiers for text generation, reasoning, coding help, and web-search-augmented answers for current-information workflows. -
Kimi 2.7 (Moonshot AI)
Moonshot AI LLM access for text reasoning and image-to-text understanding, supporting writing, coding, visual analysis, OCR-style extraction, and multimodal assistant workflows. -
GLM 5.2 (Z.ai)
Advanced Z.ai LLM for practical reasoning, coding help, structured text generation, and agent-oriented engineering workflows through a production-ready text-to-text API. -
Grok 4 (xAI)
High-performance xAI LLM for text reasoning and image understanding with strong analytical capabilities. -
Grok 4.5 (xAI)
Advanced xAI LLM for text generation, reasoning, coding assistance, and multimodal image understanding across analytical and conversational workflows. -
Kimi K3 (Moonshot AI)
Moonshot AI text-only LLM access for long-horizon coding, reasoning, knowledge work, and scalable agent workflows.