AI Models
Compare AI image models and AI video models available on FireRed Image Edit, then open the best model page or generator for your workflow.
AI Image Models
FireRed Image Edit
An AI image editing model built for targeted inpainting, object replacement, background modification, and style refinement. It is ideal for high-control edits that preserve the original composition while improving visual quality, especially for e-commerce retouching, poster updates, social media assets, and creative rework.
Nano Banana 2
Fast and affordable image generation model powered by Google Gemini Flash Image. Ideal for everyday creative tasks with flexible resolutions from 0.5K to 4K.
Nano Banana 2 Lite
Google DeepMind’s fastest and most efficient Gemini Image model for low-latency image generation and editing. Use it when you need rapid drafts, reference-guided edits, and reliable 1K visual exploration at a lighter credit cost.
Nano Banana Pro
High-quality image generation with enhanced details and better consistency
Seedream 5.0 Lite
ByteDance's next-generation lightweight image generation model that delivers flagship-level visual quality at significantly faster speeds. Supports batch generation of up to 15 images with strong prompt understanding, ideal for high-frequency creative workflows.
GPT-Image 1.5
Professional-grade photorealistic image generation model, specializing in capturing delicate light/shadow textures and natural human expressions. Supports both text-to-image and image-to-image modes, with excellent performance in fidelity and detail restoration, particularly suitable for commercial scenarios requiring high realism.
GPT Image 2
A multi-reference image generation model that supports up to 16 reference images, built for high consistency, stronger control, and more complex commercial workflows. Supports both text-to-image and image-to-image generation.
Seedream 4.5
Flagship visual generation model with ultimate natural language understanding and light/shadow expression. It no longer relies on mechanical tag stacking, but understands 'stories' and 'atmosphere', switching freely between hyper-realistic photography and surreal art.
FLUX.2
Next-generation industrial-grade model, breaking the 'text curse' of AI drawing. It has the strongest instruction following capability to date, not only perfectly generating specified words/text within images, but also precisely controlling complex hand movements and object spatial relationships.
Z-Image
A model optimized for text-to-image scenarios, dedicated to accurately mapping complex descriptions to visual elements, especially excelling in conceptualization and abstract description realization.
Grok 4.2 Image
xAI's latest Grok 4.2 image generation model, excelling at creative composition and stylized expression. Supports text-to-image and image-to-image, with outstanding performance in portraits, scene rendering, and artistic creation.
FLUX.2 Klein 9B
Black Forest Labs' ultra-fast lightweight image generation model with 9B parameters, delivering blazing speed for high-frequency batch generation and rapid iteration workflows.
Nano Banana
Google-powered image generation model with excellent detail rendering and multi-image reference support. Handles up to 14 reference images for precise style control across diverse aspect ratios.
GLM Image
Zhipu AI's text-to-image model with strong Chinese and English bilingual understanding. Fast generation at ultra-low cost, ideal for rapid content creation and batch workflows.
AI Video Models
Veo 3.1
Google's latest flagship video generation model. Veo 3.1 Quality features industry-leading physics engine and ultra-high fidelity, perfectly replicating real-world textures, dynamics and details. Supports 16:9, 9:16 and Auto aspect ratios, ideal for commercial-grade high-quality video production.
Sora 2
The standard edition of the Sora series. Maintains OpenAI's superior prompt understanding while optimizing for speed and cost. Perfect for storyboarding, social media shorts, and rapid creative iteration.
HappyHorse
HappyHorse is Alibaba's next-generation multimodal video model with native audio-video co-generation. A single unified model handles four scenes — text-to-video, image-to-video, multi-image reference-to-video, and in-place video editing — making it ideal for ads, e-commerce, short drama, and social creatives.
Wan 2.6
Wan 2.6 is an advanced video generation model supporting text-to-video, image-to-video, and video-to-video modes. Offers duration options of 5s, 10s, and 15s with 720p and 1080p resolutions. Features multi-shot capabilities for creating diverse video content.
Kling Motion Control
Kling Motion Control model precisely controls character movements and poses by uploading reference images and videos. Supports 3-30 second videos, generates character actions consistent with references, ideal for character animation and motion transfer scenarios.
Kling 2.6
Renowned for capturing complex motion and physical laws. Kling 2.6 excels at generating high-dynamic character movements, intricate object interactions, and cinematic camera movements with fluidity.
Seedance 1.5 Pro
ByteDance's advanced video generation model. Seedance 1.5 Pro excels at character animation with precise lip-sync and natural expressions. Features realistic motion physics, supports multiple aspect ratios (1:1, 21:9, 4:3, 3:4, 16:9, 9:16), and offers flexible duration options (4s, 8s, 12s) with optional audio generation.
Seedance 2
ByteDance's next-generation video model focused on high visual quality, complex motion, and multi-modal reference control. Seedance 2 supports text, image, video, and audio inputs, making it ideal for professional video production that needs stronger consistency and richer camera language.
Seedance 2.5
ByteDance's latest video model with text-to-video, first/last-frame image-to-video, and multimodal reference-to-video. Supports reference images, videos, audio, 480p/720p, 4–30 seconds, and optional audio generation.
Seedance 2 Fast
The faster and more cost-efficient version of Seedance 2. It is ideal for rapid iteration, prompt testing, and high-volume content production while still supporting image, video, and audio references.
Seedance 2 Mini
The lightest Seedance 2 variant for budget-friendly generation. Supports first/last frames and multimodal references at lower per-second cost, ideal for drafts, social clips, and high-volume testing.
Grok Imagine
Creative video generation model from xAI. Grok Imagine excels at transforming text descriptions into imaginative video content, supports multiple aspect ratios (2:3, 3:2, 1:1, 9:16, 16:9), offers three style modes (fun, normal, spicy), perfect for creative content production and rapid prototyping.
Grok Imagine 1.5 Preview
Grok Imagine 1.5 Preview is a newer xAI-style video model for fast text-to-video and image-to-video creation. It supports 16:9 and 9:16 aspect ratios, 480p and 720p resolution, and short duration controls for social clips, ads, and creative tests.
Grok Video
Grok Video is xAI's advanced video generation model supporting 6s, 10s, 12s, 16s, and 20s durations. Supports text-to-video and image-to-video with up to 5 reference images. Offers multiple aspect ratios (16:9, 9:16, 2:3, 3:2, 1:1) and up to 5000 character prompts for detailed creative control.
Gemini Omni
Gemini Omni is Google's advanced video generation model powered by Omni-Flash-Ext. Supports text-to-video, single image-to-video, and 3-image reference fusion. Offers 4/6/8/10 second durations with 16:9 and 9:16 aspect ratios.
PixVerse V6
PixVerse V6 is a flexible video model for text-to-video, image animation, first/last-frame transitions, and video extension. It supports 360p to 1080p, 1-15 second clips, and optional native audio.