FireRed Image Edit

AI Models

Compare AI image models and AI video models available on FireRed Image Edit, then open the best model page or generator for your workflow.

30 Models

AI Image Models

FireRed Image Edit
New

FireRed Image Edit

An AI image editing model built for targeted inpainting, object replacement, background modification, and style refinement. It is ideal for high-control edits that preserve the original composition while improving visual quality, especially for e-commerce retouching, poster updates, social media assets, and creative rework.

View model
Nano Banana 2
New

Nano Banana 2

Fast and affordable image generation model powered by Google Gemini Flash Image. Ideal for everyday creative tasks with flexible resolutions from 0.5K to 4K.

View model
Nano Banana 2 Lite
New

Nano Banana 2 Lite

Google DeepMind’s fastest and most efficient Gemini Image model for low-latency image generation and editing. Use it when you need rapid drafts, reference-guided edits, and reliable 1K visual exploration at a lighter credit cost.

View model
Nano Banana Pro
New

Nano Banana Pro

High-quality image generation with enhanced details and better consistency

View model
Seedream 5.0 Lite
New

Seedream 5.0 Lite

ByteDance's next-generation lightweight image generation model that delivers flagship-level visual quality at significantly faster speeds. Supports batch generation of up to 15 images with strong prompt understanding, ideal for high-frequency creative workflows.

View model
GPT-Image 1.5
New

GPT-Image 1.5

Professional-grade photorealistic image generation model, specializing in capturing delicate light/shadow textures and natural human expressions. Supports both text-to-image and image-to-image modes, with excellent performance in fidelity and detail restoration, particularly suitable for commercial scenarios requiring high realism.

View model
GPT Image 2
New

GPT Image 2

A multi-reference image generation model that supports up to 16 reference images, built for high consistency, stronger control, and more complex commercial workflows. Supports both text-to-image and image-to-image generation.

View model
Seedream 4.5
New

Seedream 4.5

Flagship visual generation model with ultimate natural language understanding and light/shadow expression. It no longer relies on mechanical tag stacking, but understands 'stories' and 'atmosphere', switching freely between hyper-realistic photography and surreal art.

View model
FLUX.2
New

FLUX.2

Next-generation industrial-grade model, breaking the 'text curse' of AI drawing. It has the strongest instruction following capability to date, not only perfectly generating specified words/text within images, but also precisely controlling complex hand movements and object spatial relationships.

View model
Z-Image
Image

Z-Image

A model optimized for text-to-image scenarios, dedicated to accurately mapping complex descriptions to visual elements, especially excelling in conceptualization and abstract description realization.

View model
Grok 4.2 Image
New

Grok 4.2 Image

xAI's latest Grok 4.2 image generation model, excelling at creative composition and stylized expression. Supports text-to-image and image-to-image, with outstanding performance in portraits, scene rendering, and artistic creation.

View model
FLUX.2 Klein 9B
Image

FLUX.2 Klein 9B

Black Forest Labs' ultra-fast lightweight image generation model with 9B parameters, delivering blazing speed for high-frequency batch generation and rapid iteration workflows.

View model
Nano Banana
Image

Nano Banana

Google-powered image generation model with excellent detail rendering and multi-image reference support. Handles up to 14 reference images for precise style control across diverse aspect ratios.

View model
GLM Image
Image

GLM Image

Zhipu AI's text-to-image model with strong Chinese and English bilingual understanding. Fast generation at ultra-low cost, ideal for rapid content creation and batch workflows.

View model

AI Video Models

Veo 3.1
New

Veo 3.1

Google's latest flagship video generation model. Veo 3.1 Quality features industry-leading physics engine and ultra-high fidelity, perfectly replicating real-world textures, dynamics and details. Supports 16:9, 9:16 and Auto aspect ratios, ideal for commercial-grade high-quality video production.

View model
Sora 2
Video

Sora 2

The standard edition of the Sora series. Maintains OpenAI's superior prompt understanding while optimizing for speed and cost. Perfect for storyboarding, social media shorts, and rapid creative iteration.

View model
HappyHorse
New

HappyHorse

HappyHorse is Alibaba's next-generation multimodal video model with native audio-video co-generation. A single unified model handles four scenes — text-to-video, image-to-video, multi-image reference-to-video, and in-place video editing — making it ideal for ads, e-commerce, short drama, and social creatives.

View model
Wan 2.6
New

Wan 2.6

Wan 2.6 is an advanced video generation model supporting text-to-video, image-to-video, and video-to-video modes. Offers duration options of 5s, 10s, and 15s with 720p and 1080p resolutions. Features multi-shot capabilities for creating diverse video content.

View model
Kling Motion Control
Video

Kling Motion Control

Kling Motion Control model precisely controls character movements and poses by uploading reference images and videos. Supports 3-30 second videos, generates character actions consistent with references, ideal for character animation and motion transfer scenarios.

View model
Kling 2.6
Video

Kling 2.6

Renowned for capturing complex motion and physical laws. Kling 2.6 excels at generating high-dynamic character movements, intricate object interactions, and cinematic camera movements with fluidity.

View model
Seedance 1.5 Pro
Video

Seedance 1.5 Pro

ByteDance's advanced video generation model. Seedance 1.5 Pro excels at character animation with precise lip-sync and natural expressions. Features realistic motion physics, supports multiple aspect ratios (1:1, 21:9, 4:3, 3:4, 16:9, 9:16), and offers flexible duration options (4s, 8s, 12s) with optional audio generation.

View model
Seedance 2
New

Seedance 2

ByteDance's next-generation video model focused on high visual quality, complex motion, and multi-modal reference control. Seedance 2 supports text, image, video, and audio inputs, making it ideal for professional video production that needs stronger consistency and richer camera language.

View model
Seedance 2.5
New

Seedance 2.5

ByteDance's latest video model with text-to-video, first/last-frame image-to-video, and multimodal reference-to-video. Supports reference images, videos, audio, 480p/720p, 4–30 seconds, and optional audio generation.

View model
Seedance 2 Fast
New

Seedance 2 Fast

The faster and more cost-efficient version of Seedance 2. It is ideal for rapid iteration, prompt testing, and high-volume content production while still supporting image, video, and audio references.

View model
Seedance 2 Mini
New

Seedance 2 Mini

The lightest Seedance 2 variant for budget-friendly generation. Supports first/last frames and multimodal references at lower per-second cost, ideal for drafts, social clips, and high-volume testing.

View model
Grok Imagine
Video

Grok Imagine

Creative video generation model from xAI. Grok Imagine excels at transforming text descriptions into imaginative video content, supports multiple aspect ratios (2:3, 3:2, 1:1, 9:16, 16:9), offers three style modes (fun, normal, spicy), perfect for creative content production and rapid prototyping.

View model
Grok Imagine 1.5 Preview
New

Grok Imagine 1.5 Preview

Grok Imagine 1.5 Preview is a newer xAI-style video model for fast text-to-video and image-to-video creation. It supports 16:9 and 9:16 aspect ratios, 480p and 720p resolution, and short duration controls for social clips, ads, and creative tests.

View model
Grok Video
New

Grok Video

Grok Video is xAI's advanced video generation model supporting 6s, 10s, 12s, 16s, and 20s durations. Supports text-to-video and image-to-video with up to 5 reference images. Offers multiple aspect ratios (16:9, 9:16, 2:3, 3:2, 1:1) and up to 5000 character prompts for detailed creative control.

View model
Gemini Omni
New

Gemini Omni

Gemini Omni is Google's advanced video generation model powered by Omni-Flash-Ext. Supports text-to-video, single image-to-video, and 3-image reference fusion. Offers 4/6/8/10 second durations with 16:9 and 9:16 aspect ratios.

View model
New

PixVerse V6

PixVerse V6 is a flexible video model for text-to-video, image animation, first/last-frame transitions, and video extension. It supports 360p to 1080p, 1-15 second clips, and optional native audio.

View model