AI Models
Choose from the latest AI video and image generation models. All available on Flashloop for iOS, Android, and web.
Video Models
Veo 3
Generate 8-second videos with native dialogue, sound effects, and ambient audio in one pass.
Kling 3.0
Create up to 15-second multi-shot videos with character consistency, 4K 60fps support, and bilingual audio.
Kling 3.0 Turbo
A faster, more affordable Kling 3.0, text-to-video and image-to-video at 720p or 1080p, up to 15 seconds.
Kling 2.6
Turn images into 10-second 1080p videos with 48fps motion control and native speaking, singing, or rapping audio.
Sora 2
Generate up to 30-second videos with strong physics, long scene continuity, and precise style control.
Seedance 2.5
ByteDance's 30-second single-shot video model, with up to 50 multimodal references and native synchronized audio.
Seedance 2.0
ByteDance's video model with a Fast/High switch, native audio, up to 7 reference images, and director-level camera control.
Seedance 2.0 Mini
ByteDance's faster, lower-cost Seedance, native audio, multi-reference, and 480p/720p output.

MiniMax H3 Max
MiniMax's flagship with always-on native audio, up to 12 mixed reference files, and character tagging that keeps the same person in every shot.

MiniMax H3 Max Turbo
H3 Max's speed-tuned sibling — the same post-trained lineage at roughly twice the speed and half the price, with native audio on every clip.
MiniMax Hailuo 3
Native stereo audio written with the picture, legible on-screen text, and native 2K.
Seedance 1.5 Pro
Generate 12-second videos with native audio, multilingual speech, and fast turnaround from 45 seconds to 3 minutes.
Wan 2.6
Generate 15-second videos with audio, strong character consistency, and a free tier on Flashloop.
Wan 3.0
One multimodal model for text, image, video, audio, web and document references, with native audio and up to 30 seconds at 1080p.
Wan 3.0 Prime
The same all-in-one multimodal video model as WAN 3.0, tuned for faster end-to-end generation.
Grok Imagine
Generate 30-second videos with native audio, 4 instant variations, and fast renders in about 17 seconds.

Gemini Omni
Multimodal video generation, text, up to 7 reference images, or a video clip. Native 4K output and up to 10-second clips.
Image Models

Nano Banana Pro
Generate 4K images, edit with plain English, and blend up to 4 references in one pass.

Flux 2
Create fast, flexible images with up to 10 references, LoRA tuning, and four speed-to-quality tiers.

Seedream 5.0 Lite
Generate 4K images in 2–3 seconds with web-aware prompts, strong reasoning, and near-perfect EN/CN text.

GPT Image 2
Generate high-quality images from text or references with OpenAI's GPT Image 2, 1K, 2K, and 4K output with 20,000-character prompts.