AI Models — 50+ Image, Video, Text, Audio & 3D Generators
Access the best AI models: FLUX, Stable Diffusion, Midjourney alternatives, Kling, Runway, GPT-4.1, Claude 4, ElevenLabs & more. Free to try!
Text & Chat AI ModelsGPT-4.1, Claude 4, Gemini, Llama 3 — Write articles, code, stories & more with cutting-edge language AI.

GPT-4.1
35 +
Advanced multimodal AI understanding text, images, video, and large files with superior coding capabilities—excellent for technical documentation, data analysi…
Multi-modalCost effectiveSupport file upload

GPT-4o
60 +
GPT-4o (GPT-4 Omni) by OpenAI is a lightning-fast, cost-efficient multimodal AI model that processes both text and images with exceptional contextual understan…
Support file upload

GPT-4o Mini
5 +
GPT-4o Mini is OpenAI's most cost-effective multimodal AI model, offering an optimal balance between performance and affordability. This compact yet powerful m…
Fast generationMulti-modalCost effective
GPT-5
20 +
GPT-5 represents OpenAI's next-generation flagship AI model with breakthrough capabilities in advanced reasoning, multimodal understanding, and sophisticated c…
High accuracyMulti-modalSupport file upload
GPT-5-mini
35 +
Efficient multimodal AI processing images and large documents at high speed—optimized for rapid content generation, summarization, and real-time applications
Fast generationCost effectiveLarge context
GPT-5.2
20 +
Latest iteration with breakthrough reasoning abilities and multimodal understanding—excels at complex problem-solving, advanced mathematics, and enterprise-lev…
High accuracyMulti-modalSupport file upload
GPT-5.4
75 +
OpenAI GPT-5.4 — latest generation language model with improved reasoning and creativity
High accuracyMulti-modalSupport file upload
GPT-5.4
160 +
OpenAI GPT-5.4 — latest generation language model with improved reasoning and creativity
High accuracyMulti-modalSupport file upload
GPT-5.6 Luna
Free +
GPT-5.6 Luna is OpenAI's budget tier in the GPT-5.6 family, offering a 1.05M token context window at a lower price point than Sol or Terra.
GPT-5.6 Sol
Free +
OpenAI's GPT-5.6 flagship reasoning model
GPT-5.6 Terra
Free +
OpenAI GPT-5.6 Terra — balanced mid-tier reasoning model between Sol (flagship) and Luna (cost-efficient)

GPT-OSS 120B
5 +
GPT-OSS 120B is OpenAI's open-weight 120-billion parameter language model designed for customization, on-premise deployment, and full enterprise control. This …
CreativeEfficientFast processing

Claude 4.1 Opus
1200 +
Enhanced AI assistant with extended memory for long-term projects—maintains context across sessions for ongoing coding work, research, and collaborative writing
Agentic toolsSafe alignmentHigh accuracy

Claude 4.5 Opus
60 +
Latest balanced AI model combining speed with intelligence—analyzes images, processes large documents, and excels at programming tasks with multilingual support
Multi-modalCost effectiveSupport file upload

Claude 4.5 Sonnet
60 +
Latest balanced AI model combining speed with intelligence—analyzes images, processes large documents, and excels at programming tasks with multilingual support
Multi-modalCost effectiveSupport file upload

Claude 4.6 Opus
60 +
Latest balanced AI model combining speed with intelligence—analyzes images, processes large documents, and excels at programming tasks with multilingual support
Multi-modalCost effectiveSupport file upload

Claude 4.6 Sonnet
60 +
Latest balanced AI model combining speed with intelligence—analyzes images, processes large documents, and excels at programming tasks with multilingual support
Multi-modalCost effectiveSupport file upload

Claude 4.7 Opus
60 +
Latest balanced AI model combining speed with intelligence—analyzes images, processes large documents, and excels at programming tasks with multilingual support
Multi-modalCost effectiveSupport file upload

Claude Fable 5
200 +
Anthropic's most capable widely released model — for the most demanding reasoning and long-horizon agentic work. A new tier above Opus, with a 1M token context…

Claude Haiku 4.5
20 +
Anthropic's fastest and most cost-effective current model — ideal for simple, speed-critical tasks. The platform's first Anthropic fast/cheap tier.

Claude Opus 4.8
100 +
Anthropic's most capable Opus-tier model — highly autonomous, state-of-the-art on long-horizon agentic work, knowledge work, and memory, with clearer, warmer w…

Claude Opus 5
100 +
Anthropic's newest Opus-tier flagship, released July 2026 — successor to Opus 4.8 with the same pricing tier, built for long-horizon agentic work, knowledge ta…

Claude Sonnet 5
160 +
Anthropic's Claude Sonnet 5 — near-Opus quality on coding and agentic work at Sonnet-tier speed and cost, with adaptive thinking.
Gemini 2.0 Flash
5 +
Next-gen multimodal AI creating text, images, audio and video from prompts—enables creative content production, interactive experiences, and multimedia generat…
High accuracyFast generationMulti-modal
Gemini 2.5 Flash
5 +
Efficient multimodal AI understanding images, audio, and video at high speed—cost-effective solution for content moderation, media analysis, and automated tran…
Low latencyMulti-modalCost effective
Gemini 2.5 Pro
40 +
Google's most capable AI processing massive multimodal datasets with advanced reasoning—ideal for academic research, complex data analysis, and scientific comp…
Fast responseHigh accuracyMulti-modal
Gemini 3.5 Flash
40 +
Google's most intelligent Flash-tier model for sustained frontier performance on agentic and coding tasks. Fully GA as of May 2026, succeeding Gemini 2.5 Flash.
Gemini 3.5 Flash-Lite
10 +
Google's fastest, cheapest current Flash tier (~350 tok/s) — GA July 2026, for high-volume low-latency text tasks.
Gemini 3.6 Flash
30 +
Google's current Flash-tier model, GA July 2026 — successor to Gemini 3.5 Flash with faster throughput and the same agentic/coding focus.
DeepSeek V4 Flash
Free +
DeepSeek's fast, low-cost flagship model with 1M token context and dual thinking/non-thinking modes
DeepSeek V4 Pro
Free +
DeepSeek's flagship reasoning model with 1M token context
Deepseek
60 +
Generates text, understands images and code; excels at reasoning
Fast reasoningHigh accuracyMulti-modal

o3
35 +
OpenAI's reasoning powerhouse with multi-step thought processes—solves advanced mathematics, writes complex algorithms, and performs deep scientific analysis
Advanced reasoningTool useHigh accuracy

o3 mini
20 +
Compact reasoning model optimized for STEM tasks—delivers accurate solutions in mathematics, physics, chemistry, and programming at lower cost
Fast responseHigh accuracyMulti-modal

o4 mini
20 +
Fast multimodal reasoning AI understanding images and text—combines visual analysis with mathematical logic for data science, engineering diagrams, and technic…
High accuracyFast generationMulti-modal

LLama 3.3 70B
5 +
Llama 3.3 70B Instruct Turbo by Meta is a powerful open-source instruction-tuned AI language model with 70 billion parameters, optimized for extended context u…
Instruction-tunedHigh accuracyMultilingual

Llama 3 8B
5 +
A powerful, conversational AI model optimized for natural language understanding and generation tasks.
High qualityOpen sourceMultilingual
Image Generation ModelsFLUX, Stable Diffusion XL, DALL-E 3, Ideogram — Create photorealistic images, art & illustrations free.
BG Remover
Free
BG Remover is a powerful AI background removal tool that automatically isolates subjects from their backgrounds with pixel-perfect precision. This image-to-ima…
Easy integrationFast processingHigh accuracy
Blend Images
150
Blend Images is an AI-powered image compositing tool that seamlessly merges multiple photos into realistic, cohesive compositions. This image-to-image model in…
High qualityFast generationSupports references
Clarity Upscaler
40 +
AI image upscaler that improves resolution, clarity, and style
Creative controlHigh qualityOpen source
Face Swap
5
Seamlessly swap faces in photos and videos with photoreal results
Supports videoHigh qualityHigh accuracy
Face to many
80 +
Identify individuals by matching one face against millions
Fast matchingScalableHigh accuracy
Image Upscaler
10
Upscale and restore images with AI for sharper, print-ready results
Fast processingScalableUser-friendly
Instant ID
150
Zero-shot identity-preserving image generation from one face
High accuracyFast generationSupports references
Instruct pix2pix
30 +
Text-guided image editor — fast, precise image-to-image edits
Fast inferenceHigh accuracyMulti-modal
QR Code Generator
5 +
Generate branded, secure QR codes with dynamic, trackable designs
CustomizableReal-time analyticsHigh accuracy
Sticker Maker
120 +
Generate stickers from text or photos — fast, editable, high‑res
High qualityFast generationMulti-modal
Style transfer
50
Create images in style of uploaded image
High qualityFast generationSupports references
Virtual Try On
160 +
Virtual Try On — realistic apparel & jewelry previews on you
Cross-platformReal-timeHigh accuracy
crystal-upscaler
200 +
Crystal Upscaler is a high-precision AI image upscaler optimized for portraits, faces, and product photography, powered by Clarity AI technology. This speciali…
Adjustable creativityBest for facesOfficial Replicate model
p-image
20 +
P-Image by Pruna AI is an ultra-fast, production-optimized text-to-image AI model generating high-quality images in under 1 second—completely free. This lightn…
Custom sizes & aspectOptional safety offSub-second generation
p-image-edit
40 +
P-Image Edit by Pruna AI is an ultra-fast, production-ready AI image editor delivering professional results in under 1 second at just $0.01 per generation. Thi…
Editing presetsMulti-referenceSub-second speed
FLUX 2 Flex
240 +
Max-quality image generation and editing with typography and up to 10 reference images.
High qualityPhotorealisticSupport file upload
FLUX 2 Klein 4B
5 +
Very fast image generation. 4-step distilled, sub-second inference for production and real-time applications.
High qualityPhotorealisticSupport file upload
FLUX 2 Max
160 +
The highest fidelity image model from Black Forest Labs. Superior prompt understanding, editing consistency, and multi-reference support.
High qualityPhotorealisticSupport file upload
FLUX 2 Pro
60 +
FLUX 2 Pro by Black Forest Labs is a professional-grade AI image generator designed for brand consistency and creative control. This advanced text-to-image and…
8 reference imagesAny aspect ratioCreative control
Flux Kontext Max
320 +
Flux Kontext Max by Black Forest Labs is an advanced multi-scene storytelling AI image generator that creates coherent visual narratives spanning multiple conn…
High consistencyFast generationMulti-modal
Flux Kontext Pro
160 +
Flux Kontext Pro by Black Forest Labs is an iterative AI image editor that generates and progressively refines visuals through multiple revision cycles, enabli…
High consistencyFast generationMulti-modal
Flux Pro 1.1
160
Flux Pro 1.1 by Black Forest Labs is a high-speed professional AI image generator producing 2K resolution outputs with exceptional prompt accuracy and rapid ge…
High qualityHigh accuracyFast generation
Flux Pro 1.1 Redux
200
Advanced image-to-image transformer specializing in style transfer and artistic remixing—reimagine photos with different aesthetics, lighting, and artistic sty…
High qualityFast generationSupports references
Flux Pro Canny
200
Precision retexturing tool maintaining original structure while applying new styles—transform images based on edge detection for architectural visualization an…
Supports references
Flux Pro Fill
200
Flux Pro Fill by Black Forest Labs is a professional AI inpainting and outpainting tool for seamless image expansion, object removal, and completion of partial…
High qualityFast generationMulti-modal
Flux Pro Ultra 1.1
240 +
Flux Pro Ultra 1.1 by Black Forest Labs is an ultra-high-resolution photorealistic AI image generator creating stunning 4-megapixel (4MP) outputs optimized for…
Versatile modesHigh qualityHigh accuracy
Flux Schnell
15 +
Flux Schnell (German for 'fast') by Black Forest Labs is an ultra-high-speed AI image generator optimized for instant visual creation. This lightning-fast text…
High qualityFast generationCost effective

SDXL Flash
20
Fast sdxl with higher quality
Diverse stylesHigh speedText to image
SDXL Pixar
15 +
Generate Pixar-style poster art from text or image inputs
CustomizableHigh qualityMulti-modal

SDXL Realism 2.0
50 +
Generate photorealistic images and portraits with cinematic lighting
Image inputHigh qualityHigh accuracy

Stable Diffusion 3
260 +
Generate high-resolution images from text and images, fast and customizable
High qualityFast generationMulti-modal

Stable Diffusion 3 Medium
140 +
Generate photorealistic images from text; runs on consumer hardware
High qualityFast generationMulti-modal

Stable Diffusion 3 Turbo
160 +
Fast text-to-image & image-to-image generation, excellent typography
TypographyHigh qualityFast generation

Stable Diffusion 3.5 Large
260 +
High-quality text-to-image and image-to-image at 1MP, strong prompt adherence
High qualityHigh accuracyMulti-modal

Stable Diffusion Core
120 +
Generate detailed images from text; inpainting, outpainting, edits
High qualityFast generation

Stable Diffusion XL
10 +
Stable Diffusion XL (SDXL) by Stability AI is a powerful open-source text-to-image and image-to-image AI model capable of generating ultra-high-resolution 1024…
High qualityFast generationMulti-modal
Ideogram 4.0 Quality
400
The highest quality Ideogram v4 model. v4 creates images with stunning realism, creative designs, and consistent styles
Ideogram Upscaler
240 +
AI upscaler doubling image resolution with enhanced detail recovery and intelligent cropping—sharpen low-res images for print and high-DPI displays
Easy integrationHigh qualityFast generation
Ideogram v1
320 +
Text-to-image AI excelling at legible typography in images—create logos, posters, infographics, and memes with accurate text rendering
High qualityHigh accuracyFast generation
Ideogram v1 Turbo
80 +
Fast image generator with clear text rendering at budget-friendly pricing—quickly produce social media graphics, banners, and promotional materials
Text-firstHigh qualityFast generation
Ideogram v2
320 +
Advanced text-to-image model with photorealistic quality and superior typography—generate professional visuals with embedded text for branding and design
Custom stylesHigh qualityHigh accuracy
Ideogram v2 Turbo
200 +
High-fidelity image generation with flexible style control and fast output—create diverse visual content from photorealistic to artistic with text integration
High qualityFast generationMulti-modal
Ideogram v3 Quality
360 +
Ideogram v3 Quality — highest fidelity variant of Ideogram v3 for premium graphic design and typography work.
Renders textBest for logosSupport file upload
Ideogram v3 Turbo
120 +
Ideogram v3 Turbo — fast, design-focused image generation with best-in-class typography, accurate text rendering and graphic design quality at a low price.
Renders textBest for logosSupport file upload
Recraft V3
160 +
Create print-ready designs with flawless text, precise layout, and vectors
Vector outputHigh qualityLarge context
Recraft V4
160 +
Recraft's latest image generation model, built around design taste. Strong prompt accuracy, art-directed composition, and integrated text rendering. Fast and c…
Renders textBest for logosAny aspect ratio
Recraft V4 Pro
1000 +
Recraft's latest image generation model at ~2048px resolution. Same design taste and prompt accuracy as V4, with higher resolution for print-ready and large-sc…
High qualityRenders textBest for logos
Recraft V4 Pro SVG
1200 +
Generate detailed SVG vector graphics from text prompts. Recraft V4 Pro's design taste with more geometric detail and finer paths — clean layers, editable outp…
High qualityRenders textBest for logos
Recraft V4 SVG
320 +
Generate production-ready SVG vector images from text prompts. Recraft V4's design taste applied to vector output — clean geometry, structured layers, and edit…
Renders textBest for logosAny aspect ratio
Nano Banana
160 +
Nano Banana by Google is an experimental AI image generator and editor that leverages natural language processing for intuitive visual creation and transformat…
Advanced image editingFast processingMulti-image
Nano Banana 2
270 +
Google's fast image generation model with conversational editing, multi-image fusion, and character consistency
High qualityAny aspect ratio
Nano Banana Pro
140 +
Nano Banana Pro by Google is the enhanced version of Google's experimental AI image generator and editor, offering improved quality and advanced natural langua…
Advanced image editingFast processingMulti-image
Seedream 4
120 +
Seedream 4 by ByteDance is a unified AI image generation and editing model capable of creating stunning images up to 4K resolution (4096×4096 pixels) from text…
4K outputAccurate editingHigh detail
Seedream 4.5
160 +
ByteDance Seedream 4.5 — upgraded unified image generation and editing model with stronger spatial understanding, world knowledge, and cinematic visuals. Up to…
High qualityPhotorealisticSupport file upload
Seedream 5 Lite
140 +
ByteDance Seedream 5 Lite — lightweight variant of Seedream 5 with built-in multi-step reasoning, example-based editing, and deep domain knowledge. Up to 3K.
High qualityPhotorealisticSupport file upload

GPT Image 1.5
40 +
OpenAI GPT Image 1.5 — latest flagship image gen/edit model with stronger preservation, precise iterative edits, up to 4x faster than GPT Image 1. Built on GPT…
High qualityPhotorealisticAny aspect ratio

GPT Image 2
35 +
OpenAI GPT Image 2 — state-of-the-art image generation and editing model. Highest performance with flexible image sizes, high-fidelity inputs, and precise iter…
High qualityPhotorealisticSupport file upload
Imagen 4 Fast
80 +
Google Imagen 4 Fast — fast, cost-effective variant of Imagen 4 with up to 2K resolution at ~$0.02/image.
High qualityPhotorealisticFast generation
Imagen 4 Ultra
240 +
Google Imagen 4 Ultra — top-tier photorealism, renders the finest details like skin texture and individual hair strands. Up to 2K resolution.
High qualityPhotorealisticAny aspect ratio
Audio & Voice AI ModelsElevenLabs, Suno — Generate realistic speech, clone voices, create music & sound effects.

ElevenLabs Music
320 +
ElevenLabs Music is a cutting-edge AI music generation platform that creates professional-quality, royalty-free music tracks from simple text descriptions. Thi…
High qualityLip syncControllable

ElevenLabs TTS
300 +
ElevenLabs Text-to-Speech (TTS) is the industry-leading AI voice synthesis platform delivering human-like speech in 29+ languages with exceptional naturalness …
Voice cloningHigh qualityLow latency

MusicGen
5 +
MusicGen by Meta AI is a sophisticated music generation model that creates high-quality, original music tracks from text descriptions or audio references. This…
High qualitySupports referencesControllable

MusicGen Remixer
500 +
AI music remixer with chord-aware controls for customizing generated tracks—remix, rearrange, and fine-tune AI music with harmonic precision
High qualityFast generationSupports references
ACE-Step Audio
10 +
ACE-Step Audio is an advanced AI music generator that transforms text prompts into professional-quality audio tracks. This text-to-music AI model excels at cre…
Fine controlHigh qualityFast generation
Google Lyria 3
160
Generate 30-second music clips from text prompts or images with Lyria 3, Google's music generation model

Stable Audio
40 +
Stability AI's music and sound generator from text or audio prompts—create professional audio tracks, ambiences, and soundscapes for creative projects
High qualityLarge context
lyria-3-pro
320
Generate full-length songs up to 3 minutes from text prompts or images with Lyria 3 Pro, Google's most capable music generation model
Video Generation ModelsKling, Runway Gen-3, Hailuo — Transform text & images into stunning AI videos.
Wan 2.1 I2V 480p
1800 +
WaveSpeed AI serving Wan 2.1 image-to-video at 480p. Cheaper draft-tier.
High qualityImage to videoSupport file upload
Wan 2.1 I2V 720p
5000 +
WaveSpeed AI serving Wan 2.1 image-to-video at 720p.
High qualityImage to videoSupport file upload
Wan 2.1 T2V 720p
4800 +
WaveSpeed AI serving Wan 2.1 text-to-video at 720p.
High qualityImage to video9:16 vertical
Wan 2.2 I2V Fast
200 +
Generate cinematic videos from images with fast, accurate control
High qualityFast generationMulti-modal
Wan 2.2 T2V Fast
200 +
Alibaba Wan 2.2 T2V Fast — fast text-to-video at 480p. Cheap and quick for iteration.
High qualityImage to videoFast generation
Wan 2.5 Image to Video
1000 +
Wan 2.5 Image to Video (I2V) by Alibaba is a powerful AI animator that transforms static images into cinematic videos with optional background audio synchroniz…
Audio syncCustom durationFlexible resolution
Wan 2.5 T2V
1000 +
Wan 2.5 T2V (Text-to-Video) by Alibaba is an advanced AI video generator that creates cinematic videos from text descriptions with optional audio synchronizati…
Audio syncFast processingFlexible resolution
Wan 2.5 T2V Fast
1360 +
Alibaba Wan 2.5 T2V Fast — fast variant of Wan 2.5 text-to-video for quick iteration.
High qualityImage to videoFast generation
Wan 2.7 T2V
2000 +
Alibaba Wan 2.7 — newest 27B open-weights video model. Generates up to 15s 1080p with native audio-sync from text prompts.
High qualityImage to video9:16 vertical
Wan 2.7 VideoEdit
800 +
Alibaba Wan 2.7 VideoEdit — instruction-based video editing model. Modify an existing clip with natural-language instructions: background swaps, lighting, styl…
High qualityImage to videoSupport file upload
Kling 2.5 Turbo Pro
1405 +
Kling 2.5 Turbo Pro — image-to-video with cinematic motion and precise intent following. Up to 10s.
Image to videoFast generationSupport file upload
Kling 2.6
1405 +
Kling 2.6 — text-to-video and image-to-video with native audio. 5-10s with motion prompts and negative prompts.
High qualityImage to videoSupport file upload
Kling 3 Pro I2V
3365 +
Kling Video v3 Pro image-to-video on fal.ai — cinematic visuals, fluid motion, native audio, element referencing. Up to 15s.
High qualityImage to videoSupport file upload
Kling V3 Motion Control
1405 +
Transfer motion from a reference video to any character image with improved consistency and quality.
High qualityImage to videoSupport file upload
Kling V3 Omni Video
2020 +
Unified multimodal video generation with reference images, video editing, native audio, and multi-shot control.
High qualityImage to videoSupport file upload
Kling V3 Video
2020 +
Cinematic videos up to 15 seconds with multi-shot control, native audio, and improved consistency.
High qualityImage to videoSupport file upload
SeeDANCE 1 Pro
240 +
Create cinematic 5-10 second 1080p videos from text or images—ByteDance's professional video generator for ads, social content, and previews
Supports referencesHigh qualityFast generation
SeeDANCE 1 Pro Fast
120 +
Generate cinematic 5–10s 1080p videos from text or images
Supports referencesHigh qualityFast generation
Seedance 1 Lite
145 +
Seedance 1 Lite by ByteDance is an affordable, high-speed AI video generator that creates professional cinematic videos from text prompts or static images with…
5–10 sec clipsAny aspect ratioFlexible resolution
Seedance 2.0
3600 +
ByteDance's multimodal video generation model with native audio, multimodal reference inputs, and intelligent duration control. Supports text-to-video and imag…
High qualityImage to videoCinematic
Seedance 2.0 Fast
1000 +
ByteDance Seedance 2.0 Fast — speed-optimized variant of Seedance 2.0 with native audio-video joint generation. Supports text-to-video and image-to-video at 48…
High qualityImage to videoCinematic
seedance-1.5-pro
105 +
SeeDANCE 1.5 Pro by ByteDance is a revolutionary joint audio-video AI model that generates cinematic videos with synchronized audio from text descriptions or i…
Duration, FPS & aspectOptional synced audioText & image to video
Omni-Human
500 +
Omni-Human by ByteDance is a revolutionary AI video generator that creates highly realistic lip-synced talking head videos from a single static photograph and …
High qualityHigh accuracyMulti-modal
Pyramid Flow
200 +
Generate short high-quality videos from text or images quickly—budget-friendly option for social media content and rapid prototyping
High qualityMulti-modalCost effective
ToonCrafter
80
Cartoon animation generator creating smooth transitions between keyframe images—produce animated shorts, explainer videos, and motion comics
User-friendlyHigh qualityHigh accuracy
Video Morpher
80
Blend multiple images with seamless morphing transitions—create mesmerizing visual effects for music videos, presentations, and artistic projects
Supports references
Video Upscaler
120 +
AI video upscaler enhancing low-resolution footage to 1080p/4K/8K with detail reconstruction—restore old videos and enhance quality for modern displays
Fast processingNoise reductionHigh quality
PixVerse V6
1000 +
PixVerse's flagship video generation model. Generate cinematic videos with synchronized audio, multi-shot sequences, and precise camera control.
PixVerse v4
1000 +
PixVerse v4 — video generation model supporting text-to-video and image-to-video up to 1080p.
High qualityImage to videoSupport file upload
PixVerse v4.5
1200 +
PixVerse v4.5 — upgraded version with improved motion quality and prompt adherence.
High qualityImage to videoSupport file upload
PixVerse v5
1405 +
PixVerse v5 — latest PixVerse with lifelike physics and striking visuals.
High qualityImage to videoSupport file upload
Runway Gen-3 Alpha Turbo
1000 +
Runway's flagship image-to-video model with cinematic camera controls—transform still images into dynamic professional footage for film and advertising
Cinematic controlHigh qualityFast generation
Runway Gen-4 Turbo
1000 +
Runway Gen-4 Turbo — fast, high-quality video generation from images with improved motion coherence and cinematic controls
High qualityImage to videoFast generation
Runway Gen-4.5
2400 +
Runway Gen-4.5 — highest quality video generation model with state-of-the-art motion and visual fidelity
High qualityImage to video9:16 vertical
gen-4.5
12000 +
State-of-the-art video motion quality, prompt adherence and visual fidelity
High fidelityHigh-quality motionImage to video
animate-diff
30 +
🎨 AnimateDiff (w/ MotionLoRAs for Panning, Zooming, etc): Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
High quality9:16 vertical
animate-diff-og
30 +
Animate Your Personalized Text-to-Image Diffusion Models
High quality9:16 vertical
animatediff-illusions
30 +
Monster Labs' Controlnet QR Code Monster v2 For SD-1.5 on top of AnimateDiff Prompt Travel (Motion Module SD 1.5 v2)
High quality9:16 vertical
animatediff-prompt-travel
30 +
🎨AnimateDiff Prompt Travel🧭 Seamlessly Navigate and Animate Between Text-to-Image Prompts for Dynamic Visual Narratives
High quality9:16 vertical
Hailuo 2.3 Pro I2V
1960 +
MiniMax Hailuo 2.3 Pro image-to-video (1080p) on fal.ai — flagship tier for quality.
Image to videoFast generationSupport file upload
MiniMax Hailuo 2.3
1120 +
MiniMax Hailuo 2.3 — latest flagship video model with improved motion, physics, and prompt adherence. Supports text-to-video and image-to-video up to 1080p.
High qualityFast generationSupport file upload
MiniMax Hailuo 2.3 Fast
760 +
MiniMax Hailuo 2.3 Fast — faster and cheaper variant of Hailuo 2.3 for quick iterations.
High qualityFast generationSupport file upload
Google Veo 3.1
3200 +
Google Veo 3.1 — state-of-the-art text-to-video model with native audio generation, natural lip sync, cinematic quality, and strong prompt adherence. Up to 108…
Image to videoCinematicSupport file upload
Google Veo 3.1 Fast
2000 +
Google Veo 3.1 Fast — faster, cheaper variant of Veo 3.1 with native audio. Great for high-volume production.
Image to videoCinematicSupport file upload

Luma Ray 2
6400 +
Luma Ray 2 represents the next generation of AI video synthesis, delivering photorealistic, professional-grade video generation from text descriptions or image…
Cinematic controlHigh qualityMulti-modal

Luma Ray 2 Flash
2200 +
Fast photorealistic video creation from text or images with rapid turnaround—perfect for quick content production and social media campaigns
Cinematic controlHigh qualityMulti-modal
Mochi v1
1600 +
Text-to-video AI creating high-fidelity realistic motion and scenes—excellent for storytelling, advertising, and creative video projects
CustomizableHigh qualityHigh accuracy
mochi-1
140
Mochi 1 preview is an open video generation model with high-fidelity motion and strong prompt adherence in preliminary evaluation
High quality9:16 vertical
OpenAI Sora 2
1600 +
OpenAI Sora 2 — flagship video model with cinematic quality, strong physics simulation, and precise prompt adherence. 4-12s at 720p/1080p.
Image to videoCinematicSupport file upload
OpenAI Sora 2 Pro
4800 +
OpenAI Sora 2 Pro — highest-fidelity Sora 2 tier for premium cinematic video. Text-to-video and image-to-video.
Image to videoCinematicSupport file upload
Google Veo 3.1 Lite
800 +
Google's cost-efficient video generation model with native audio, optimized for high-volume applications
MMAudio Foley
200 +
Adds AI-generated sound effects, foley and ambience to a silent video, synchronized to the visual content. Describe the desired sound in the prompt.
controlvideo
140 +
Training-free Controllable Text-to-Video Generation
High quality9:16 vertical
damo-text-to-video
30
Multi-stage text-to-video generation
High quality9:16 vertical
hotshot-xl
30 +
😊 Hotshot-XL is an AI text-to-GIF model trained to work alongside Stable Diffusion XL
High quality9:16 vertical
i2vgen-xl
30
RESEARCH/NON-COMMERCIAL USE ONLY: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models
High qualityImage to video9:16 vertical
ltx-2.3-pro
1920 +
High-fidelity video generation with portrait support, audio-to-video, retake, and extend. Text, image, and audio-driven creation up to 4K at 50 FPS.
pia
140 +
Personalized Image Animator
High qualityImage to video9:16 vertical

stable-diffusion-animation
30 +
Animate Stable Diffusion by interpolating between two prompts
High quality9:16 vertical

stable_diffusion_infinite_zoom
30 +
Use Runway's Stable-diffusion inpainting model to create an infinite loop video
High quality9:16 vertical
text2video-zero
560
Text-to-Image Diffusion Models are Zero-Shot Video Generators
High quality9:16 vertical
tile-morph
30 +
Create tileable animations with seamless transitions
High quality9:16 vertical








