Every image, video, motion, and audio model on TheFluxTrain. Open any model for a short description, how to use it, capabilities, and FAQs — sign in only when you generate.
Quick answer: TheFluxTrain model pages list every available AI image, video, motion, and audio model so you can compare capabilities and generate from a dedicated URL.
Speed-Optimized Detail
OpenAI image model on fal — sharp detail and typography
4x faster, lower cost, better quality
Google's Flagship Generation Model
ByteDance Seedream — cinematic text-to-image at 1K/2K
Professional-grade image upscaling
Premium text-to-image detail and composition
Uncensored image generation
Animated product demo from ordered app screenshots with AI voiceover
Avatar speaking/singing from image + audio
Quick generation with good quality
Continue a clip with FLUX.3 native audio, 5–20s, 720p/1080p
Fast 720p FLUX.3 continuation preview
Interpolate start and end frames with FLUX.3 native audio, 5–20s
Fast 720p start/end interpolation preview
Black Forest Labs FLUX.3 — animate a still with native audio, 5–20s, 720p/1080p
Fast 720p FLUX.3 preview from a still; enhance later at 1080p
Pin start and end frames on a FLUX.3 timeline (24 fps)
Fast 720p keyframe interpolation preview
Black Forest Labs FLUX.3 text to video with native audio, 5–20s, 720p/1080p
Fast 720p FLUX.3 preview from text; enhance later at 1080p
Google Gemini Omni Flash — image to video with synced audio, 720p, 3–10s
Google Gemini Omni Flash — refs as <IMAGE_REF_0>, synced audio, 720p, 3–10s
Google Gemini Omni Flash text to video with synced audio, 720p, 3–10s
Google Gemini Omni Flash — conversational video edits, 720p
Alibaba Happy Horse — image to video with native audio
Alibaba Happy Horse — refs as character1, character2…
Alibaba Happy Horse text to video with native audio
Alibaba Happy Horse natural-language video edits
Talking-head video with AI infographic overlays timed to voiceover
Professional grade video editing
Standard quality video editing
Professional grade video synthesis
High-fidelity cinematic motion
Professional grade video synthesis from reference
High-fidelity cinematic motion from reference
Professional grade text to video
High-fidelity text to video
Professional grade video reference
Standard quality video reference
Video from audio + optional image
Extend video forward or backward
LTX-2.3 image-to-video with native audio
Text-to-video with native audio
Restyle video from prompt
Generate from reference video
Standard quality generation
Multi-stage talking avatar with segmented motion and 720p/1080p output
Storyline + screenshots → scored promo with review gates, GPT Image 2 assets/storyboards, and Seedance 2.0 video
Bytedance Seedance with audio
Bytedance Seedance with audio
ByteDance Seedance 2.0 — fast tier, native audio
Up to 9 images / video+audio refs — fast tier
ByteDance Seedance 2.0 — fast text-to-video + audio
ByteDance Seedance 2.0 — up to 1080p, native audio
Up to 9 images / video+audio refs — quality tier
ByteDance Seedance 2.0 — quality text-to-video + audio
Auto-route to script or film based on input maturity
Story + hero character image → 35–40s animated short (4 scenes, Seedance reference-to-video)
Develop a story document from an idea with evaluate/improve loop
Professional-grade video upscaling
Alibaba Wan 2.7 Pro — image to video with optional audio, 720p/1080p, 2–15s
Alibaba Wan 2.7 Pro — reference images/videos, optional audio, 720p/1080p, 2–10s
Alibaba Wan 2.7 Pro text to video with optional audio, 720p/1080p, 2–15s
Alibaba Wan 2.7 Pro — natural-language video edits, 720p/1080p
Browsing the catalog and reading model pages does not spend credits. Credits are used only when you start a generation after signing in.