Kling 2.5 T2V (Pro) by Kuaishou. Text-to-1080p video at Turbo speed. Choose 5 or 10 seconds in 16:9, 9:16, or 1:1. Strong prompt adherence and dynamic motion.
Models
All Models
Kling 2.5 I2V (Pro) by Kuaishou turns a still image into a smooth 1080p video clip, with optional last-frame anchoring for precise start-to-end control.
Sync Lipsync 2 (Pro) by Sync Labs syncs mouth movements to any audio track with studio-grade detail preservation, up to 4K resolution.
Luma Modify Video by Luma Labs rewrites existing footage with a text prompt across nine Adhere, Flex, and Reimagine strength levels.
Sync Lipsync 2 by Sync Labs matches mouth movements in any video to a new audio track. Works on live-action, animation, and AI-generated faces.
Google Lyria 2 generates high-fidelity instrumental music from text prompts, producing up to 30-second clips at 48kHz stereo across a wide range of genres.
Pixverse 5 by PixVerse. Text-to-video and image-to-video at up to 1080p, with first/last frame transitions, 15 creative effects, and 5 or 8-second outputs.
Luma Video Reframe by Luma Labs resizes any video across 7 aspect ratios by outpainting beyond the original frame, preserving your subject.
Dreamina 3.1 by ByteDance generates images up to 2K with precise style control, rich detail, and strong text rendering. Nine aspect ratios or custom dimensions.
Qwen Image by Alibaba is a text-to-image and image-to-image model with strong text rendering and broad artistic style support. Up to 2048x2048.
Wan 2.2 T2V by Alibaba is an open-source text-to-video model with a Mixture-of-Experts architecture. Generate 480p or 720p clips with motion speed and seed controls.
Wan 2.2 I2V by Alibaba. Open-source image-to-video with first and last frame conditioning, 480p or 720p output, and a motion speed control.
Runway Gen4 Turbo by Runway ML is a fast image-to-video model. Animate a first frame into clips of 2 to 10 seconds across six aspect ratios, from 21:9 to 9:16.
PartCrafter by PKU turns a single image into up to 16 separate, semantically distinct 3D meshes in one pass, no segmentation required.
Minimax Video 02 by MiniMax generates realistic video from text or images, with natural motion, physics accuracy, and first and last frame anchoring.
Minimax Image 01 by MiniMax generates photorealistic images with advanced lighting, natural skin rendering, and optional character reference support.
Luma Photon by Luma AI generates images with strong prompt adherence across seven aspect ratios, with reference controls for content, style, and character.
Luma Photon Flash by Luma Labs is a fast text-to-image model built for rapid concept iteration, with reference image, style, and character guidance controls.
Ideogram 3 (Quality) by Ideogram generates images with precise text rendering and complex layouts. Supports inpainting, up to 4 style references, and 60+ style presets.
Ideogram 3 (Balanced) by Ideogram generates images and handles inpainting with strong text rendering, 60+ style presets, and up to 4 style reference images.
Ideogram 3 (Turbo) by Ideogram. Fast text-to-image with inpainting, 15 aspect ratios, 60+ style presets, up to 4 style references, and multilingual Magic Prompt.
Rodin Gen-1 (HighPack) by Deemos Technology generates detailed 3D models from text or images, with 4K textures and up to 500k polygon geometry.
Rodin Gen-1 by Deemos Technology generates textured 3D models from text or images, with PBR materials, multi-view input, pose control, and quality settings.
Hunyuan 3D 2.1 by Tencent converts a single image into a textured 3D mesh with production-ready PBR materials, precise geometry, and adjustable face count.
Seedance 1 (Pro) by ByteDance generates cinematic video up to 1080p, from text or an image, with first and last frame control and flexible aspect ratios.
Wan 2.1 (1.3B) by Alibaba. Lightweight open-source text-to-video model generating 480p clips up to 5 seconds in 16:9 or 9:16, built for consumer GPUs.
Minimax 01 Director by MiniMax is a text- and image-to-video model with 15 bracketed camera commands for precise shot control. Generates 720p clips up to 6 seconds.
Minimax Video 01 by MiniMax: a foundational text-to-video and image-to-video model with stable motion. Good for early prototyping and experimentation.
Luma Ray 2 Flash (720p) by Luma Labs generates 720p videos from text or a first frame image, with 5 or 9 second clips and 34 camera movement controls.
Luma Ray 2 Flash (540p) by Luma Labs. Fast text-to-video and image-to-video at 540p, 5 or 9 seconds, seven aspect ratios, and 34 camera moves.
Pixverse 4.5 by PixVerse generates videos from text or images with first/last-frame control, built-in audio, style presets, and motion effects.