FLUX 2 (Pro) Edit by Black Forest Labs. Instruction-based image editor with up to 8 reference images, 4 MP output, and flexible aspect ratios.
Models
All Models
LongCat Image by Meituan edits photos via plain-language instructions, no masks needed. Supports 15 edit types and accurate text rendering in Chinese and English.
Sparc3D (Portrait) by Hitem3D converts 1-4 face photos into detailed 3D head models, up to 1536³ Pro resolution, as mesh-only or fully textured assets.
Minimax Hailuo 2.3 (Fast) by MiniMax. Text-to-video and image-to-video with optimized speed. Choose 6 or 10-second clips at 768p, or 6 seconds at 1080p.
Minimax Hailuo 2.3 by MiniMax generates text and image-to-video at 768p or 1080p, with strong motion coherence, temporal consistency, and anime style support.
A Flux Kontext LoRA that ages clean building images into worn, crumbling ruins, preserving the original structure, layout, and perspective.
Turns a character image into a 9-emotion expression sheet, preserving their look and art style.
Reve Remix by Reve AI combines up to 6 reference images into a single result using a text prompt, with 7 aspect ratios to choose from.
Clarity Crystal Upscaler by Clarity AI enhances images with an adjustable scale factor and a creativity dial, optimized for portraits, faces, and product photos.
Flux Kontext LoRA that transforms 3D blockout shapes into photorealistic renders, preserving the original geometry and spatial layout.
Veo 3.1 (Fast) by Google. Fast text-to-video and image-to-video with native audio, first/last frame anchoring, and subject-consistent generation from reference images.
Generate a four-view character turnaround sheet from any character image, showing every side on a clean background.
Flux Kontext LoRA that converts photos of buildings into isometric 3D models placed on square tiles, ideal for game assets and top-down scene design.
Hunyuan Image 3 by Tencent. An 80B MoE text-to-image model with exceptional multilingual text rendering, complex scene understanding, and photorealistic output.
SeedVR2 - Image Upscale by Bytedance scales images up to 4K using one-step diffusion super-resolution, with factor or target resolution modes and a noise slider.
SeedVR2 - Video Upscale by ByteDance upscales videos up to 4K using one-step diffusion, restoring fine detail and temporal consistency with minimal hallucination.
Omni Human 1.5 by ByteDance animates a single photo into a talking avatar, syncing lip movements, expressions, and natural gestures to your audio.
Wan 2.5 I2V by Alibaba animates any image into a 720p or 1080p video clip of 5 or 10 seconds, with optional audio synchronization.
Wan 2.5 T2V by Alibaba generates 720p or 1080p videos from text, in 16:9 or 9:16, with synchronized audio support, in 5 or 10 second clips.
Wan 2.2 Animate (Replace) by Alibaba swaps a person in a video with your reference character image, preserving the original motion and audio.
Wan 2.2 Animate (Move) by Alibaba brings any character image to life by transferring body motion and expressions from a reference video.
Wan 2.2 Reframe by Alibaba converts any video to a new aspect ratio, keeping the subject intelligently framed. Outputs in 16:9, 1:1, or 9:16 at up to 720p.
Lucy Edit (Pro) by Decart edits videos via text: swap outfits, change objects, or replace scenes while preserving motion and identity. Up to 720p.
Lucy Edit (Dev) by Decart AI edits videos from text prompts, swapping outfits, characters, or scenes while preserving original motion and composition.
Expand any video's borders left, right, up, or down with Wan 2.2 Outpainting by Alibaba. New content is generated to match your scene seamlessly.
Sync Lipsync 2 (Pro) by Sync Labs syncs mouth movements to any audio track with studio-grade detail preservation, up to 4K resolution.
Luma Modify Video by Luma Labs rewrites existing footage with a text prompt across nine Adhere, Flex, and Reimagine strength levels.
ElevenLabs 3 (Alpha) is an expressive text-to-speech model for 70+ languages, using inline tags like [whispers] or [excited] to steer emotion, emphasis, and delivery.
Pixverse Lipsync by PixVerse syncs any audio track to a speaker's mouth movements in your video, with built-in TTS across 14 voices.
Meta MusicGen by Meta generates music from text prompts or a reference melody. Choose stereo or melody-guided variants and produce clips up to 30 seconds.
Creatify Lipsync by Creatify syncs realistic mouth movements to any audio track, letting you dub or revoice video content for marketing and social.
ElevenLabs Multilingual 2 by ElevenLabs converts text to natural, emotionally rich speech in 29 languages. Choose from 21 built-in voices or bring your own cloned voice.
ElevenLabs Turbo 2.5 by ElevenLabs. Low-latency text-to-speech in 32 languages. Choose from 21 built-in voices or bring a cloned voice, with speed and style controls.
ElevenLabs Sound Effects 2 by ElevenLabs generates custom SFX from text descriptions, with clips up to 30 seconds, seamless looping, and adjustable guidance.
Sync Lipsync 2 by Sync Labs matches mouth movements in any video to a new audio track. Works on live-action, animation, and AI-generated faces.
Kling Lipsync by Kuaishou syncs lip movement in any video to a new audio track or typed text, animating facial motion while keeping the background intact.
Pixverse 5 by PixVerse. Text-to-video and image-to-video at up to 1080p, with first/last frame transitions, 15 creative effects, and 5 or 8-second outputs.
Flux.1 LoRA for stylized 3D fantasy characters. Generates armored warriors with sculpted detail, expressive faces, and dragon companions.
A Flux.1 LoRA for 3D cartoon RPG enemies. Produces stylized monsters, orcs, goblins, and skeletons with glowing eyes, bold colors, and simple display bases.
A Flux.1 LoRA for painterly fantasy environments: ancient ruins, stone arenas, and dramatic outdoor scenes with rich lighting and vibrant colors.
Upscale images 2x to 8x with FLUX. Choose between precise, balanced, or creative presets to preserve structure or add imaginative detail.
Topaz Image Upscale by Topaz Labs enlarges photos up to 6x with five specialized enhancement modes, subject detection, and face enhancement controls.
A Flux.1 LoRA that renders characters and scenes in a contemporary animated feature style: vibrant colors, cinematic lighting, and elaborate medieval-inspired costumes.
A Flux.1 LoRA that renders characters in dynamic poses with a collectible action figure aesthetic. Bold colors, exaggerated features, clean backgrounds.
A Flux.1 LoRA for game UI assets. Generates buttons, HUDs, shop screens, inventory panels, and character stat menus in sci-fi or fantasy styles.