FLUX 1.1 (Pro) by Black Forest Labs is a fast text-to-image model that tops prompt-adherence and aesthetics benchmarks, generating photorealistic images in seconds.
Models
All Models
FLUX 1.1 (Pro Ultra) by Black Forest Labs generates images at up to 4MP across 11 aspect ratios, with a Raw mode for natural, candid photography looks.
Kling AI Avatar 2 (Pro) by Kuaishou turns a single image and an audio track into a lifelike talking avatar video with precise lip sync.
Kling 2.6 I2V (Pro) by Kuaishou animates a still into 1080p video with its strongest motion yet, stable character identity, and optional native audio synced in one pass.
Kling 2.6 T2V (Pro) by Kuaishou Technology generates 1080p video clips of 5 or 10 seconds with optional native audio, including voice, sound effects, and ambience.
Kling O1 I2V by Kuaishou animates a start image into video with optional end-frame control for precise transitions. Choose 5 or 10 second clips.
Voxel Crafter 1.0 generates blocky 3D voxel models from text or images, with grid size controls (16 to 256) for width, height, and depth.
Grid Maker arranges multiple images into a clean grid layout. Control columns, rows, padding, background color, and cell aspect ratio.
Texture Converter turns a flat image into a surface material, with sliders for how raised, shiny, polished, and angular it looks, plus an option to invert the relief.
Extract individual frames from any video as PNG, JPEG, or WebP images. Choose a frame interval or pull every frame at once.
Assemble up to 1,000 images into an MP4 or GIF with control over frame rate, compression, looping, and optional audio.
Scenario Gemini Upscale by Google uses multimodal reasoning to sharpen and enhance images up to 4K, with an optional prompt and creativity dial.
A Flux.1 LoRA that renders any landscape or environment in a bold-line cartoon illustration style, with clean outlines and smooth cel shading.
Flux.1 LoRA for glossy, candy-colored surreal landscapes. Generates dripping liquids, swirling textures, and vibrant floral scenes.
Flux.1 LoRA for stylized 3D cartoon characters with vibrant colors and exaggerated proportions, built for games and animation assets.
Flux LoRA that renders complete costume sets as flat-lay collections, from medieval armor to sci-fi gear, in muted, semi-realistic tones.
Z-Image Turbo by Tongyi-MAI is a distilled open-source image model built for near-instant generation with strong photorealism and accurate text rendering in images.
Bria Remove Background by Bria isolates subjects with clean cutouts, preserving fine details like hair. Outputs a transparent PNG or a fully opaque image.
Image Slicer by Scenario splits any image into a custom grid of up to 6x6 sections, outputting each tile as its own file.
Pixel Snapper cleans up pixel art by snapping every pixel to a consistent grid and reducing colors to a strict palette, removing blurry AI artifacts.
Retro Diffusion Tile by Retro-Diffusion generates seamless pixel art tilesets and game assets in six modes, from full tilesets to placeable objects.
Retro Diffusion Plus by Retro-Diffusion generates crisp, grid-aligned pixel art in 18 styles, from isometric assets to Minecraft textures.
Retro Diffusion Animation by Retro-Diffusion generates pixel art character walk cycles, idle states, small sprites, and VFX effects as animated GIFs or spritesheets.
FLUX 2 (Dev) by Black Forest Labs is an open-weight image model built for fine-tuning. Stack up to 6 LoRAs, add up to 5 reference images, and generate up to 2048x2048.
Meshy Rigging by Meshy adds a skeleton and skin weights to humanoid GLB characters automatically, making them animation-ready in one step.
Hunyuan 3D Part by Tencent segments an existing GLB mesh into clean, organized sub-parts. Especially effective on mechanical and hard-surface models, ready for editing.
Gemini 3.0 Pro by Google. Instruction-based image generation and editing with advanced reasoning, multi-image fusion, and outputs up to 4K resolution.
Flash VSR by Tsinghua/Shanghai AI Lab upscales video 2x to 4x using one-step diffusion, delivering sharp detail restoration with real-time streaming speed.
Pixverse Swap by PixVerse replaces people or backgrounds in any video using a reference image, with Person and Background modes up to 720p.
Hunyuan 3D 3.0 Pro (Sketch) by Tencent converts hand-drawn sketches into textured 3D meshes with adjustable face count and optional PBR materials.
Hunyuan 3D Pro 3.0 by Tencent converts a single image into a detailed 3D mesh with up to 1.5M faces, PBR textures, and a choice of triangle or quad topology.
Hunyuan 3D 3.0 Pro (Multiview) by Tencent generates 3D assets from up to 4 reference views, with PBR materials, face count control, and tri or quad topology.
Qwen Edit Multi-Angle by Alibaba edits images using camera controls: rotate, zoom, vertical tilt, and wide-angle, with optional text prompts.
Gemini 2.5 Flash Image by Google edits and generates images from text instructions, with up to 6 reference images and aspect ratios from ultrawide to portrait.
Flux Kontext by Black Forest Labs. Edit images with text instructions, preserving characters and style across local or global edits. Supports up to 10 reference images.
GPT Image 1.5 by OpenAI. Instruction-driven image editing with precise reference fidelity, transparent background support, and aspect ratio control.
Seedream 4.5 by ByteDance edits images from natural language instructions, with up to 10 reference images, 4K output, and strong subject preservation.
P-Image Edit by Pruna AI is a fast, instruction-driven image editor. Edit, composite, or transform images using text and up to 10 reference images.
Meshy Remesh by Meshy rebuilds your 3D model's geometry with clean triangle or quad topology and a target polycount you control, from 100 to 300,000 polygons.
Meshy Image-to-3D by Meshy converts one photo or up to four multiview images into a textured, PBR-ready 3D mesh with full geometry coverage.
Meshy Text-to-3D by Meshy generates textured, game-ready 3D assets from a text prompt using Meshy 5 or 6.
Meshy Retexture by Meshy applies new textures to existing 3D models from a text prompt or reference image, with optional full PBR map generation.
Flux.2 [klein] 4B by Black Forest Labs. A 4B-parameter model distilled for sub-second image generation and editing, with LoRA support.
FLUX.2 [klein] 4B Base by Black Forest Labs. Compact, undistilled 4B image model built for LoRA compatibility, fine-tuning, and precise prompt control.
Flux 2 (Turbo) Edit by Black Forest Labs: fast, instruction-based image editing. Describe the change, get a transformed image in seconds.
FLUX.2 [klein] 9B Base by Black Forest Labs. Undistilled 9B foundation model for text-to-image and image-to-image, built for LoRA fine-tuning and high output diversity.
FLUX 2 (Max) by Black Forest Labs. Top-tier image editing with up to 8 reference images, up to 4 MP output, and precise consistency across colors, faces, and objects.