Added May 10, 2026
GPT-Image-2 is OpenAI's image generation model, first announced on April 21, 2026, and available through the API in early May 2026. It improves prompt adherence, text rendering, and photorealism compared to earlier DALL-E generations.
Why: GPT-Image-2 is OpenAI's most capable image model to date, with notably better text-in-image accuracy. It is a natural choice for OpenAI API users who want image generation alongside text and audio in a single platform.
Added May 19, 2026
ComfyUI is an open-source node-based interface for diffusion models. Instead of a prompt box, a generation is a graph: loaders, samplers, conditioning, upscalers and masks wired together, each step inspectable and re-runnable. Workflows save as JSON and can be shared, which is why most published Stable Diffusion and FLUX pipelines circulate as ComfyUI graphs.
Why: If you want to know exactly what happened between the prompt and the image, this is the tool that shows you. Every step is a node you can open, change and re-run, and a workflow someone else built arrives as a file you can load rather than a screenshot you have to reverse engineer. That reproducibility is why it became the format the rest of the field builds around.
Added Aug 8, 2026
OpenArt is a hosted generative image platform with a node-based workflow builder alongside a conventional prompt interface. Workflows can be built from templates or from scratch, run on hosted compute, and published for other people to reuse.
Why: It is the shortest path from wanting a node workflow to having one running. ComfyUI asks you to bring a GPU and set it up; OpenArt hosts the compute and ships a template library, so the graph is something you edit rather than something you first have to stand up.
Added Aug 8, 2026
Weavy is a browser-based node canvas for chaining hosted generative models into a single pipeline, mixing image, video and editing steps from different providers in one graph rather than moving files between tools.
Why: Most canvases are built around one model family. This one treats the model as a node, so a pipeline can pass through several providers without leaving the graph. That matters when the best step for a job is not all from the same vendor.
Added Aug 8, 2026
Invoke is an open-source generative image platform combining a unified canvas with inpainting, outpainting and layer control, plus a node editor for building repeatable workflows. It can be self-hosted or used through a commercial hosted tier aimed at studios.
Why: Its centre of gravity is the canvas rather than the graph, which suits the way a lot of real work happens: generate something, then keep editing regions of it. The node editor is there when a job needs repeating, instead of being the only way in.
Added Feb 5, 2026
Generates high-aesthetic images from text prompts with strong artistic style and composition. Produces variations and allows style exploration through Discord-based workflow with iterative refinement. Supports multiple aspect ratios, style parameters (--style, --stylize), and advanced features like remix mode for composition control. Known for exceptional artistic taste and cinematic quality output suitable for professional concept art and creative projects.
Why: Consistently strong artistic style and taste, making it the go-to choice for concept art and aesthetic image generation.
Added Jan 31, 2026
Flora is a collaborative AI design canvas that moves beyond the prompt box and into node-based workflow orchestration. It allows designers to build complex, repeatable AI creative engines by connecting different 'Nodes', such as Sketch-to-Image, ControlNet, and multi-model refinement layers. Unlike traditional AI tools, Flora is built for teams, offering real-time collaborative spaces where multiple creators can design and iterate on the same AI canvas simultaneously. It represents the shift from simple prompting to professional AI design systems.
Why: Flora is built for more than one person working on the same graph at the same time, which most node canvases are not. If the bottleneck in your work is handing a workflow to a colleague rather than the workflow itself, that is what it solves. ComfyUI gives you more control and Invoke gives you a better editing canvas, so pick this one for the collaboration.
Added May 15, 2026
FLUX.2 [max] is Black Forest Labs' flagship image generation model. It builds on the FLUX architecture with improved prompt adherence, anatomy, text rendering, and aesthetic quality for professional image creation.
Why: FLUX.2 [max] continues the FLUX lineage of excellent prompt adherence and typography. It is a top choice for designers, advertisers, and developers who need reliable, high-quality image generation.
Added Feb 5, 2026
Generates images from text prompts with exceptional typography and text rendering capabilities. Produces high-quality text-in-image designs, logos, and poster-style visuals with accurate text placement and readability. Supports multiple aspect ratios, style controls, and advanced typography features. Generates professional-grade output suitable for marketing materials, brand assets, and design projects with precise text rendering that other models struggle with.
Why: Great for posters, logos, and brand mockups where accurate text rendering is critical.
Added Feb 5, 2026
Generates and edits images with a creator-friendly UI and extensive model library. Provides image variations, inpainting, outpainting, and production workflows with multiple AI models and style options. Supports multiple aspect ratios, resolution up to 1024x1024, and advanced editing tools. Generates professional-quality output suitable for concept art, game assets, and design projects with comprehensive workflow features.
Why: Good all-around image tool with comprehensive workflow features for concept art and production pipelines.
Added Feb 5, 2026
Animates images into stylized video clips with motion presets and artistic effects. Creates music-video style animations with fast aesthetic transformations and creative motion patterns. Supports multiple animation styles, motion intensity controls, and artistic filters. Produces unique stylized videos suitable for music videos, creative projects, and social media content with distinctive visual aesthetics.
Why: Great for music-video style animations and fast aesthetics with unique stylized motion effects.
Added Feb 5, 2026
Generates and edits images with native integration into Adobe Creative Cloud workflows. Provides generative fill, text-to-image, and style transfer directly within Photoshop, Illustrator, and other Adobe applications. Supports commercial-safe content generation, multiple style options, and seamless workflow integration. Produces professional-grade output suitable for commercial design work with full Creative Cloud compatibility.
Why: Great when you already live in Adobe apps and need seamless integration with existing design workflows.
Added May 20, 2026
Recraft V4 is a design-focused image generation model from Recraft, released in 2026. It specializes in brand-consistent visuals, vector graphics, illustrations, and marketing assets with precise style control.
Why: Recraft V4 is built for designers rather than casual prompt users. Its emphasis on brand consistency, vector output, and editable design assets makes it unique among image generation tools.
Added Feb 5, 2026
Helps generate and refine images with creator-oriented workflows and real-time preview. Provides image generation, variations, and refinement tools with fast iteration cycles for creative exploration. Features real-time AI preview that shows results as you type, allowing instant visual feedback. Supports multiple generation modes, style transfer, and creative enhancement tools optimized for rapid prototyping and artistic experimentation.
Why: Good for fast creative iteration and image refinement with real-time preview and creator-focused features.
Added Feb 5, 2026
Black Forest Labs' FLUX.1 [pro] is a state-of-the-art image generation model that outperforms almost everything in prompt adherence, human anatomy, and complex text rendering within images.
Why: FLUX.1 [pro] is the 'Master Artist' for AI images. Most AI tools are bad at writing words inside pictures, but Flux is perfect at it. It's the best tool for designers who need high-quality posters, logos, and photos that look 100% real.
Added Oct 2, 2024
FLUX.1 Fill [pro] is a specialized Black Forest Labs model for high-quality image inpainting, outpainting, and content-aware editing. It uses a masked conditioning approach to seamlessly integrate new content into existing images.
Added Oct 2, 2024
FLUX.1 Canny is a control-oriented Black Forest Labs model that uses Canny edge maps to guide image generation and structure-preserving edits. It is useful for maintaining pose, composition, and object outlines while changing styles or content.
Added Oct 2, 2024
FLUX.1 Depth is a control model from Black Forest Labs that uses depth maps to guide new image generation or editing. It preserves the spatial structure of a scene while allowing changes to objects, lighting, and style.
Added Jun 26, 2025
BRIA FIBO is an 8B-parameter DiT text-to-image model trained on long structured JSON captions. It turns short prompts into detailed structured schemas and generates images with precise, reproducible control over composition, lighting, camera, and color. It is also available in an image-to-image 'Inspire' mode and is trained entirely on licensed data for commercial safety.
Why: FIBO stands out for native JSON structured prompting and fully licensed training data, making it the strongest open-source choice for enterprises that need predictable, legally safe image generation.
Added Nov 11, 2025
BRIA FIBO Lite is a lightweight variant of the FIBO image generation pipeline. It pairs an open-source FIBO-VLM bridge with a smaller FIBO Lite model to enable rapid inference and fully local, on-prem deployment for privacy-critical environments.
Why: FIBO Lite gives teams a FIBO-family option optimized for speed and data sovereignty, with a fully local deployment path that the full FIBO pipeline does not emphasize.
Added Jul 24, 2025
BRIA RMBG 2.0 is a dichotomous image segmentation model that produces a grayscale alpha matte for high-quality background removal. It is trained on over 15,000 fully licensed, manually labeled high-resolution images and is designed for e-commerce, advertising, gaming, and enterprise content workflows.
Why: RMBG 2.0 is a widely adopted, source-available background removal model with strong commercial licensing and a dedicated GitHub presence, filling a clear gap alongside BRIA's eraser tools.
Added Sep 28, 2025
Native multimodal MoE image generation model with 80 billion total parameters and 13 billion active parameters. It unifies multimodal understanding and generation within an autoregressive framework, supports image editing and multi-image fusion, and is released as the largest open-source image generation model.
Why: The largest open-source image generation model, combining high parameter counts with efficient MoE inference.
Added Mar 1, 2025
Kling Image 2.0 is Kling AI's image generation model, optimized for high-quality text-to-image and image-to-image generation with strong style and composition control.
Added Jun 15, 2026
Luma Uni-1.1 is a multimodal reasoning model that understands intention, follows reference images, and generates or edits images with style and brand consistency. It supports text-to-image, image-to-image, and multi-reference generation, and ranks highly in human preference benchmarks for overall quality, style and editing, and reference-based generation.
Why: Uni-1.1 ties a reasoning model directly to pixel generation, making it unusually good at following brand references and complex creative direction in images.
Added May 14, 2026
Recraft V4.1 is Recraft's latest image generation model, released in May 2026. It improves photorealism, short-prompt understanding, and illustration quality, and ships with Standard, Pro, Vector, and Utility variants for different creative and production needs.
Why: Recraft V4.1 is the current flagship model, offering more natural photorealism, refined illustration quality, and dedicated Utility and Vector variants for production design workflows.
Added Jun 1, 2024
Upscales and enlarges images using AI while recovering natural detail and reducing artifacts. Offers specialized models for photos, art, text, low-resolution sources, and face recovery.
Why: The standalone upscaling specialist in Topaz's lineup, with dedicated models for different image types and up to 8x enlargement.
Added Apr 28, 2026
Reinterprets and enhances AI-generated images and digital art up to 8x, adding texture, lighting, and realism while preserving the original composition. Uses adjustable creativity and realism controls.
Why: A dedicated creative upscaler for AI-generated imagery, bridging the gap between raw AI output and production-ready assets.
Added Jun 1, 2025
Runs Topaz image enhancement tools directly in the browser with unlimited cloud rendering. Includes denoise, sharpen, upscale, face enhancement, background removal, colorization, and creative upscaling.
Why: The no-install, browser-based entry point to Topaz image enhancement with a wide workflow menu and cloud rendering.
Added Apr 28, 2026
Brings Topaz Photo AI enhancement capabilities to iPhone, allowing mobile photographers to upscale, sharpen, denoise, and enhance images directly on their device.
Why: Extends Topaz's photo enhancement models to iPhone, giving mobile creators access to desktop-quality AI polish.
New this month
Added Sep 10, 2026
ChatGPT Images 2.5 is OpenAI's latest image generation model announced September 8, 2026. Delivers sharper detail, more natural lighting and texture, better preservation of people and products in reference photos, and 50% faster generation than Images 2.0. Reliable editing across multi-turn conversations. Available in ChatGPT (all tiers), Codex, and API with two variants: Flare (high-quality, low-latency default) and Sunburst (premium control for production workflows).
Why: Represents significant iterative improvement in image quality, speed, and editing reliability. Two API variants show sophisticated tuning for different use cases. 50% speed improvement is substantial for production workflows.
Added Feb 5, 2026
Generates design assets including logos, vectors, and brand visuals with clean, usable outputs. Produces vector-style graphics, illustrations, and design elements optimized for production workflows. Specializes in creating scalable vector graphics, logo designs, and brand assets that maintain quality at any size. Supports multiple design styles, aspect ratios, and export formats suitable for professional design work and brand identity projects.
Why: Great for design assets when you want clean, usable outputs with vector-style graphics and brand-ready visuals.
Added Feb 5, 2026
Generates and edits images with context awareness for better coherence using Flux Kontext model. Understands image context and relationships to produce more coherent variations, edits, and style transfers with improved consistency. Advanced context understanding enables the model to maintain visual relationships, preserve important elements, and create coherent edits that respect the original image's context. Ideal for image editing, variations, and style transfer tasks requiring consistency.
Why: Context-aware generation for more coherent results, making it superior for image editing and variation tasks requiring consistency.
Added Feb 5, 2026
Enhances and upscales images with AI-powered detail boost and quality improvement. Provides advanced upscaling, detail enhancement, and final polish tools for creators refining their outputs to production quality. Supports upscaling up to 8x resolution with intelligent detail generation, creative enhancement modes, and fine-tuned control over enhancement intensity. Produces professional-grade results suitable for print, digital media, and high-resolution displays.
Why: High-quality enhancement for creators polishing outputs with exceptional detail preservation and quality improvement.
Added Feb 5, 2026
Generates image variations and edits using Wan 2.6 architecture with improved quality and style control. Produces coherent variations, style transfers, and image edits with enhanced visual quality and better prompt adherence. Latest iteration of Wan's image-to-image technology with superior quality, better style control, and improved prompt understanding. Suitable for creating variations, applying styles, and editing images with high visual fidelity.
Why: Latest Wan iteration for I2I with improved quality, representing the current state-of-the-art in Wan's image-to-image capabilities.
Added Feb 5, 2026
Publishes the FLUX family of state-of-the-art image generation models including FLUX.1, FLUX.1-dev, FLUX.2, and specialized variants. Provides open-source models with exceptional quality and prompt adherence for modern image generation workflows. FLUX models represent cutting-edge diffusion technology with superior text rendering, style control, and image quality. Offers multiple model variants optimized for different use cases including speed, quality, and specialized applications.
Why: Important modern image model family to know and track, representing the cutting edge of open-source image generation.
Added Feb 5, 2026
Removes unwanted objects from images with high fidelity and minimal artifacts using BRIA's advanced inpainting technology. Produces clean results with seamless background reconstruction and natural-looking edits. Advanced AI inpainting understands image context to generate plausible replacements for removed objects, maintaining visual consistency and natural appearance. Ideal for professional image cleanup, background editing, and object removal workflows requiring high-quality results.
Why: Best-in-class object removal with clean results, making it the top choice for professional image cleanup and editing workflows.
New this month
Added Sep 6, 2026
ChatGPT Images 2.0 is OpenAI's updated image generation model integrated into ChatGPT. Improves upon prior versions with better detail rendering, faster generation speed (up to 4x faster), and notably improved text rendering in images. Can generate images from text prompts and edit existing photos with more precision. Ranked second globally in text-to-image generation and image editing (behind previous versions of its own gpt-image-2 model). Available to ChatGPT users with Plus/Pro subscriptions.
Why: Shows iterative improvement in text-to-image quality and speed. Better text rendering is significant for design and creative use. Demonstrates practical advancement in generative image quality.
Added Feb 5, 2026
Generates and edits images via an open model ecosystem including Stable Diffusion models and community tools. Provides local generation, API access, and extensive customization options with fine control over generation parameters. Supports multiple model versions, LoRA fine-tuning, ControlNet for precise control, and a vast ecosystem of community models and tools. Enables complete workflow customization from local deployment to cloud API integration, making it the foundation for many custom image generation pipelines.
Why: Core ecosystem for customizable image workflows with open-source flexibility and extensive community support.
Added Feb 5, 2026
Helps create designs and generate assets inside a familiar, user-friendly editor with built-in AI features. Provides text-to-image, background removal, and design automation tools integrated into a comprehensive design platform. Offers extensive template library, drag-and-drop interface, and AI-powered design suggestions. Supports social media graphics, presentations, marketing materials, and print designs with seamless AI integration for non-designers and professionals alike.
Why: Best mainstream design workflow for non-designers with intuitive interface and integrated AI generation features.
Added Feb 5, 2026
Enhances photos with strong AI-powered denoise, sharpen, and upscale tools using advanced image processing algorithms. Provides professional photo cleanup, detail enhancement, and quality improvement for final image polish. Combines multiple AI models for face recovery, denoising, sharpening, and upscaling in a unified workflow. Supports batch processing, automatic model selection, and fine-tuned control over enhancement parameters for professional photography workflows.
Why: Great finishing tool for polishing images with exceptional denoising and sharpening capabilities for professional workflows.
Added Feb 5, 2026
A 7B parameter multimodal model developed by ByteDance-Seed, capable of generating both text and images. Supports text-to-image generation, image-to-image editing, and image understanding in a unified framework. Provides versatile capabilities for content creation and image manipulation workflows. Multimodal architecture enables seamless integration of text and image generation with editing capabilities, making it ideal for complex content creation workflows requiring multiple modalities in a single model.
Why: Unique multimodal capabilities combining text and image generation with editing, making it versatile for complex content creation workflows requiring multiple modalities.