PRICING • CURATED
Freemium AI Tools
These tools offer a free tier with additional paid features. Great for getting started, with options to upgrade as your needs grow.
RESULTS
Standalone agent-first platform with CLI, SDK, and managed agents
AI-powered IDE built as a fork of Visual Studio Code, designed with an 'agent-first' paradigm where autonomous AI agents plan, execute, and validate code
Why: Antigravity 2.0 is Google's most credible bid for the agentic IDE seat. The new CLI and SDK make it competitive with Cursor, Claude Code, and Codex for terminal-first and automation workflows.
Freemium
Best for Google-Native Agents
Visit
The Native Agentic Layer: The Browser as an OS
Google Chrome has evolved from a simple browser into a native agentic layer powered by Gemini 3
Why: We added Chrome to the agentic category because it represents the first time a mainstream browser has integrated a native reasoning engine that can autonomously navigate the web on behalf of the user.
Freemium
Best for Native Web Automation
Visit
The Invisible OS: Pure Execution via Messaging
Moltbot (also known as Clawdbot) is the spearhead of the 'Invisible OS' movement, a shift away from fragmented apps and toward pure, autonomous execution via messaging
Why: Moltbot represents the death of the 'app for everything' era. We picked it because it's the first agentic assistant to prove that reasoning-based execution through simple chat is more powerful than manual task management in 10+ different apps.
Freemium
Best for Agentic Automation
Visit
OpenAI's latest image generation model
GPT-Image-2 is OpenAI's image generation model, first announced on April 21, 2026, and available through the API in early May 2026
Why: GPT-Image-2 is OpenAI's most capable image model to date, with notably better text-in-image accuracy. It is a natural choice for OpenAI API users who want image generation alongside text and audio in a single platform.
Freemium
Best for OpenAI Image API
Visit
30-second 4K video with native audio and up to 50 reference inputs
Seedance 2
Why: The longest single-run generation of any current video model at 4K, and the 50-reference input system is the most direct answer yet to character consistency, the problem that breaks most AI video work. Availability is the constraint: it ships inside ByteDance's own apps first, and the previous generation's international rollout was postponed indefinitely.
Freemium
Best for Long Clips
Visit
Platform for prototyping with Google's Gemini models
Web-based integrated development environment for prototyping and building applications with Google's generative AI models
Why: Official Google platform providing direct access to Gemini models with excellent developer tools and seamless API integration.
Freemium
Best for Gemini Models
Visit
OpenAI's AI browser with agent mode for autonomous tasks
An AI-powered web browser developed by OpenAI, built on Chromium and integrating ChatGPT directly into the browsing experience
Why: OpenAI's flagship agentic browser with powerful Agent Mode for autonomous task execution and seamless ChatGPT integration.
Freemium
Best for Automation
Visit
Anthropic's frontier model, currently first on the Artificial Analysis Intelligence Index
Claude Opus 5 is Anthropic's flagship model, released 24 July 2026 with a 1M-token context window and five selectable effort levels (low, medium, high, xhigh, max)
Why: It is the current number one on the independent Artificial Analysis Intelligence Index, and it got there while costing less per task than the model it displaced: $2.03 average per index task against Fable 5's $2.75. The effort dial is the reason to pick it over a fixed-tier model, because one integration covers cheap high-volume calls and expensive long-horizon agent runs.
Freemium
Best Frontier Model Overall
Visit
The frontier of cinematic video synthesis
Kling AI is a state-of-the-art video generation platform capable of producing high-fidelity cinematic content
Why: Kling AI is the current king of AI movies. It can create high-quality video clips that are 2 minutes long, which is like an eternity in AI time, while keeping the characters and the physics (like how water splashes or hair moves) looking perfectly real. It's the first tool that lets professional filmmakers create a whole scene without the video 'glitching' halfway through.
Freemium
Best for Filmmaking
Visit
Node workflows without running your own GPU
OpenArt is a hosted generative image platform with a node-based workflow builder alongside a conventional prompt interface
Why: It is the shortest path from wanting a node workflow to having one running. ComfyUI asks you to bring a GPU and set it up; OpenArt hosts the compute and ships a template library, so the graph is something you edit rather than something you first have to stand up.
Freemium
Best for hosted workflows
Visit
The AI-native IDE that redefined software engineering
Cursor is a fork of VS Code built specifically for AI-pair programming
Why: Cursor is a special coding tool that actually 'reads' your entire folder of files. Imagine having a partner who remembers every single line of code you've ever written and can tell you exactly where a bug is hiding. It's the top choice for developers because it makes building apps 10 times faster by doing the boring 'search and find' work for you.
Freemium
Best for AI Coding
Visit
Moonshot AI's 2.8-trillion-parameter open-weight flagship
Kimi K3 is Moonshot AI's open-weight model released July 16, 2026, built at roughly 2
Why: Kimi K3 is one of the most credible open-weight challengers to closed frontier models this year, aggressive enough on pricing and scale that it moved markets, Fortune covered it as a 'DeepSeek shock' moment for AI stocks.
Freemium
Best for Open-Weight Frontier Performance
Visit
One canvas, many models, wired together
Weavy is a browser-based node canvas for chaining hosted generative models into a single pipeline, mixing image, video and editing steps from different providers in one graph rather than moving files ...
Why: Most canvases are built around one model family. This one treats the model as a node, so a pipeline can pass through several providers without leaving the graph. That matters when the best step for a job is not all from the same vendor.
Freemium
Best for mixing models
Visit
Open-source node canvas built around the edit, not the prompt
Invoke is an open-source generative image platform combining a unified canvas with inpainting, outpainting and layer control, plus a node editor for building repeatable workflows
Why: Its centre of gravity is the canvas rather than the graph, which suits the way a lot of real work happens: generate something, then keep editing regions of it. The node editor is there when a job needs repeating, instead of being the only way in.
Freemium
Best for iterative editing
Visit
AI-powered full-stack development platform
AI-powered platform that enables users to build full-stack applications using natural language descriptions
Why: Best platform for non-technical users to build full-stack applications through natural language.
Freemium
Best for Rapid Prototyping
Visit
Design platform with multiple AI tools and licensed content
Graphic design platform offering multiple AI-powered tools including F Lite image generator (trained on licensed data), image editing, video generation, icon generation, AI image classification, and a...
Why: Unique combination of AI tools and licensed content, ensuring commercial compliance for design projects.
Freemium
Best for Licensed Content
Visit
The Efficiency Revolution: Frontier Intelligence at 1/100th the Cost
DeepSeek is the architect of the 'DeepSeek movement,' a fundamental shift in AI development that prioritizes extreme efficiency over raw compute
Why: DeepSeek changed the game by proving that 'expensive' doesn't always mean 'better.' We picked it because it's the first model family to offer true frontier-level reasoning (R1), general intelligence (V3), and advanced vision/OCR (VL2) with an open-weight philosophy and an API price point that makes proprietary models look obsolete.
Freemium
Best for Cost-Efficiency
Visit
Text-to-music & vocals with fast iteration
Generates complete songs from text prompts, including both instrumental music and vocal tracks
Why: Suno is the current gold standard for mainstream text-to-music generation, offering unparalleled speed for creating full song drafts with high-fidelity vocals. Its ability to maintain musical structure across various genres while allowing for rapid iteration makes it the premier choice for creators needing instant, high-quality audio content.
Freemium
Best for Music
Visit
Anthropic's cheaper, near-Opus everyday model
Claude Sonnet 5 is Anthropic's mid-tier model released June 30, 2026, replacing Sonnet 4
Why: Sonnet 5 is the practical default for most day-to-day work: it closes much of the gap to Opus-tier performance while staying meaningfully cheaper, and Anthropic made it the automatic replacement for Sonnet 4.6 across the free and Pro tiers.
Freemium
Best for Everyday Agentic Work
Visit
Cloud-based online IDE for web development
Cloud-based online IDE focused on web application development
Why: Best cloud IDE for web development with instant setup and collaboration.
Freemium
Best for Web Development
Visit
European open-source and commercial LLM
Mistral AI provides high-performance large language models with both open-source and commercial offerings
Why: European LLM provider with strong open-source offerings, multilingual capabilities, and focus on data privacy and compliance.
Freemium
Best for Europe
Visit
High-quality TTS and voice tools
Generates realistic text-to-speech voiceovers with natural intonation and emotion
Why: Best voice quality combined with reliable API for production pipelines requiring consistent, natural-sounding narration.
Freemium
Best for Narration
Visit
Online IDE by Google with AI assistance
Online IDE developed by Google, based on Visual Studio Code and running on Google Cloud infrastructure
Why: Best cloud IDE for Google Cloud development with integrated AI and Android emulation.
Freemium
Best for Google Cloud
Visit
Alibaba's multilingual open-source LLM
Qwen is Alibaba Cloud's family of large language models with multiple versions: Qwen-1
Why: Alibaba's high-performance multilingual LLM with strong Chinese language support, cost-efficient pricing, and comprehensive open-source availability.
Freemium
Best for Multilingual
Visit
The 'Next DeepSeek' Movement: o1-Level Reasoning at 1/100th the Cost
Kimi k1
Why: Kimi k1.5 is the first model to prove that o1-level reasoning is achievable through efficient, open-weight architectures. We selected it because it consistently matches or exceeds Claude 4.5 in technical benchmarks (AIME, MATH-500) while offering a 2M context window and a significantly lower API price point, making frontier intelligence accessible to everyone.
Freemium
Best for Technical Reasoning
Visit
The Open Vision Frontier: 124B Multimodal Power
Pixtral Large is Mistral AI's flagship 124B parameter multimodal model, designed to compete directly with GPT-4o and Claude 3
Why: We added Pixtral Large because it represents the peak of European open-weight AI. It is one of the few open models that truly matches the visual reasoning depth of the top proprietary models, making it essential for the Open Frontier movement.
Freemium
Best for Complex Visual Reasoning
Visit
OpenAI's fast default ChatGPT model from May 2026
GPT-5
Why: GPT-5.5 Instant is the model most ChatGPT users will interact with by default. Its balance of speed and capability makes it a practical baseline for writing, analysis, coding help, and general assistant tasks.
Freemium
Best for Everyday ChatGPT
Visit
Google's fast, capable multimodal model from I/O 2026
Gemini 3
Why: Gemini 3.5 Flash hits a practical sweet spot for developers and creators who need more capability than entry-level models but do not require the full cost of an Ultra model. Its native multimodal design makes it especially useful for mixed-media tasks.
Freemium
Best for Fast Multimodality
Visit
xAI's agentic coding CLI for autonomous software engineering
Grok Build is an agentic command-line coding assistant from xAI that understands natural-language project descriptions, generates and edits code across files, runs commands, and iterates until tasks a...
Why: Grok Build brings xAI's frontier reasoning directly into the terminal, making it a strong alternative to other agentic coding CLIs. It is particularly useful for developers already embedded in the X and xAI ecosystem who want a fast, opinionated agent.
Freemium
Best for Agentic CLI
Visit
Google's unified multimodal generation model
Gemini Omni is a single Google model announced at I/O 2026 that can generate and reason across text, images, video, and audio from unified prompts
Why: Gemini Omni represents Google's push toward a single model for all media types. For teams building multimodal products, it simplifies architecture by replacing multiple specialized endpoints with one interface.
Freemium
Best for Unified Generation
Visit
Text/image-to-video with Pikaffects (squish, melt, explode)
Generates short-form videos from text or images with punchy motion and creative effects
Why: Great for quick social clips with unique Pikaffects that create viral-style transformations and motion effects.
Freemium
Best for Effects
Visit
Fast video generation from Luma Dream Machine
Creates realistic visuals with natural, coherent motion using Luma's Ray2 Flash model optimized for speed
Why: Speed + quality balance for quick iterations with fast generation times and reliable motion quality.
Freemium
Best for Speed
Visit
Google's personal AI agent for proactive assistance
Gemini Spark is a personal AI agent announced at Google I/O on May 19, 2026
Why: Gemini Spark is Google's answer to the emerging personal-agent category. By integrating deeply with Gmail, Calendar, Maps, and Android, it can automate everyday tasks that previously required switching between apps.
Freemium
Best for Personal Agent
Visit
Mistral's unified work and coding agent
Mistral Vibe is Mistral's unified agent for work and coding, rebranded from Le Chat and launched on May 28, 2026
Why: Mistral Vibe brings Mistral's strong European model lineage into a competitive all-in-one agent. It is a good choice for users who want a privacy-conscious alternative to US-centric assistants with solid coding skills.
Freemium
Best for European AI Assistant
Visit
DeepSeek's open-weight model with permanent pricing
DeepSeek V4-Pro is a high-performance language model from DeepSeek
Why: DeepSeek V4-Pro stands out for combining frontier-level performance with transparent, permanent pricing and open weights. It is a practical choice for teams that want to self-host or avoid unpredictable API costs.
Freemium
Best for Predictable Pricing
Visit
The Workflow Canvas: Figma for Generative AI
Flora is a collaborative AI design canvas that moves beyond the prompt box and into node-based workflow orchestration
Why: Flora is built for more than one person working on the same graph at the same time, which most node canvases are not. If the bottleneck in your work is handing a workflow to a colleague rather than the workflow itself, that is what it solves. ComfyUI gives you more control and Invoke gives you a better editing canvas, so pick this one for the collaboration.
Freemium
Best for AI Design Workflows
Visit
Text-to-image with strong typography (varies by model)
Generates images from text prompts with exceptional typography and text rendering capabilities
Why: Great for posters, logos, and brand mockups where accurate text rendering is critical.
Freemium
Best for Images
Visit
Moonshot's specialized coding model
Kimi K2
Why: Kimi K2.7-Code is one of the strongest coding models from a Chinese AI lab, with particular strength in long-context understanding and bilingual code tasks. It is a good addition for teams evaluating global coding models.
Freemium
Best for Bilingual Coding
Visit
Image generation with workflows and models
Generates and edits images with a creator-friendly UI and extensive model library
Why: Good all-around image tool with comprehensive workflow features for concept art and production pipelines.
Freemium
Best for Images
Visit
MiniMax's 1M-context agentic frontier model
MiniMax M3 is a 1-million-token-context agentic frontier model released on May 31, 2026
Why: MiniMax M3's 1M context window makes it competitive for tasks that require digesting entire codebases, books, or video transcripts in a single pass. It is a strong option for long-context agentic applications.
Freemium
Best for 1M Context
Visit
Ultra-fast photorealistic image generation with bilingual text rendering
Generates high-quality photorealistic images from text prompts using Tongyi-MAI's Z-Image model with Single-Stream Diffusion Transformer (S3-DiT) architecture
Why: Ultra-fast photorealistic generation with superior bilingual text rendering, making it ideal for designs requiring text-in-image accuracy.
Freemium
Best for Speed
Visit
StepFun's 198B MoE vision-language model
StepFun Step 3
Why: Step 3.7 Flash offers a competitive Chinese-frontier multimodal model with an MoE architecture that balances capability and inference cost. It is a useful option for vision-language applications and for teams exploring alternatives to US models.
Freemium
Best for Efficient VLM
Visit
Design and brand image generation with vector support
Recraft V4 is a design-focused image generation model from Recraft, released in 2026
Why: Recraft V4 is built for designers rather than casual prompt users. Its emphasis on brand consistency, vector output, and editable design assets makes it unique among image generation tools.
Freemium
Best for Brand Design
Visit
Creative image workflows (and some video features)
Helps generate and refine images with creator-oriented workflows and real-time preview
Why: Good for fast creative iteration and image refinement with real-time preview and creator-focused features.
Freemium
Best for Images
Visit
The Open Image Standard: The Midjourney Killer
FLUX
Why: FLUX.2 represents the shift toward 'High-End Open Source.' We picked it because it matches Midjourney's aesthetic quality while offering the transparency and customizability that only an open-weight model can provide.
Freemium
Best for Open-Weight Quality
Visit
AI music generation with professional controls
ElevenLabs Music v2, released on May 26, 2026, is the company's next-generation AI music generator
Why: Music v2 extends ElevenLabs' voice and audio strengths into complete song generation. For creators who already use ElevenLabs for voice, it offers a natural path to full music production.
Freemium
Best for AI Music Production
Visit
Text/image-to-video with effects, transitions & swaps
Generates short videos from text prompts or images with an extensive effects library, smooth transitions between scenes, and advanced object/person/background swapping capabilities
Why: Comprehensive effects library + seamless transitions + object swapping in one platform, making it ideal for creative video work requiring multiple transformation capabilities.
Freemium
Best for Effects
Visit
Google's fast text-to-image model via Fal
Nano Banana 2 is Google's fast text-to-image model, available in part through Fal's model hosting platform
Why: Nano Banana 2 fills the need for a lightning-fast diffusion-style model from a major lab. Its availability on Fal makes it easy for developers to drop into existing inference pipelines without managing their own GPU infrastructure.
Freemium
Best for Fast Google Image Gen
Visit
AI-powered video dubbing in multiple languages
ElevenLabs Dubbing v2, released on May 28, 2026, automatically translates and dubs video content into multiple languages while preserving the original speaker's voice characteristics and lip-sync timi...
Why: Dubbing v2 makes multilingual video production far more accessible. It is especially valuable for creators, educators, and businesses that want to localize content without hiring voice actors for every language.
Freemium
Best for AI Dubbing
Visit
The first agentic IDE with Flow-state intelligence
Codeium's Windsurf is an agentic IDE that features 'Flow', a system where the AI and developer work in a continuous, shared context
Why: Windsurf is like a 'Mind-Reading Partner' for coders. It uses a special 'Flow' mode where it stays perfectly in sync with what you're doing. It doesn't just suggest code; it actually understands the 'why' behind your work and helps you fix big problems automatically.
Freemium
Best for Agentic Flow
Visit
Generative UI for React, Tailwind, and Shadcn UI
Vercel's v0
Why: v0.dev is like a 'Magic Sketchbook' for websites. You just describe what you want your site to look like, and it draws it and writes the code instantly. It's the fastest way in the world to go from a simple idea to a beautiful, working website.
Freemium
Best for Gen-UI
Visit
Full-stack web applications in the browser
Bolt
Why: Bolt.new is like an 'App Factory' in your browser. You don't need to install anything on your computer; you just tell it what app you want to build, and it builds it, runs it, and puts it on the internet for you in seconds.
Freemium
Best for MVPs
Visit
The conversational search engine that replaced traditional search
Perplexity uses frontier LLMs to browse the web in real-time and provide cited, accurate answers to complex queries
Why: Perplexity AI is the 'Death of the Search Engine.' Instead of giving you a list of 10 links to click on, it just reads the whole internet for you and gives you a single, cited answer. It's like having a personal researcher who never sleeps.
Freemium
Best for Research
Visit
AI search engine for peer-reviewed scientific research
Consensus searches over 200 million scientific papers to provide evidence-based answers
Why: The 'Truth' layer for AI. It solves the hallucination problem in research by grounding every answer in peer-reviewed science.
Freemium
Best for Science
Visit
Instant high-quality 3D modeling from text and images
Tripo AI v3 generates high-fidelity 3D meshes with clean topology and PBR textures in seconds
Why: The fastest path to 3D. Its v3 engine produces meshes that are actually usable in production pipelines without massive manual cleanup.
Freemium
Best for 3D Speed
Visit
High-fidelity 3D asset generation from Luma Labs
Genie is Luma's specialized 3D generation engine
Why: The 'Midjourney' of 3D. It prioritizes aesthetic quality and texture detail, making it the best for visual-first 3D projects.
Freemium
Best for 3D Detail
Visit
High-speed, high-realism video generation
Luma's Dream Machine v2 is a highly efficient video model known for its extreme realism and fast generation speeds
Why: Luma Dream Machine v2 is the 'Speed Demon' of AI video. It can turn a simple photo into a realistic 5-second video clip faster than almost any other tool. It's perfect for when you need to see your ideas come to life instantly.
Freemium
Best for Realism
Visit
The creative suite for physics-defying video effects
Pika 2
Why: Pika 2.0 is the 'Fun Lab' for AI video. It has special 'Pikaffects' that let you do crazy things like squish, melt, or explode objects in your videos. It's the best tool for making funny, viral videos for social media.
Freemium
Best for Viral Content
Visit
Production-ready 3D assets in under 60 seconds
Meshy v3 is the fastest text-to-3D and image-to-3D engine, producing high-topology meshes with PBR textures
Why: The bridge between AI and Game Engines. It generates usable, textured meshes that can be dropped directly into Unity or Unreal without manual cleanup.
Freemium
Best for Game Dev
Visit
Generate and refine 3D assets from text or images
Generates 3D meshes from text prompts or images using AI-powered reconstruction
Why: Meshy AI provides the fastest professional speed-to-3D workflow, enabling artists to iterate from a simple text prompt or 2D image to a usable, textured mesh in under a minute. Its high-quality PBR texture generation and clean topology make it the most efficient tool for game developers and 3D prototypers looking to bypass manual modeling bottlenecks.
Freemium
Best for 3D Assets
Visit
Multi-image to production-grade 3D on next-gen Meshy
Meshy 6 continues Meshy's focus on fast, usable 3D assets with emphasis on multi-image conditioning: feed several views or references so the model better infers shape, materials, and proportions for g...
Why: Teams outgrew 'cool sculpt from one photo' and need consistent assets from multiple references, Meshy 6 is explicitly positioned for that workflow.
Freemium
Best for Multi-Ref 3D
Visit
AI music generation with stems and inpainting
Udio v4 is the 2026 release of Udio's AI music platform, adding stem separation, audio inpainting, and more precise editing controls
Why: Udio v4 gives musicians more granular control over AI-generated music. Stems and inpainting move it closer to a real production tool rather than a one-shot generator.
Freemium
Best for Music Editing
Visit
Alibaba's strongest vision-language model
Qwen-VL-Max is a high-performance vision-language model from Alibaba, capable of understanding images, charts, and documents, and answering questions about them
Freemium
Best for Vision-Language
Visit
Open-source JSON-native text-to-image model built for controllable, enterprise-safe generation
BRIA FIBO is an 8B-parameter DiT text-to-image model trained on long structured JSON captions
Why: FIBO stands out for native JSON structured prompting and fully licensed training data, making it the strongest open-source choice for enterprises that need predictable, legally safe image generation.
Freemium
Best for Controllable Image Generation
Visit
Fast, lightweight FIBO pipeline designed for speed, efficiency, and on-prem deployment
BRIA FIBO Lite is a lightweight variant of the FIBO image generation pipeline
Why: FIBO Lite gives teams a FIBO-family option optimized for speed and data sovereignty, with a fully local deployment path that the full FIBO pipeline does not emphasize.
Freemium
Best for Fast Local Image Generation
Visit
High-accuracy background removal model trained on a licensed, professionally labeled dataset
BRIA RMBG 2
Why: RMBG 2.0 is a widely adopted, source-available background removal model with strong commercial licensing and a dedicated GitHub presence, filling a clear gap alongside BRIA's eraser tools.
Freemium
Best for Background Removal
Visit
High-volume DeepSeek inference with a 1M-token context window
DeepSeek V4-Flash is the efficient sibling of V4-Pro, offering a 1M-token context window and configurable thinking modes at a fraction of the API cost
Why: V4-Flash delivers the same 1M context and thinking modes as V4-Pro at roughly one-third the API cost, making it the practical default for most production workloads.
Freemium
Best for High-Volume APIs
Visit
The open-weight reasoning model that sparked the efficiency revolution
DeepSeek R1 is a 671B-parameter open-weight reasoning model that matches o1-class performance on math, code, and logic benchmarks through reinforcement learning on verifiable tasks
Why: R1 proved that open-weight models can match proprietary reasoning systems at a fraction of the cost, making it a landmark for reproducible AI research.
Freemium
Best for Open Reasoning
Visit
The 128K-context MoE flagship that introduced sparse attention
DeepSeek V3
Why: V3.2 introduced DeepSeek Sparse Attention and unified thinking modes, making it the architectural bridge that enabled the later 1M-context V4 family.
Freemium
Best for Long-Context MoE
Visit
Emotionally-aware multilingual text-to-speech across 29 languages
Produces natural, lifelike text-to-speech with rich emotional range and contextual understanding across 29 languages
Why: ElevenLabs' most emotionally-aware multilingual model, ideal for projects that need a consistent, expressive voice across many languages.
Freemium
Best for Emotional Multilingual TTS
Visit
Ultra-low-latency text-to-speech for real-time voice agents
Delivers high-quality speech synthesis with approximately 75ms latency across 32 languages, optimized for real-time voice agents, chatbots, interactive applications, and large-scale TTS processing
Why: The fastest ElevenLabs TTS model for production voice agents and real-time interactive experiences where latency matters.
Freemium
Best for Real-Time Voice
Visit
Real-time multilingual voice conversion that preserves emotion and content
Converts one voice into another while preserving intonation, emotion, accent, and spoken content across 29 languages
Why: A dedicated voice conversion model that keeps emotion, accent, and content intact across languages.
Freemium
Best for Voice Conversion
Visit
Generate custom synthetic voices from text descriptions
Creates entirely new synthetic voices from a text prompt describing the desired age, gender, accent, personality, and style, supporting 70+ languages
Why: Lets creators design unique voices from a written description, eliminating the need for recorded samples.
Freemium
Best for Voice Design
Visit
Google's long-context multimodal flagship with up to 2M tokens
Gemini 1
Freemium
Best for Long Context
Visit
Fast, cost-efficient multimodal model with a 1M context window
Gemini 1
Freemium
Best for Fast Multimodal Tasks
Visit
Google's low-latency agentic model with native tool use
Gemini 2
Freemium
Best for Agentic Apps
Visit
Google's high-performance reasoning model with advanced coding
Gemini 2
Freemium
Best for Complex Reasoning
Visit
OpenAI's balanced GPT-5.6 model for intelligence and cost
GPT-5
Why: Terra is the sensible default for most GPT-5.6 work: it delivers the lion's share of Sol's capability at roughly 40% of the cost and is the default model for ChatGPT Free and Go users.
Freemium
Best for Balanced Cost and Capability
Visit
OpenAI's cost-optimized GPT-5.6 model for high-volume workloads
GPT-5
Why: Luna brings GPT-5.6-scale capabilities to high-volume applications at roughly one-tenth of Sol's cost, with strong enough performance for everyday tasks and broad API availability.
Freemium
Best for Cost-Sensitive Workloads
Visit
Tencent's Mamba-powered deep-thinking reasoning model
Hybrid Mamba-Transformer MoE reasoning model released March 2025, built on Hunyuan TurboS with 52 billion active parameters and a 256K context window
Why: One of the first ultra-large Mamba-Transformer MoE reasoning models, offering strong benchmark scores and a 256K context window.
Freemium
Best for Reasoning
Visit
Tencent's fast, cost-efficient flagship Hunyuan model
A 200K-context open-weight Hunyuan model optimized for speed while maintaining strong performance on general chat, coding, and agentic tasks
Why: Speed-optimized Hunyuan flagship with a 200K context window and strong price/performance for production APIs.
Freemium
Best for Speed
Visit
Tencent's general-purpose instruction-tuned Hunyuan 2.0 model
Open-weight instruction-tuned variant of Tencent's Hunyuan 2
Why: Versatile instruction-tuned Hunyuan model balancing capability and context for a wide range of tasks.
Freemium
Best for General-Purpose Chat
Visit
The deep-thinking variant of Hunyuan 2.0
Open-weight reasoning variant of Hunyuan 2
Why: Hunyuan 2.0's reasoning mode for tasks that benefit from longer thought chains.
Freemium
Best for Reasoning
Visit
Tencent's efficient small-scale MoE instruct model
A compact open-weight MoE instruction model with a 131K context window, listed as a cost-efficient everyday Hunyuan option via OpenRouter and Tencent Cloud
Why: Smallest listed Hunyuan instruct model, making it attractive for budget-conscious long-context deployments.
Freemium
Best for Cost-Efficient Inference
Visit
Tencent's latest open-source MoE flagship with tool use
Open-weight preview of Hunyuan 3 (Hy3), a 295B-parameter MoE model with 21B active parameters and a 262K context window
Why: Tencent's strongest open-source Hunyuan model to date, with competitive coding and agentic benchmarks.
Freemium
Best for Coding and Agents
Visit
Realistic images, flexible styles, and reliable typography in one prompt
Generates photorealistic and stylized images from text prompts with a major leap in realism, prompt adherence, and text rendering over the first Ideogram model
Why: Ideogram 2.0 was the release that made Ideogram a serious alternative to Midjourney for realistic, text-heavy marketing imagery before V3 arrived.
Freemium
Best for Realistic Marketing Images
Visit
Fast, low-cost generation for rapid creative exploration
A speed-optimized variant of Ideogram 2
Why: Ideogram 2a gives creators a faster, cheaper way to produce the same text-in-image style when iteration speed matters more than pixel-perfect quality.
Freemium
Best for Fast Iteration
Visit
Moonshot's open-weight multimodal generalist with agent swarms
Kimi K2
Why: Kimi K2.5 was Moonshot's first widely available open-weight multimodal generalist and remains a notable reference point for the K2 family before K2.6 and K3 arrived.
Freemium
Best for Open Multimodal Agents
Visit
Moonshot's open-weight multimodal successor with long-context coding stability
Kimi K2
Why: Kimi K2.6 improves on K2.5 with stronger long-context coding and is a practical open-weight alternative for teams that want multimodal agents without the cost of closed frontier models.
Freemium
Best for Long-Context Coding
Visit
Faster inference variant of Kimi's coding specialist
Kimi K2
Why: Kimi K2.7 Code Highspeed is the latency-optimized version of an already strong coding model, making it a good pick for interactive coding agents and live pair-programming workflows.
Freemium
Best for Fast Coding
Visit
Kling's first widely available video generation model
Kling 1
Freemium
Best for Early Kling Video
Visit
Improved physics and expressive movement in Kling video
Kling 2
Freemium
Best for Expressive Motion
Visit
Kling's image generation model with style control
Kling Image 2
Freemium
Best for Styled Images
Visit
Controllable cinematic video model with multi-keyframe direction and motion transfer
Luma Ray 3
Why: Ray 3.2 gives professional teams frame-level control over video generation, including motion transfer and EXR export, making it a strong contender for production pipelines.
Freemium
Best for Cinematic Control
Visit
Multimodal reasoning model that generates brand-consistent images and edits
Luma Uni-1
Why: Uni-1.1 ties a reasoning model directly to pixel generation, making it unusually good at following brand references and complex creative direction in images.
Freemium
Best for Brand-Consistent Images
Visit
Fast text- and image-to-3D for concept exploration
Meshy 5 is a 2024-generation model that turns text prompts or reference images into textured 3D meshes in about 45 seconds
Why: It is the fast-iteration sibling in Meshy's current model family, still available for creators who want usable concepts in under a minute.
Freemium
Best for Fast Iteration
Visit
Clean, controllable game-ready topology in ~10 seconds
Smart Topology is Meshy's 2026 in-house model that generates 3D models with native, cleanly structured geometry and a controllable polygon count from 100 to 15,000
Why: It brings explicit polygon budgets and clean topology to AI-generated meshes, making it Meshy's most game-engine-ready model.
Freemium
Best for Game-Ready Topology
Visit
Conversational AI agent for end-to-end 3D creation
Meshy 3D Agent is a chat-first AI assistant that brainstorms, refines, and generates 3D models from text, images, or sketches inside a single ongoing conversation
Why: It preserves creative context across multiple steps, helping small teams produce stylistically consistent asset sets without restarting every prompt.
Freemium
Best for Conversational 3D Workflows
Visit
Microsoft's everyday AI assistant across web, PC, and mobile
Microsoft Copilot is the free consumer AI assistant formerly known as Bing Chat
Why: The free, broadly available Microsoft AI assistant that brings search, chat, and image generation into one cross-platform experience.
Freemium
Best for Everyday AI
Visit
Free AI design and image generation app powered by DALL-E
Microsoft Designer is a browser-based and mobile design app that generates images from text prompts, creates social graphics, and combines AI-generated visuals with templates
Why: Microsoft's free, template-driven AI design tool that pairs DALL-E image generation with practical layout tools.
Freemium
Best for Social Graphics
Visit
Recursive self-improvement language model for real-world engineering
MiniMax M2
Why: MiniMax M2.7 is the current production language model below M3 and is explicitly listed as beginning recursive self-improvement, making it a notable addition to the family.
Freemium
Best for Engineering Tasks
Visit
Same M2.7 performance with significantly faster inference
MiniMax M2
Why: The Highspeed variant is a current, actively promoted option for developers who need M2.7 capability with lower latency.
Freemium
Best for Low-Latency Coding
Visit
Ultra-realistic multilingual text-to-speech with sound tags
MiniMax Speech 2
Why: Speech 2.8 HD is the current quality-tier MiniMax voice model, replacing the earlier Speech 2.6 / Speech-02 series.
Freemium
Best for Realistic Speech
Visit
Fast multilingual text-to-speech with natural flow
MiniMax Speech 2
Why: Speech 2.8 Turbo is the current speed-tier MiniMax voice model, distinct from the HD quality variant.
Freemium
Best for Real-Time TTS
Visit
Music generation with humanized vocals and elevated sound
MiniMax Music 3
Why: Music 3.0 is the current MiniMax music generation model, replacing the legacy Music 2.0 entry already in the directory.
Freemium
Best for Vocal Music
Visit
Mistral's flagship open-weight multimodal frontier model
A 675B-parameter sparse mixture-of-experts model with 41B active parameters and a 262K context window, released under Apache 2
Why: Mistral Large 3 is one of the most capable permissive open-weight models available, offering frontier performance with the deployment flexibility of Apache 2.0 licensing.
Freemium
Best for Open-Weight Frontier
Visit
Mistral's mid-tier workhorse for reasoning, coding, and instruction
A mid-tier model that balances performance and cost, optimized for instruction following, reasoning, and coding
Why: Mistral Medium 3.5 delivers strong performance at a lower cost than the flagship, making it the sensible default for most business and development workloads.
Freemium
Best for Everyday Workloads
Visit
Unified open-source small model for chat, reasoning, vision, and coding
A 119B-parameter MoE model with 6B active parameters and a 256K context window, released under Apache 2
Why: Small 4 packs flagship-class reasoning, vision, and coding into a single open-source model that is efficient enough for high-throughput and local deployments.
Freemium
Best for Efficient Open Multimodal
Visit
Mistral's code-specialist model with fill-in-the-middle support
A code generation model optimized for latency-sensitive fill-in-the-middle completion and chat, supporting 80+ programming languages
Why: Codestral 25.08 improves accepted completions and reduces runaway generations, making it a strong open-weight option for production IDE assistants.
Freemium
Best for IDE Code Completion
Visit
Mistral's edge family of small, dense open-source models
A family of 3B, 8B, and 14B parameter dense models released under Apache 2
Why: Ministral 3 brings Mistral's open-weight lineage to edge devices, offering a strong 14B reasoning option and smaller variants for local and on-device use.
Freemium
Best for Edge Deployment
Visit
Mistral's open-weight speech understanding and TTS models
A family of open-weight speech models including a 24B production variant and a 3B edge variant, released under Apache 2
Why: Voxtral offers open-weight speech understanding and synthesis at a fraction of the cost of proprietary alternatives, making it practical for production voice agents.
Freemium
Best for Voice AI
Visit
Pika's refined model with stronger realism and camera control
Pika 2
Freemium
Best for Realistic Pika Video
Visit
Second-generation designer-first image generation model
Recraft V2 is the second-generation image generation model released by Recraft in March 2024
Why: Recraft V2 was the first generational upgrade that explicitly positioned Recraft as a designer-first model with strong style and anatomy control.
Freemium
Best for Design Assets
Visit
Recraft's most advanced image model with photorealistic, vector, and utility variants
Recraft V4
Why: Recraft V4.1 is the current flagship model, offering more natural photorealism, refined illustration quality, and dedicated Utility and Vector variants for production design workflows.
Freemium
Best for Photorealistic Design
Visit
Runway's first generation of text- and image-to-video
Runway Gen-2 is an earlier-generation video foundation model that generates short video clips from text prompts or images
Freemium
Best for Early AI Video
Visit
Faster, cheaper Gen-3 Alpha for rapid video iteration
Runway Gen-3 Alpha Turbo is a faster and more cost-efficient variant of Gen-3 Alpha, designed for creators who need to iterate quickly on video concepts without sacrificing too much quality
Freemium
Best for Fast Iteration
Visit
Image generation model with strong style control
Runway Frames is a dedicated image generation model from Runway, designed to create stylized images with strong consistency and to serve as the starting frame for video generations
Freemium
Best for Style-Locked Images
Visit
Fast, inference-efficient 1080p video with native multi-shot storytelling
Seedance 1
Why: Seedance 1.0 established the foundation for ByteDance's video generation line with a strong emphasis on inference speed and native multi-shot coherence.
Freemium
Best for Fast 1080p Video
Visit
Cinematic audio-video joint generation with lip-sync and dialect support
Seedance 1
Why: Seedance 1.5 pro moved the family from silent video to native audio-visual generation, with strong lip-sync and dialect support that makes it practical for short-form drama and advertising.
Freemium
Best for Audio-Visual Sync
Visit
Browser-based AI image enhancement workflows
Runs Topaz image enhancement tools directly in the browser with unlimited cloud rendering
Why: The no-install, browser-based entry point to Topaz image enhancement with a wide workflow menu and cloud rendering.
Freemium
Best for Browser Image Enhancement
Visit
Topaz image enhancement on iPhone
Brings Topaz Photo AI enhancement capabilities to iPhone, allowing mobile photographers to upscale, sharpen, denoise, and enhance images directly on their device
Why: Extends Topaz's photo enhancement models to iPhone, giving mobile creators access to desktop-quality AI polish.
Freemium
Best for Mobile Photo Enhancement
Visit
Earlier generation of Tripo's text- and image-to-3D pipeline
Tripo 2
Freemium
Best for Generation Pipeline
Visit
Design-forward image generation (logos, vectors, assets)
Generates design assets including logos, vectors, and brand visuals with clean, usable outputs
Why: Great for design assets when you want clean, usable outputs with vector-style graphics and brand-ready visuals.
Freemium
Best for Design
Visit
Fast text-to-3D and image-to-3D generation
Tripo v3
Why: Tripo v3.1 refines one of the fastest production-ready 3D generators. It is a practical choice for game developers, product designers, and AR/VR creators who need usable assets quickly.
Freemium
Best for Fast 3D Assets
Visit
Google's coding workhorse, three weeks after 3.6 Flash
Gemini 3
Why: It is the cheapest route to a current-generation Google coding model. Introductory pricing runs at half the standard Flash rate until the end of 2026, and the published jump over 3.6 Flash is large enough to matter on exactly the agentic and web-development work Flash-tier models are usually bought for.
Freemium
Best for Coding Value
Visit
Hyper3D's image-to-3D generation model
Rodin Gen-2, also referred to as Rodin 2
Why: Rodin Gen-2 offers a focused image-to-3D pipeline that is easy to use for artists and developers. It is a solid option when you have a 2D concept and need a 3D starting point quickly.
Freemium
Best for Image-to-3D
Visit
Advanced video editing and effects
Provides video editing, effects, and generation capabilities with advanced control using Runway's Gen-3 Alpha model
Why: Runway's latest generation model with enhanced editing features, representing the cutting edge of integrated video generation and editing.
Freemium
Best for Editing
Visit
Design suite with built-in AI generation features
Helps create designs and generate assets inside a familiar, user-friendly editor with built-in AI features
Why: Best mainstream design workflow for non-designers with intuitive interface and integrated AI generation features.
Freemium
Best for Design
Visit
Audio/video editing with AI features
Edits audio and video like a document with creator-friendly AI features including transcription, text-based editing, and automated workflows
Why: Great all-in-one editor for creators who want speed with text-based editing and AI-powered automation.
Freemium
Best for Editing
Visit
3D design tool (with AI features depending on product)
Helps design 3D scenes and assets in a browser-based workflow with real-time rendering and collaboration
Why: Great for interactive 3D design + rapid iteration with browser-based workflow and real-time collaboration features.
Freemium
Best for 3D Design
Visit
AI-powered video editing with enhancement features
Provides comprehensive video editing with AI-powered features including video enhancement, upscaling, stabilization, color correction, and frame interpolation
Why: Popular commercial video editing platform with extensive AI-powered enhancement features, widely used by content creators for professional video production.
Freemium
Best for Editing
Visit
Open-source MoE LLM with strong Chinese NLP and multimodal capabilities
Baidu ERNIE 4
Why: Leading Chinese LLM with strong multilingual capabilities, open-source availability, and cost-efficient MoE architecture.
Freemium
Best for Chinese
Visit
Advanced multilingual LLM with enhanced reasoning and long-context support
GLM-4
Why: Advanced Chinese LLM with strong multilingual capabilities, efficient inference, and comprehensive deployment options.
Freemium
Best for Multilingual
Visit
Meta's closed-weight agentic model, and its first paid model API
Muse Spark 1
Why: This is the release where Meta stopped giving models away. Muse Spark is closed, metered and sold through Meta's own API, and it lands at rank 15 on the independent index while undercutting comparable models on price. Worth tracking for that reason alone if your stack assumed Meta meant open weights.
Freemium
Best Value for Agentic Multimodal Work
Visit
Google's faster, sharper agentic-coding upgrade to 3.5 Flash
Gemini 3
Why: Gemini 3.6 Flash is the clearest upgrade path for teams already running high-volume agentic and coding workloads on Flash-tier pricing. It delivers a real benchmark jump over 3.5 Flash without moving up to Ultra-tier cost.
Freemium
Best for Fast Agentic Coding
Visit
753B open-weight MoE coding model with a 1M-token context, MIT licensed
GLM-5
Why: The strongest open-weight coding model published to date: 62.1 on SWE-bench Pro against GPT-5.5's 58.6, and 81.0 on Terminal-Bench 2.1, at roughly a sixth of GPT-5.5's API price. The MIT licence carries no regional restrictions, so the weights can genuinely be self-hosted commercially, which is the reason to choose it over a closed model of similar strength.
Freemium
Best Open-Weight Coder
Visit