AI Model Release Tracker
Curated foundation model and API releases across modalities. Updated as new drops are added to the directory.
July 2026 (9)
OpenCode
Jul 7, 2026Open-source, model-agnostic terminal coding agent
Kimi K2.7-Code
Jul 7, 2026Moonshot's specialized coding model
TRELLIS 2
Jul 7, 2026Microsoft Research's open image-to-3D model
Microsoft MAI Models (Build 2026)
Jul 7, 2026Microsoft's unified AI model family from Build 2026
NVIDIA Cosmos 3
Jul 7, 2026Open physical-AI omnimodel for robotics and AV
Claude Fable 5
Jul 7, 2026Anthropic's Mythos-class creative model
Gemini 3.6 Flash
Jul 21, 2026Google's faster, sharper agentic-coding upgrade to 3.5 Flash
Kimi K3
Jul 16, 2026Moonshot AI's 2.8-trillion-parameter open-weight flagship
Qwen 3.8-Max
Jul 19, 2026Alibaba's 2.4-trillion-parameter flagship, currently in preview
May 2026 (26)
GPT-5.5 Instant
May 5, 2026OpenAI's fast default ChatGPT model from May 2026
Cursor Composer 2.5
May 18, 2026Cursor's agentic coding model for multi-file software engineering
Gemini 3.5 Flash
May 19, 2026Google's fast, capable multimodal model from I/O 2026
Grok Build
May 14, 2026xAI's agentic coding CLI for autonomous software engineering
Claude Opus 4.8
May 28, 2026Anthropic's powerful enterprise model from May 2026
Gemini Omni
May 19, 2026Google's unified multimodal generation model
Runway Aleph 2.0
May 21, 2026In-context video editing model and Edit Studio
Gemini Spark
May 19, 2026Google's personal AI agent for proactive assistance
Mistral Vibe
May 28, 2026Mistral's unified work and coding agent
DeepSeek V4-Pro
May 31, 2026DeepSeek's open-weight model with permanent pricing
GPT-Image-2
May 10, 2026OpenAI's latest image generation model
FLUX.2 [max]
May 15, 2026Black Forest Labs' top-tier image generation model
MiniMax M3
May 31, 2026MiniMax's 1M-context agentic frontier model
StepFun Step 3.7 Flash
May 29, 2026StepFun's 198B MoE vision-language model
Recraft V4
May 20, 2026Design and brand image generation with vector support
ElevenLabs Music v2
May 26, 2026AI music generation with professional controls
Nano Banana 2
May 22, 2026Google's fast text-to-image model via Fal
ElevenLabs Dubbing v2
May 28, 2026AI-powered video dubbing in multiple languages
Udio v4
May 15, 2026AI music generation with stems and inpainting
Tripo v3.1
May 25, 2026Fast text-to-3D and image-to-3D generation
Rodin Gen-2
May 27, 2026Hyper3D's image-to-3D generation model
GPT Image 2
May 3, 2026OpenAI quality-first image generation and editing
HappyHorse 1.0
May 3, 2026Alibaba flagship video with joint audio and multilingual lip-sync
NVIDIA Nemotron 3 Nano Omni
May 3, 2026One multimodal model for text, vision, audio, and video reasoning
Meshy 6
May 3, 2026Multi-image to production-grade 3D on next-gen Meshy
Decart Lucy 2.1 VTON
May 3, 2026Real-time virtual try-on in video
February 2026 (105)
Claude Opus 4.6
Feb 6, 2026The ceiling of enterprise autonomy with 1M context
NotebookLM
Feb 4, 2026Google's AI Research Assistant: The Ultimate Study Tool
Kling AI
Feb 6, 2026The frontier of cinematic video synthesis
Grok
Feb 5, 2026xAI's real-time AI assistant
DeepSeek
Feb 5, 2026The Efficiency Revolution: Frontier Intelligence at 1/100th the Cost
Suno
Feb 5, 2026Text-to-music & vocals with fast iteration
Llama
Feb 5, 2026Meta's open-source large language model
Mistral AI
Feb 5, 2026European open-source and commercial LLM
ElevenLabs
Feb 5, 2026High-quality TTS and voice tools
Cohere
Feb 5, 2026Enterprise-focused LLM platform
Qwen
Feb 5, 2026Alibaba's multilingual open-source LLM
Microsoft Phi
Feb 5, 2026Microsoft's efficient small language models
Gemma
Feb 5, 2026Google's open-source lightweight LLM
DBRX
Feb 5, 2026Databricks' high-performance open-source LLM
Resemble AI
Feb 5, 2026Voice generation and cloning tools
Veo 3.1
Feb 5, 2026Google's state-of-the-art video generation model
Sora 2
Feb 5, 2026OpenAI's state-of-the-art video model with audio
Kling 2.6 Pro
Feb 5, 2026Top-tier image-to-video with native audio generation
Kling AI
Feb 5, 2026Text/image-to-video generation (availability varies)
Runway
Feb 5, 2026Text/image-to-video creation suite with editing tools
Pika
Feb 5, 2026Text/image-to-video with Pikaffects (squish, melt, explode)
HeyGen
Feb 5, 2026Avatar and talking-head video generation
Ray2 Flash
Feb 5, 2026Fast video generation from Luma Dream Machine
Synthesia
Feb 5, 2026AI avatar video creation for teams
Hailuo 2.3 Fast
Feb 5, 2026Fast 1080p image-to-video from MiniMax
Midjourney
Feb 5, 2026High-end image generation with strong aesthetics
OmniHuman v1.5
Feb 5, 2026Audio-driven human animation from ByteDance
D-ID
Feb 5, 2026Talking avatar videos from images and scripts
Wan 2.1
Feb 5, 2026Open-source image-to-video with LoRA support
Hunyuan Video
Feb 5, 2026Tencent's high-quality open video model
Ideogram
Feb 5, 2026Text-to-image with strong typography (varies by model)
Leonardo AI
Feb 5, 2026Image generation with workflows and models
Wan 2.6 Text-to-Video
Feb 5, 2026Latest Wan model for text-to-video generation
Kaiber
Feb 5, 2026Stylized image/video animation for creators
Hunyuan Video 1.5
Feb 5, 2026Tencent's latest text-to-video model
Adobe Firefly
Feb 5, 2026Generative image tools inside Adobe ecosystem
LTX-2
Feb 5, 2026Fast text-to-video with audio support
Hunyuan 3D
Feb 5, 2026Tencent's high-quality 3D generation engine
Seedance 2.0
Feb 12, 2026ByteDance's Next-Gen Video Model with Native Audio-Video Joint Generation
Krea
Feb 5, 2026Creative image workflows (and some video features)
PixVerse
Feb 5, 2026Text/image-to-video with effects, transitions & swaps
Vidu Q2
Feb 5, 2026Shengshu's advanced image-to-video with better control
Viggle
Feb 5, 2026Character motion and meme-style video creation
GPT-Image 1.5
Feb 5, 2026OpenAI's high-fidelity image generation
Meshy AI
Feb 5, 2026Generate and refine 3D assets from text or images
Flux 2 Flex
Feb 5, 2026Fine-tuned control with adjustable inference
Recraft
Feb 5, 2026Design-forward image generation (logos, vectors, assets)
Flux Kontext
Feb 5, 2026Context-aware image generation and editing
Stable Diffusion 3.5
Feb 5, 2026Open-source image generation with flexibility
Magnific
Feb 5, 2026AI upscaling and enhancement for images
Wan 2.6 Image-to-Image
Feb 5, 2026Latest Wan for image variations and editing
Black Forest Labs
Feb 5, 2026FLUX image model family (provider site)
BRIA Eraser
Feb 5, 2026High-fidelity object removal from images
Chatterbox Turbo
Feb 5, 2026Ultra-fast text-to-speech for real-time voice AI
Microsoft TRELLIS
Feb 5, 2026Microsoft's advanced 3D generation from text or images
BRIA Video Eraser
Feb 5, 2026Object removal from video with high fidelity
Kaedim
Feb 5, 20262D-to-3D conversion for game assets
LightX Recamera
Feb 5, 2026Relight and recamera videos
3DFY.ai
Feb 5, 2026Text-to-3D for product-style assets
Runway Gen-3 Alpha
Feb 5, 2026Advanced video editing and effects
MiniMax Music 2.0
Feb 5, 2026Advanced AI music generation with high-quality compositions
Stable Audio 2.5
Feb 5, 2026High-quality music and sound effects generation
Luma AI
Feb 5, 20263D capture + creative tools (incl. 3D/Video features)
ElevenLabs TTS Eleven-v3
Feb 5, 2026Multilingual text-to-speech with natural voice synthesis
Stable Diffusion
Feb 5, 2026Open image generation ecosystem (model + tools)
Lyria 2
Feb 5, 2026Google's latest music generation model
Canva
Feb 5, 2026Design suite with built-in AI generation features
Sonauto v2.2
Feb 5, 2026CD-quality music with superior vocals
Descript
Feb 5, 2026Audio/video editing with AI features
ElevenLabs Sound Effects v2
Feb 5, 2026Advanced sound effects generation
Flux 1 [schnell]
Feb 5, 2026Fast Flux variant for rapid image generation
Imagen 3
Feb 5, 2026Google's high-quality text-to-image model
Recraft V3
Feb 5, 2026Vector art and brand-style image generation
Ideogram V3
Feb 5, 2026Exceptional typography and text rendering
Topaz Photo AI
Feb 5, 2026Image enhancement (denoise/sharpen/upscale)
Flux 1 [dev]
Feb 5, 2026Development Flux for advanced control
Spline
Feb 5, 20263D design tool (with AI features depending on product)
Ovis Image
Feb 5, 2026Quick text rendering for marketing graphics
LongCat Image
Feb 5, 2026Multilingual text rendering and photorealism
Bagel
Feb 5, 20267B multimodal model for text and images
Flux Realism LoRA
Feb 5, 2026Photorealistic Flux with LoRA fine-tuning
Flux LoRA
Feb 5, 2026Customizable Flux with LoRA fine-tuning
MiniMax TTS
Feb 5, 2026Multilingual text-to-speech with streaming
Shap-E
Feb 5, 2026OpenAI's conditional 3D model generation
Point-E
Feb 5, 2026OpenAI's fast point cloud generation
DreamFusion
Feb 5, 2026Text-to-3D via NeRF with score distillation
Get3D
Feb 5, 2026NVIDIA's high-quality 3D mesh generation
Topaz Video Enhance AI
Feb 5, 2026Professional video upscaling and enhancement
CapCut
Feb 5, 2026AI-powered video editing with enhancement features
Zero-1-to-3
Feb 5, 2026View-consistent image-to-3D generation
Instant3D
Feb 5, 2026Fast single-image 3D generation
Qwen3-Coder-Next
Feb 6, 202680B parameter open-weight coding powerhouse
GPT-5.3 Codex
Feb 6, 2026The frontier model for complex reasoning and software architecture
Claude 4.6 Sonnet
Feb 6, 2026The industry standard for coding and nuanced instruction following
Runway Gen-4.5
Feb 6, 2026The industry standard for cinematic AI video generation
Gemini 3 Ultra
Feb 5, 2026Native multimodal intelligence with a 10M context window
Perplexity AI
Feb 5, 2026The conversational search engine that replaced traditional search
Consensus
Feb 5, 2026AI search engine for peer-reviewed scientific research
Tripo AI v3
Feb 5, 2026Instant high-quality 3D modeling from text and images
Luma Genie
Feb 5, 2026High-fidelity 3D asset generation from Luma Labs
Luma Dream Machine v2
Feb 5, 2026High-speed, high-realism video generation
FLUX.1 [pro]
Feb 5, 2026The new gold standard for prompt adherence and text rendering
Pika 2.0
Feb 5, 2026The creative suite for physics-defying video effects
Meshy AI v3
Feb 5, 2026Production-ready 3D assets in under 60 seconds
SAM3D v2
Feb 5, 2026Meta's Segment Anything 3D for high-fidelity reconstruction
January 2026 (13)
Flora
Jan 31, 2026The Workflow Canvas: Figma for Generative AI
Kimi k1.5
Jan 31, 2026The 'Next DeepSeek' Movement: o1-Level Reasoning at 1/100th the Cost
Qwen 2.5-VL
Jan 31, 2026The Open Vision-Reasoner: SOTA Multimodal Performance
Llama 3.2 Vision
Jan 31, 2026Meta's Open Multimodal Standard
Pixtral Large
Jan 31, 2026The Open Vision Frontier: 124B Multimodal Power
InternVL 2.5
Jan 31, 2026The Open-Source Vision Giant: 78B Multimodal Leader
Z-Image
Jan 1, 2026Ultra-fast photorealistic image generation with bilingual text rendering
Qwen-Image
Jan 1, 2026Open-source 20B model with commercial-grade text rendering and advanced image editing
FLUX.2 Pro
Jan 1, 2026The Open Image Standard: The Midjourney Killer
Baidu ERNIE 4.5
Jan 1, 2026Open-source MoE LLM with strong Chinese NLP and multimodal capabilities
GLM-4.5
Jan 1, 2026Advanced multilingual LLM with enhanced reasoning and long-context support
Hymotion 1.0
Jan 1, 2026Open-source text-to-3D motion model with 200+ motion categories and production-ready exports
Manus AI
Jan 1, 2026Autonomous AI agent for complex multi-step workflows and research automation