BEST FOR • CURATED

Best AI Tools for AI Open Source

Best for AI Open Source

We've curated 104 top AI tools specifically selected for ai open source use cases. Each tool is evaluated for quality, reliability, and unique capabilities that make it well-suited for ai open source workflows.

WHY THESE TOOLS

These tools are selected because they excel at ai open source. When choosing, consider:

  • How the tool's specific features align with your ai open source needs
  • Whether the tool offers the right balance of quality, speed, and cost for your use case
  • Integration capabilities if you need to incorporate into existing workflows
  • Scalability for your production requirements
RESULTS
104 tools • curated
Standalone agent-first platform with CLI, SDK, and managed agents
Added May 19, 2026
AI-powered IDE built as a fork of Visual Studio Code, designed with an 'agent-first' paradigm where autonomous AI agents plan, execute, and validate code. Features two primary views: Editor View (traditional IDE with agent sidebar) and Manager View (control center for orchestrating multiple parallel agents across workspaces). Agents generate verifiable 'Artifacts' including task lists, implementation plans, screenshots, and browser recordings. Supports multiple AI models including Gemini 3 Pro, Gemini 3 Deep Think, Gemini 3 Flash, Claude Sonnet 4.5, and open-source GPT variants. Agents have direct access to editor, terminal, and integrated browser, and learn from previous interactions.
Why: Antigravity 2.0 is Google's most credible bid for the agentic IDE seat. The new CLI and SDK make it competitive with Cursor, Claude Code, and Codex for terminal-first and automation workflows.
Freemium Best for Google-Native Agents Visit
The LLM-Ready Web Scraper: Turn Websites into Markdown
Added Jan 31, 2026
Firecrawl is the industry-standard tool for turning entire websites into clean, LLM-ready markdown. It handles all the 'messy' parts of web scraping, including JavaScript rendering, proxy rotation, and anti-bot bypass, automatically. Designed specifically for AI developers, it can crawl entire domains and output structured data that is perfectly formatted for RAG (Retrieval-Augmented Generation) or fine-tuning. It acts as the bridge between the unstructured web and the structured needs of modern AI agents.
Why: Firecrawl is the leader of the 'LLM-Data' movement. We picked it because it's the first scraper that actually understands what AI models need: clean, noise-free markdown without the overhead of traditional scraping libraries.
Enterprise Best for AI Data Extraction Visit
The Invisible OS: Pure Execution via Messaging
Added Jan 27, 2026
Moltbot (also known as Clawdbot) is the spearhead of the 'Invisible OS' movement, a shift away from fragmented apps and toward pure, autonomous execution via messaging. Operating entirely through WhatsApp and Telegram, Moltbot uses advanced reasoning to manage your digital life without a traditional UI. It handles complex, multi-service tasks like clearing your inbox, coordinating calendars, and managing travel logistics (including flight check-ins) autonomously. With persistent memory and a deep 'Persona Onboarding' process, it learns your work patterns and preferences, acting as a unified reasoning layer across your existing services.
Why: Moltbot represents the death of the 'app for everything' era. We picked it because it's the first agentic assistant to prove that reasoning-based execution through simple chat is more powerful than manual task management in 10+ different apps.
Freemium Best for Agentic Automation Visit
The Open-Source Scraping Engine: High-Performance LLM Crawling
Added Jan 31, 2026
Crawl4AI is an open-source, high-performance web crawling and scraping engine specifically optimized for large language models. It provides a robust, asynchronous architecture that can handle complex JavaScript-heavy websites, dynamic content, and multi-page crawls with ease. Unlike traditional scrapers, Crawl4AI focuses on 'semantic extraction', automatically identifying the core content of a page and converting it into structured markdown or JSON that is ready for RAG pipelines. It is designed to be deeply integrated into Python-based AI workflows, offering native support for Playwright and advanced proxy management.
Why: Crawl4AI is the leading open-source alternative to proprietary scraping APIs. We picked it because it offers the most powerful 'local-first' crawling experience, giving developers full control over their data extraction pipeline without the per-page costs of cloud services.
Free Best for Open-Source Crawling Visit
The node graph the rest of the field is measured against
Added May 19, 2026
ComfyUI is an open-source node-based interface for diffusion models. Instead of a prompt box, a generation is a graph: loaders, samplers, conditioning, upscalers and masks wired together, each step inspectable and re-runnable. Workflows save as JSON and can be shared, which is why most published Stable Diffusion and FLUX pipelines circulate as ComfyUI graphs.
Why: If you want to know exactly what happened between the prompt and the image, this is the tool that shows you. Every step is a node you can open, change and re-run, and a workflow someone else built arrives as a file you can load rather than a screenshot you have to reverse engineer. That reproducibility is why it became the format the rest of the field builds around.
Free Best for control Visit
API access to thousands of models on Hugging Face
Added Feb 5, 2026
Provides API access to thousands of machine learning models hosted on the Hugging Face Hub. Supports models for text generation, image generation, audio synthesis, computer vision, and more. Simple REST API for easy integration. Pay-per-use pricing based on model and compute requirements. Includes both open-source and proprietary models. Suitable for developers wanting access to the vast Hugging Face model ecosystem without local deployment. Offers inference endpoints for production use and serverless inference for quick testing.
Why: Largest model repository with API access, making it the go-to platform for accessing diverse AI models.
Enterprise Best for Model Variety Visit
Fast inference platform for AI models
Added Feb 5, 2026
High-performance inference platform providing ultra-fast API access to large language models and other AI models. Optimized for speed using custom hardware (LPU - Language Processing Unit). Supports popular open-source models including Llama, Mixtral, Mistral, and Gemma. Offers REST API with streaming support and extremely low latency. Focuses on speed optimization, making it ideal for real-time applications. Provides dedicated endpoints for specific models and shared infrastructure. Suitable for developers needing fast inference for production applications, chatbots, and real-time AI interactions. Pay-per-use pricing with competitive rates.
Why: Fastest inference platform available, making it ideal for real-time applications requiring low latency.
Enterprise Best for Speed Visit
Moonshot AI's 2.8-trillion-parameter open-weight flagship
Added Jul 16, 2026
Kimi K3 is Moonshot AI's open-weight model released July 16, 2026, built at roughly 2.8 trillion parameters. It's available now via kimi.com, Kimi Work, Kimi Code, and the Kimi API, priced at $0.30 per million input tokens (cached), $3.00 per million (uncached), and $15.00 per million output tokens. Moonshot published the full open weights on Hugging Face on July 27, 2026, as promised at launch. Moonshot positions it as approaching Anthropic Fable 5-tier performance at a fraction of the cost, though the company acknowledges a tendency toward 'excessive proactivity' on long-running tasks.
Why: Kimi K3 is one of the most credible open-weight challengers to closed frontier models this year, aggressive enough on pricing and scale that it moved markets, Fortune covered it as a 'DeepSeek shock' moment for AI stocks.
Freemium Best for Open-Weight Frontier Performance Visit
Open-source node canvas built around the edit, not the prompt
New this month Added Aug 8, 2026
Invoke is an open-source generative image platform combining a unified canvas with inpainting, outpainting and layer control, plus a node editor for building repeatable workflows. It can be self-hosted or used through a commercial hosted tier aimed at studios.
Why: Its centre of gravity is the canvas rather than the graph, which suits the way a lot of real work happens: generate something, then keep editing regions of it. The node editor is there when a job needs repeating, instead of being the only way in.
Freemium Best for iterative editing Visit
The Efficiency Revolution: Frontier Intelligence at 1/100th the Cost
Added Feb 5, 2026
DeepSeek is the architect of the 'DeepSeek movement,' a fundamental shift in AI development that prioritizes extreme efficiency over raw compute. Founded by High-Flyer Quant, they proved that architectural innovations like Multi-head Latent Attention (MLA) and DeepSeekMoE could match the performance of $100B models like GPT-4o and Claude 3.5 while costing 95% less to train and run. Their ecosystem includes the flagship DeepSeek-V3, the reasoning-heavy DeepSeek-R1, and the state-of-the-art DeepSeek-VL2 for high-fidelity OCR and vision tasks. DeepSeek is committed to the open-source community, regularly releasing model weights and technical papers that have democratized frontier-level AI for developers globally.
Why: DeepSeek changed the game by proving that 'expensive' doesn't always mean 'better.' We picked it because it's the first model family to offer true frontier-level reasoning (R1), general intelligence (V3), and advanced vision/OCR (VL2) with an open-weight philosophy and an API price point that makes proprietary models look obsolete.
Freemium Best for Cost-Efficiency Visit
Meta's open-source large language model
Added Feb 5, 2026
Llama is Meta AI's open-source large language model family with multiple versions: Llama (February 2023), Llama 2 (July 2023), Llama 3 (April 2024), Llama 3.1 405B (405B parameters, July 2024), Llama 3.3 (December 2024), Llama 4 Maverick (April 2026), and Llama 4 Scout (April 2026). Designed for research and commercial use with strong performance across text generation, reasoning, and code tasks. Available in various sizes from 7B to 405B parameters. Supports multiple languages and extended context windows. Available through Meta's official channels, Hugging Face, and various cloud providers. Open-source licensing allows for local deployment and customization.
Why: Meta's flagship open-source LLM with strong performance, extensive model sizes, and permissive licensing for research and commercial use.
Free Best for Open Source Visit
European open-source and commercial LLM
Added Feb 5, 2026
Mistral AI provides high-performance large language models with both open-source and commercial offerings. Models include Mistral 7B, Mistral 8x7B (Mixtral), Mistral Large, Mistral Large 2.1, Mistral Small, Pixtral (multimodal, 123B parameters), and the Magistral family (June 2026) - reasoning models designed for enhanced accuracy through increased computational power during inference. Designed for efficiency and performance with strong multilingual capabilities, particularly for European languages. Offers both open-source models for local deployment and commercial API access. Available through Mistral AI's platform, Hugging Face, and various cloud providers. Strong focus on European data privacy and compliance.
Why: European LLM provider with strong open-source offerings, multilingual capabilities, and focus on data privacy and compliance.
Freemium Best for Europe Visit
Alibaba's multilingual open-source LLM
Added Feb 5, 2026
Qwen is Alibaba Cloud's family of large language models with multiple versions: Qwen-1.5 (February 2024), Qwen2 (2024), Qwen2.5 (January 3, 2026) with seven dense models from 0.5B to 72B parameters plus MoE variants, and Qwen3 (April 29, 2026) with variants Qwen3-Next, Qwen3-Max, and Qwen3-Omni focusing on context length scaling and parameter efficiency. Designed for multilingual applications with strong support for Chinese, English, and other languages. Excels at code generation, mathematical problem-solving, and structured data understanding. Pre-trained on significantly larger datasets than predecessors. Available through Alibaba Cloud API (DashScope), Hugging Face, and open-source model weights for local deployment. Offers both commercial API access and open-source licensing.
Why: Alibaba's high-performance multilingual LLM with strong Chinese language support, cost-efficient pricing, and comprehensive open-source availability.
Freemium Best for Multilingual Visit
Microsoft's efficient small language models
Added Feb 5, 2026
Microsoft Phi is a family of small, efficient language models designed for high performance with minimal parameters. Available models include Phi-1, Phi-2 (December 2023, 2.7B parameters), Phi-3 (April 2024), Phi-3.5, and Phi-4 (2026, 14B parameters) with variants: Phi-4-base, Phi-4-reasoning, Phi-4-reasoning-plus, and Phi-4-mini. Marketed as 'small language models' specializing in complex reasoning tasks. Optimized for reasoning tasks, code generation, and efficient inference. Released under MIT license for unrestricted use and modification. Available through Azure OpenAI Service, Hugging Face, and open-source model weights. Designed for edge devices, mobile applications, and cost-effective deployments.
Why: Microsoft's efficient small language models with strong reasoning capabilities, MIT licensing, and optimized for resource-constrained environments.
Free Best for Efficiency Visit
Open-Source Coding Agents for Private, Fine-Tuned Development
Added Feb 5, 2026
SERA is a family of open-source coding agents developed by the Allen Institute for AI (AI2). It allows developers to customize and fine-tune models on private codebases without exposing sensitive data to external servers. SERA uses synthetic training data to achieve performance comparable to much larger proprietary models at a significantly lower cost.
Why: We added SERA because it is the leading open-source alternative for privacy-conscious developers. It empowers teams to build their own custom coding assistants that understand their specific architectural patterns.
Free Best for Developers Visit
Google's open-source lightweight LLM
Added Feb 5, 2026
Gemma is Google DeepMind's family of open-source large language models, serving as lightweight versions of Gemini. Available models include Gemma 1 (February 2024), Gemma 2 (June 2024), and Gemma 3 (March 2026) with variants like PaliGemma for vision-language tasks and MedGemma for medical applications. Available in multiple sizes (2B, 7B, and larger variants). Designed for research, education, and commercial applications with permissive licensing. Trained on similar data and methods as Gemini models but optimized for open-source deployment. Available through Hugging Face, Kaggle, and Google Cloud Vertex AI.
Why: Google's open-source LLM family with strong performance, permissive licensing, and specialized variants for vision and medical applications.
Free Best for Research Visit
The 'Next DeepSeek' Movement: o1-Level Reasoning at 1/100th the Cost
Added Jan 31, 2026
Kimi k1.5 is a multimodal large language model from Moonshot AI, specifically engineered for high-fidelity technical reasoning and long-context processing. It is a key player in the 'DeepSeek movement,' matching the reasoning performance of frontier models like GPT-5.2 Codex and Claude 4.5 while remaining significantly more cost-effective. It features a massive 2 million token context window and joint text-vision reasoning, making it ideal for complex coding, mathematical proofs, and large-scale document analysis. The model is built using advanced Reinforcement Learning (RL) to achieve deep 'Chain-of-Thought' capabilities.
Why: Kimi k1.5 is the first model to prove that o1-level reasoning is achievable through efficient, open-weight architectures. We selected it because it consistently matches or exceeds Claude 4.5 in technical benchmarks (AIME, MATH-500) while offering a 2M context window and a significantly lower API price point, making frontier intelligence accessible to everyone.
Freemium Best for Technical Reasoning Visit
The Open Vision-Reasoner: SOTA Multimodal Performance
Added Jan 31, 2026
Qwen 2.5-VL is Alibaba's state-of-the-art open-weight multimodal model, designed to bridge the gap between open source and proprietary vision-language models. It features advanced 'NaViVi' (Native Dynamic Resolution) architecture, allowing it to process images of any resolution and videos of any length with extreme precision. It excels at complex visual reasoning, document understanding (OCR), and real-time video analysis, matching or exceeding GPT-4o in many multimodal benchmarks while remaining fully open for the community to build upon.
Why: We added Qwen 2.5-VL to the Open Frontier movement because it is currently the highest-performing open-weight vision model. It proves that open source can lead in multimodal reasoning, especially for tasks requiring high-resolution OCR and long-form video understanding.
Free Best for Open Vision Reasoning Visit
Databricks' high-performance open-source LLM
Added Feb 5, 2026
DBRX is a mixture-of-experts transformer model developed by Databricks and Mosaic ML. Released on March 27, 2024, with 132 billion total parameters (36B active parameters per token). Available in base and instruction-tuned (dbrx-instruct) variants. Outperforms other open-source models in various benchmarks including language understanding, programming, and mathematics. Uses fine-grained mixture-of-experts (MoE) architecture with 16 experts and 4 active per token for efficient inference. Trained at approximately $10 million cost. Released under Databricks Open Model License (permissive for research and commercial use). Available through Databricks Foundation Models API, Hugging Face, and open-source model weights.
Why: Databricks' high-performance open-source LLM with strong benchmark results, efficient MoE architecture, and permissive licensing.
Enterprise Best for Performance Visit
Meta's Open Multimodal Standard
Added Jan 31, 2026
Llama 3.2 Vision is Meta's first open-weight multimodal model family, bringing high-fidelity image reasoning to the Llama ecosystem. It integrates vision and text into a unified transformer architecture, enabling it to understand images, charts, and diagrams with the same ease as text. Available in 11B and 90B versions, it is designed for efficiency and edge deployment, making it the industry standard for developers building open multimodal applications that require deep reasoning and broad community support.
Why: We included Llama 3.2 Vision because it is the most widely supported open multimodal model in the world. Its integration into almost every AI tool and framework makes it the 'default' choice for open-weight vision reasoning.
Free Best for Open Ecosystem Support Visit
The Open Vision Frontier: 124B Multimodal Power
Added Jan 31, 2026
Pixtral Large is Mistral AI's flagship 124B parameter multimodal model, designed to compete directly with GPT-4o and Claude 3.5 Sonnet. Built on the Mistral Large 2 foundation, it features a native vision encoder that allows it to reason across text and images with extreme precision. It excels at complex diagram understanding, mathematical reasoning with visual context, and high-fidelity image captioning. Pixtral Large is released under the Mistral Research License, allowing developers to explore frontier-level vision-language capabilities with open weights.
Why: We added Pixtral Large because it represents the peak of European open-weight AI. It is one of the few open models that truly matches the visual reasoning depth of the top proprietary models, making it essential for the Open Frontier movement.
Freemium Best for Complex Visual Reasoning Visit
The Open-Source Vision Giant: 78B Multimodal Leader
Added Jan 31, 2026
InternVL 2.5 is a world-class open-source multimodal large language model (MLLM) that consistently tops the leaderboards for open-weight vision reasoning. It features a powerful 78B parameter architecture with a specialized vision-language alignment that excels at OCR, document understanding, and complex visual Q&A. It is designed to bridge the gap between open models and GPT-4V, offering exceptional performance across a wide range of multimodal benchmarks while remaining fully open for community development.
Why: We included InternVL 2.5 because it is a consistent leaderboard champion. It often outperforms much larger models in visual reasoning and OCR, making it a critical tool for developers who need GPT-4 level vision without the proprietary lock-in.
Free Best for Leaderboard-Topping Vision Visit
Open-source, model-agnostic terminal coding agent
Added Jul 7, 2026
OpenCode is an open-source terminal-based coding agent that connects to multiple language models and performs autonomous software engineering tasks from the command line. It edits files, runs tests, manages git workflows, and iterates on code with minimal human intervention. The v1.17.8 release from June 2026 improves tool calling reliability, multi-file refactoring, and support for local and remote model backends.
Why: OpenCode is the best 'bring-your-own-model' coding agent for developers who want full control. Because it is open source and runs in the terminal, it fits naturally into existing CI/CD and shell-centric workflows without locking you into a specific vendor.
Free Best for Terminal Coding Visit
DeepSeek's open-weight model with permanent pricing
Added May 31, 2026
DeepSeek V4-Pro is a high-performance language model from DeepSeek. The model itself was released on April 24, 2026, and permanent pricing was announced on May 31, 2026. It offers strong reasoning and coding performance at a competitive price point, with open weights available for local deployment.
Why: DeepSeek V4-Pro stands out for combining frontier-level performance with transparent, permanent pricing and open weights. It is a practical choice for teams that want to self-host or avoid unpredictable API costs.
Freemium Best for Predictable Pricing Visit
Open-source image-to-video with LoRA support
Added Feb 5, 2026
Generates high-quality videos with motion diversity from images using Wan 2.1 open-source model. Supports LoRA customization for fine-tuned control, enabling advanced users to adapt the model for specific styles and use cases. Provides full source code availability, allowing self-hosting, customization, and integration into custom workflows. Enables fine-tuning with LoRA (Low-Rank Adaptation) for specialized motion styles, character consistency, or domain-specific video generation.
Why: Open-source + LoRA customization for advanced users who need fine-tuned control and self-hosting capabilities.
Free Best for Open Source Visit
Tencent's high-quality open video model
Added Feb 5, 2026
High-quality image-to-video generation from Tencent using open-source Hunyuan Video models. Produces realistic motion, coherent scene dynamics, and production-ready video output with full source code availability. Provides open-source alternative with strong quality for self-hosting and customization. Supports both research and production use cases with comprehensive documentation and active community support. Enables complete control over the generation pipeline for advanced users.
Why: Strong open-source option with good quality, making it ideal for self-hosting and customization workflows.
Free Best for Open Source Visit
Open-source 20B model with commercial-grade text rendering and advanced image editing
Added Jan 1, 2026
Generates high-quality images from text prompts using Alibaba's Tongyi Qianwen 20-billion parameter MMDiT model. Excels at complex text rendering with commercial-grade quality, supporting multi-line layouts and paragraph-level text generation in both Chinese and English. Provides advanced image editing capabilities including style transfer, object insertion/removal, and detail enhancement. Ranks first in multiple public benchmark tests, surpassing similar open-source models with superior prompt understanding and visual quality.
Why: Top-performing open-source model with exceptional text rendering and advanced image editing capabilities, optimized for efficient deployment.
Free Best for Text Rendering Visit
The Open Image Standard: The Midjourney Killer
Added Jan 1, 2026
FLUX.2 Pro is the definitive answer to closed-source image generators like Midjourney. Developed by Black Forest Labs (the original creators of Stable Diffusion), it represents the pinnacle of high-fidelity, open-weight image generation. It features a massive 12B parameter 'Flow' architecture that produces photorealistic textures, perfect human anatomy, and industry-leading text rendering. Unlike its competitors, FLUX is built for the open-source community, supporting LoRA training, ControlNet, and local deployment, allowing creators to maintain full control over their artistic style and data.
Why: FLUX.2 represents the shift toward 'High-End Open Source.' We picked it because it matches Midjourney's aesthetic quality while offering the transparency and customizability that only an open-weight model can provide.
Freemium Best for Open-Weight Quality Visit
80B parameter open-weight coding powerhouse
Added Feb 6, 2026
Alibaba's latest open-weight model specialized for coding. At 80B parameters, it matches proprietary performance for local development and autonomous coding agents.
Why: Qwen3-Coder-Next is the best 'Private Brain' for coders. Most AI tools send your secret code to the internet, but this one can live entirely on your own computer. It's just as smart as the big paid tools, but it keeps your work 100% private and safe.
Free Best for Open Coding Visit
The first research stack built entirely by AI agents
Added Feb 6, 2026
An open-source research stack spanning Python, JS, C++, and CUDA, engineered from the ground up by autonomous AI coding agents. Optimized for high-performance tensor operations.
Why: A glimpse into the future of engineering. It's the first major technical stack where the AI wasn't just a helper, but the lead architect and builder.
Free Best for AI Research Visit
The open-source Firebase alternative with Vector support
Added Feb 5, 2026
Supabase provides a unified backend stack including a Postgres database, authentication, and storage. Their native Vector support makes it the premier choice for building RAG-based AI applications.
Why: Supabase is the 'All-in-One Toolbox' for building AI apps. It gives you a database, a way for users to log in, and a place for the AI to store its memory all in one spot. It's the easiest way to go from an idea to a working app without needing 10 different services.
Enterprise Best for Backend Visit
The global infrastructure for open-source model deployment
Added Feb 5, 2026
Now integrated with Cloudflare's global network, Replicate allows you to run and fine-tune open-source models (Flux, Llama, Whisper) with a single API call and zero infrastructure management.
Why: The 'GitHub' of model deployment. It democratizes access to the world's best open-source models with enterprise-grade scaling.
Enterprise Best for Open Source Visit
Meta's Segment Anything 3D for high-fidelity reconstruction
Added Feb 5, 2026
SAM3D v2 leverages Meta's latest Segment Anything technology to reconstruct 3D geometry from single or multiple images with extreme precision. It is the industry standard for research-grade 3D reconstruction.
Why: The most precise open-source 3D reconstruction tool. Its boundary awareness makes it unbeatable for complex object modeling.
Free Best for Research Visit
Alibaba's open-source MoE flagship with thinking modes
Added Apr 28, 2025
Qwen 3 is a 2025 open-weight Mixture-of-Experts model family from Alibaba Cloud, ranging from 0.6B to 235B parameters. It supports both thinking and non-thinking modes, strong multilingual performance, and agentic tool use, and is released under permissive licenses.
Free Best for Open-Source Agents Visit
Alibaba's open coding-specialist model
Added Nov 12, 2024
Qwen 2.5-Coder is a code-focused open-weight model from Alibaba, available in sizes from 1.5B to 32B parameters. It is optimized for code generation, completion, and debugging across many programming languages and is competitive with closed coding models.
Free Best for Open Coding Visit
Open mathematical reasoning specialist
Added Aug 1, 2024
Qwen-Math is a family of open-weight models specialized for mathematical reasoning and problem solving, derived from Qwen 2.5 and optimized on math datasets. It is available in several sizes and is competitive on math benchmarks.
Free Best for Math Reasoning Visit
Earlier open-source Wan video generation model
Added Mar 1, 2025
Wan 2.0 is an earlier open-source video generation model from Alibaba, predecessor to Wan 2.1. It supports text-to-video and image-to-video generation and laid the groundwork for the Wan family's open-source release.
Free Best for Open Video Generation Visit
Fast local FLUX.2 generation for personal hardware
Added Jun 1, 2025
FLUX.2 [schnell] is the fastest open-weights FLUX.2 variant, designed for 4-8 step local inference on consumer hardware. It retains strong prompt adherence and text rendering while being freely available for local and commercial use.
Free Best for Fast Local Generation Visit
Open-weight FLUX.2 for research and commercial use
Added Jun 1, 2025
FLUX.2 [dev] is an open-weight FLUX.2 model from Black Forest Labs, released for non-commercial and commercial research. It offers a strong balance of quality and efficiency, making it the base for many fine-tunes and community LoRAs.
Free Best for Open Customization Visit
Open-source JSON-native text-to-image model built for controllable, enterprise-safe generation
Added Jun 26, 2025
BRIA FIBO is an 8B-parameter DiT text-to-image model trained on long structured JSON captions. It turns short prompts into detailed structured schemas and generates images with precise, reproducible control over composition, lighting, camera, and color. It is also available in an image-to-image 'Inspire' mode and is trained entirely on licensed data for commercial safety.
Why: FIBO stands out for native JSON structured prompting and fully licensed training data, making it the strongest open-source choice for enterprises that need predictable, legally safe image generation.
Freemium Best for Controllable Image Generation Visit
Fast, lightweight FIBO pipeline designed for speed, efficiency, and on-prem deployment
Added Nov 11, 2025
BRIA FIBO Lite is a lightweight variant of the FIBO image generation pipeline. It pairs an open-source FIBO-VLM bridge with a smaller FIBO Lite model to enable rapid inference and fully local, on-prem deployment for privacy-critical environments.
Why: FIBO Lite gives teams a FIBO-family option optimized for speed and data sovereignty, with a fully local deployment path that the full FIBO pipeline does not emphasize.
Freemium Best for Fast Local Image Generation Visit
High-volume DeepSeek inference with a 1M-token context window
Added Apr 24, 2026
DeepSeek V4-Flash is the efficient sibling of V4-Pro, offering a 1M-token context window and configurable thinking modes at a fraction of the API cost. It is optimized for high-throughput chat, classification, bulk extraction, and agentic coding workloads, with a July 2026 update that boosted agent and coding benchmarks.
Why: V4-Flash delivers the same 1M context and thinking modes as V4-Pro at roughly one-third the API cost, making it the practical default for most production workloads.
Freemium Best for High-Volume APIs Visit
The open-weight reasoning model that sparked the efficiency revolution
Added Jan 20, 2025
DeepSeek R1 is a 671B-parameter open-weight reasoning model that matches o1-class performance on math, code, and logic benchmarks through reinforcement learning on verifiable tasks. It exposes chain-of-thought reasoning and is available as MIT-licensed local weights and via API, with the R1-0528 update in May 2025 further improving math and code reasoning.
Why: R1 proved that open-weight models can match proprietary reasoning systems at a fraction of the cost, making it a landmark for reproducible AI research.
Freemium Best for Open Reasoning Visit
The 128K-context MoE flagship that introduced sparse attention
Added Dec 1, 2025
DeepSeek V3.2 is a 128K-context mixture-of-experts model that unified thinking and non-thinking modes in the V3 line and introduced DeepSeek Sparse Attention. It served as the December 2025 flagship before V4 and remains available for self-hosting and as a historical comparison point.
Why: V3.2 introduced DeepSeek Sparse Attention and unified thinking modes, making it the architectural bridge that enabled the later 1M-context V4 family.
Freemium Best for Long-Context MoE Visit
744B-parameter open-weight MoE flagship for agentic planning and execution
Added Feb 11, 2026
GLM-5 is Zhipu AI's (Z.ai) first 2026 flagship, a 744B-parameter sparse mixture-of-experts model with roughly 40B active parameters per token. It is built for high-intelligence reasoning, agentic planning, and long-context execution, with a 200K context window and 128K maximum output.
Why: GLM-5 anchors the open-weight GLM line as a commercially-usable Chinese flagship with a permissive license and a strong reasoning profile.
Paid Best for Open-Weight Frontier Visit
MIT-licensed MoE flagship for 8-hour autonomous coding sessions
Added Apr 7, 2026
GLM-5.1 is Z.ai's refinement flagship released in April 2026, a 744B-parameter MoE model with 40B active parameters per token. It targets long-horizon agentic coding, multi-file refactoring, and terminal work, sustaining up to 8-hour autonomous tasks through a 200K context window and 128K maximum output.
Why: GLM-5.1 is the open-weight coding release that made Z.ai competitive on long-horizon agentic work while remaining MIT-licensed for unrestricted commercial use.
Paid Best for Long-Horizon Coding Visit
Google's open multimodal model for research and developers
Added Mar 12, 2025
Gemma 3 is an open-weights family of multimodal models from Google, ranging from 1B to 27B parameters. It supports text and image input, a 128K context window, and is released under a permissive license for research and commercial use.
Free Best for Open Multimodal Visit
Tencent's 389B-parameter open-source MoE language model
Added Nov 4, 2024
Open-source Transformer-based Mixture-of-Experts language model with 389 billion total parameters and 52 billion active parameters. Supports instruction-tuned and long-context pretraining checkpoints up to 256K tokens, distributed via Hugging Face and GitHub.
Why: Largest open-source Transformer-based MoE model from Tencent, ideal for researchers and builders who want to self-host a capable long-context LLM.
Free Best for Open-Source LLM Workloads Visit
Tencent's fast, cost-efficient flagship Hunyuan model
Added Jan 10, 2026
A 200K-context open-weight Hunyuan model optimized for speed while maintaining strong performance on general chat, coding, and agentic tasks. It also serves as the base for the Hunyuan T1 reasoning model.
Why: Speed-optimized Hunyuan flagship with a 200K context window and strong price/performance for production APIs.
Freemium Best for Speed Visit
Tencent's general-purpose instruction-tuned Hunyuan 2.0 model
Added Jun 20, 2025
Open-weight instruction-tuned variant of Tencent's Hunyuan 2.0 series, offering a 131K-token context window and strong everyday performance for chat, content creation, coding, and enterprise workflows.
Why: Versatile instruction-tuned Hunyuan model balancing capability and context for a wide range of tasks.
Freemium Best for General-Purpose Chat Visit
The deep-thinking variant of Hunyuan 2.0
Added Sep 15, 2025
Open-weight reasoning variant of Hunyuan 2.0 with a 131K context window, designed for complex problem-solving, math, and long-context reasoning workflows.
Why: Hunyuan 2.0's reasoning mode for tasks that benefit from longer thought chains.
Freemium Best for Reasoning Visit
Tencent's efficient small-scale MoE instruct model
Added Nov 15, 2025
A compact open-weight MoE instruction model with a 131K context window, listed as a cost-efficient everyday Hunyuan option via OpenRouter and Tencent Cloud.
Why: Smallest listed Hunyuan instruct model, making it attractive for budget-conscious long-context deployments.
Freemium Best for Cost-Efficient Inference Visit
Tencent's latest open-source MoE flagship with tool use
Added Apr 22, 2026
Open-weight preview of Hunyuan 3 (Hy3), a 295B-parameter MoE model with 21B active parameters and a 262K context window. Supports reasoning, function calling, and tool use, and is available via OpenRouter and Tencent Cloud TI Platform.
Why: Tencent's strongest open-source Hunyuan model to date, with competitive coding and agentic benchmarks.
Freemium Best for Coding and Agents Visit
Tencent's open-source bilingual text-to-image diffusion transformer
Added May 14, 2024
Open-source text-to-image diffusion transformer with fine-grained Chinese and English understanding, multi-turn prompt refinement, ControlNet, LoRA, and IP-Adapter support. Available via Hugging Face, Diffusers, ComfyUI, and a web demo.
Why: Leading open-source bilingual text-to-image model with strong Chinese prompt understanding and a rich ecosystem.
Free Best for Chinese Text-to-Image Visit
Tencent's 80B-parameter open-source MoE image generator
Added Sep 28, 2025
Native multimodal MoE image generation model with 80 billion total parameters and 13 billion active parameters. It unifies multimodal understanding and generation within an autoregressive framework, supports image editing and multi-image fusion, and is released as the largest open-source image generation model.
Why: The largest open-source image generation model, combining high parameter counts with efficient MoE inference.
Free Best for High-Resolution Image Generation Visit
Tencent's open-source immersive 3D world generator
Added Jul 26, 2025
Generates explorable, interactive 3D worlds from text prompts or images using panoramic proxies and semantic layering. Supports mesh export for game engines, VR/AR, and interactive content creation.
Why: Rare open-source pipeline for generating explorable 3D worlds from text or images, useful for games and immersive media.
Free Best for 3D Worlds Visit
Tencent's latest open-source high-fidelity 3D asset generator
Added Jun 13, 2025
Open-source system for generating high-resolution textured 3D assets from text or images. Includes a shape-generation diffusion transformer, texture synthesis pipeline, PBR support, and a Blender add-on, with mini and multiview variants for different hardware.
Why: Current open-source Hunyuan 3D pipeline with professional texture, PBR, and Blender integration.
Free Best for Production 3D Assets Visit
Moonshot's open-weight multimodal generalist with agent swarms
Added Jan 27, 2026
Kimi K2.5 is Moonshot AI's open-weight multimodal model released January 27, 2026. It supports text, image, and video input, thinking and non-thinking modes, and a 256K context window, with strong performance on agent, coding, and vision tasks. Moonshot announced the kimi-k2.5 API will be retired on August 31, 2026, so production workloads should plan a migration path to K2.6 or K3.
Why: Kimi K2.5 was Moonshot's first widely available open-weight multimodal generalist and remains a notable reference point for the K2 family before K2.6 and K3 arrived.
Freemium Best for Open Multimodal Agents Visit
Moonshot's open-weight multimodal successor with long-context coding stability
Added Apr 21, 2026
Kimi K2.6 is Moonshot AI's open-weight multimodal model released April 21, 2026. It supports text, image, and video input, thinking and non-thinking modes, and a 256K context window, with improved long-context coding stability and agent-task performance.
Why: Kimi K2.6 improves on K2.5 with stronger long-context coding and is a practical open-weight alternative for teams that want multimodal agents without the cost of closed frontier models.
Freemium Best for Long-Context Coding Visit
Meta's open-weight flagship with native multimodal reasoning
Added Apr 5, 2025
Llama 4 Maverick is Meta's flagship open-weight model, released in April 2025 as part of the Llama 4 family. It uses a Mixture-of-Experts architecture with 17B active parameters and around 400B total parameters, natively understands text and images, and supports a 1M-token context window.
Why: Maverick is the top open-weight model Meta actually ships today, with strong multimodal reasoning and a practical API ecosystem, making it the default choice for open Llama deployments.
Free Best for Open Multimodal Reasoning Visit
Long-context, efficient open multimodal model for edge and single-GPU use
Added Apr 5, 2025
Llama 4 Scout is Meta's efficient Llama 4 variant, released in April 2025. It is a Mixture-of-Experts model with 17B active parameters and 109B total parameters across 16 experts, natively multimodal for text and images, and supports an industry-leading 10M-token context window.
Why: Scout is notable for its extreme 10M-token context window and efficient single-GPU deployment, making it the standout open model for very long documents and memory-heavy applications.
Free Best for Long Context Visit
Efficient 70B open model matching 405B quality
Added Dec 6, 2024
Llama 3.3 is Meta's 70B-parameter multilingual instruction-tuned model, released in December 2024. It delivers text-only performance comparable to the much larger Llama 3.1 405B model while running far more efficiently, with a 128K-token context window and broad language support.
Why: Llama 3.3 is the practical sweet spot in the Llama family: it gives users near-frontier open-model quality in a 70B package that is far cheaper to host and fine-tune than the 405B model.
Free Best for Efficient Open LLMs Visit
The first frontier-scale open-weight language model
Added Jul 23, 2024
Llama 3.1 405B is Meta's 405-billion-parameter dense open-weight model, released in July 2024. It was the first openly available model to reach frontier-level performance on reasoning, coding, and multilingual tasks, with a 128K-token context window and native tool-use support.
Why: Llama 3.1 405B remains a landmark open release: it proved open weights could compete with proprietary frontier models and still serves as a high-quality baseline for research and synthetic-data generation.
Free Best for Frontier Open Research Visit
Mistral's flagship open-weight multimodal frontier model
Added Apr 15, 2026
A 675B-parameter sparse mixture-of-experts model with 41B active parameters and a 262K context window, released under Apache 2.0. It handles text and vision tasks, supports strong multilingual performance, and is designed for both research and enterprise deployment.
Why: Mistral Large 3 is one of the most capable permissive open-weight models available, offering frontier performance with the deployment flexibility of Apache 2.0 licensing.
Freemium Best for Open-Weight Frontier Visit
Unified open-source small model for chat, reasoning, vision, and coding
Added May 1, 2026
A 119B-parameter MoE model with 6B active parameters and a 256K context window, released under Apache 2.0. It unifies instruct, reasoning, multimodal, and agentic coding capabilities in a single efficient model with configurable reasoning effort.
Why: Small 4 packs flagship-class reasoning, vision, and coding into a single open-source model that is efficient enough for high-throughput and local deployments.
Freemium Best for Efficient Open Multimodal Visit
Mistral's edge family of small, dense open-source models
Added Apr 15, 2026
A family of 3B, 8B, and 14B parameter dense models released under Apache 2.0, optimized for performance-to-cost ratio at the edge. The 14B variant includes reasoning capabilities, making the family suitable for on-device and local deployments.
Why: Ministral 3 brings Mistral's open-weight lineage to edge devices, offering a strong 14B reasoning option and smaller variants for local and on-device use.
Freemium Best for Edge Deployment Visit
Mistral's open-weight speech understanding and TTS models
Added Jul 1, 2025
A family of open-weight speech models including a 24B production variant and a 3B edge variant, released under Apache 2.0. It supports transcription, audio understanding, summarization, Q&A, and function calling from voice with multilingual support.
Why: Voxtral offers open-weight speech understanding and synthesis at a fraction of the cost of proprietary alternatives, making it practical for production voice agents.
Freemium Best for Voice AI Visit
Compact 30B open-weight model with configurable reasoning for agents
Added Jun 4, 2026
A 30B total / 3B active parameter hybrid Mamba-2 + Transformer MoE language model built for efficient on-device and edge agentic tasks. It features a 1M-token context window, reasoning ON/OFF modes with configurable thinking budgets, and up to 4× faster throughput than its predecessor.
Why: The smallest open-weight member of the Nemotron 3 family, giving teams frontier-style reasoning and tool-use without data-center hardware.
Free Best for Efficient Agents Visit
120B open-weight hybrid MoE for efficient multi-agent reasoning
Added Jun 4, 2026
A 120B total / 12B active parameter hybrid Mamba-Transformer MoE language model with LatentMoE, multi-token prediction, and native NVFP4 pretraining. Optimized for complex multi-agent applications with a 1M-token context window and up to 5× higher throughput than the previous Nemotron Super.
Why: Fills the gap between Nano and Ultra with a strong efficiency-to-accuracy ratio for agentic orchestration and latency-sensitive serving.
Free Best for Multi-Agent Efficiency Visit
Efficient 8B physical-AI omni-model for workstations
Added Jun 1, 2026
An 8B-class physical-AI omni-model (8B reasoner + 8B generator) optimized for efficient inference on workstation-grade NVIDIA hardware such as the RTX PRO 6000. It unifies vision reasoning, world generation, and action prediction for robotics and physical AI prototyping.
Why: Brings Cosmos 3 physical-AI capabilities to smaller hardware so labs and individual developers can experiment without a data center.
Free Best for Workstation Physical AI Visit
4B physical-AI omni-model for real-time edge robotics
Added Jun 1, 2026
A 4B-class physical-AI omni-model (2B reasoner + 2B generator) optimized for real-time robotic policy and visual reasoning at the edge. It unifies world generation, vision reasoning, and action prediction in a compact form factor for embedded deployment.
Why: The smallest Cosmos 3 variant, built for real-time robotic perception and policy where latency and power matter most.
Free Best for Edge Robotics Visit
NVIDIA-aligned 253B Llama 3.1 for helpfulness and instruction following
Added Dec 1, 2024
A 253B-parameter variant of Llama 3.1 fine-tuned by NVIDIA using the HelpSteer2 datasets to improve helpfulness and instruction adherence. It is the largest member of the Llama-3.1-Nemotron family of community collaboration models.
Why: NVIDIA's largest aligned Llama collaboration, offering a strong open-weight alternative for teams already standardizing on Llama architectures.
Free Best for Aligned Llama Performance Visit
NVIDIA-aligned 49B Llama 3.1 for balanced performance
Added Dec 1, 2024
A 49B-parameter variant of Llama 3.1 fine-tuned by NVIDIA using the HelpSteer2 datasets to improve helpfulness and instruction adherence. It is the mid-size member of the Llama-3.1-Nemotron family, optimized for a strong performance-to-size ratio.
Why: A mid-size aligned Llama model that balances capability and deployment cost for teams using NVIDIA tooling.
Free Best for Balanced Llama Deployment Visit
NVIDIA-aligned 8B Llama 3.1 for efficient inference
Added Dec 1, 2024
An 8B-parameter variant of Llama 3.1 fine-tuned by NVIDIA using the HelpSteer2 datasets to improve helpfulness and instruction adherence. It is the smallest member of the Llama-3.1-Nemotron family, optimized for efficient on-device and edge deployment.
Why: A compact, NVIDIA-aligned Llama model for teams that need HelpSteer-tuned instruction following on limited hardware.
Free Best for Efficient Aligned Llama Visit
Open foundation model for humanoid robot reasoning and control
Added Jun 4, 2026
NVIDIA's open foundation model for humanoid robot reasoning and control, combining an Eagle-based vision-language backbone with a diffusion transformer (DiT) action head for language-conditioned manipulation across diverse embodiments.
Why: NVIDIA's open contribution to humanoid robot foundation models, enabling language-conditioned manipulation across diverse robot embodiments.
Free Best for Humanoid Robotics Visit
The foundational open-source text-to-image model
Added Oct 20, 2022
Stable Diffusion 1.5 is the landmark open-source latent diffusion model released by Stability AI in 2022. It established the open image generation ecosystem and remains the base for countless fine-tunes, LoRAs, and ControlNet models.
Free Best for Foundation Ecosystem Visit
High-resolution open-source image generation
Added Jul 26, 2023
Stable Diffusion XL (SDXL) is a 2023 open-source text-to-image model that generates higher-quality, higher-resolution images than SD 1.5. It uses a two-stage base-plus-refiner pipeline and is widely used for production image workflows.
Free Best for High-Resolution Open Images Visit
Fast one-step SDXL for real-time generation
Added Nov 28, 2023
Stable Diffusion XL Turbo is a distilled, fast variant of SDXL that can generate images in a single step or a few steps. It is optimized for low-latency applications and real-time interactive generation.
Free Best for Fast Open Images Visit
Efficient SD3 variant for consumer hardware
Added Jun 12, 2024
Stable Diffusion 3 Medium is a 2B-parameter version of SD3 designed to run well on consumer GPUs. It offers a balance of SD3 quality and efficiency, making it accessible for local creators.
Free Best for Local SD3 Visit
Stability AI's largest 3.5 model with best quality
Added Oct 22, 2024
Stable Diffusion 3.5 Large is an 8B-parameter open-weights model in the SD 3.5 family, offering the highest quality and best prompt adherence of the 3.5 series for demanding image generation tasks.
Free Best for SD3.5 Quality Visit
Compact open-source image-to-3D model from Microsoft
Added Dec 1, 2024
TRELLIS Mini is a smaller, faster variant of the TRELLIS family from Microsoft Research, designed for efficient image-to-3D generation on limited hardware while preserving the core architecture's quality.
Free Best for Fast Local 3D Visit
High-quality open-source image-to-3D from Microsoft
Added Dec 1, 2024
TRELLIS Large is the larger, higher-quality variant of the TRELLIS family, producing detailed 3D assets from single images with better geometry and texture fidelity than the smaller variants.
Free Best for High-Quality 3D Visit
Fast open-source image-to-3D from Stability AI and Tripo
Added Mar 5, 2024
TripoSR is an open-source, feed-forward image-to-3D model developed by Stability AI and Tripo AI. It generates textured 3D meshes from a single image in under a second on a single GPU and is released under an MIT license.
Free Best for Fast Open 3D Visit
Minimalist, container-isolated personal AI agent framework
New this month Added Aug 20, 2026
NanoClaw is an open-source personal AI agent runtime built by Gavriel Cohen as a smaller, auditable alternative to OpenClaw. Each agent session runs in its own Docker container with scoped permissions and self-destructs when the task ends. In August 2026 it added a Slack integration that lets you provision persistent AI agent teams and colleagues from a single message, running on customer infrastructure.
Why: The agent landscape is polarised between all-in-one platforms with huge codebases and small, custom rigs. NanoClaw occupies the small, auditable end: roughly 500 lines of TypeScript, container isolation by default, and a fork-and-own model that makes the agent's capabilities explicit rather than hidden behind plugins.
Free Best for Auditable Agents Visit
Open-source image generation with flexibility
Added Feb 5, 2026
Generates images from text with open-source flexibility and community support using Stable Diffusion 3.5 model. Provides extensive customization options, community models, LoRA support, and self-hosting capabilities for complete workflow control. Latest version of the Stable Diffusion ecosystem with improved quality, better prompt understanding, and enhanced capabilities. Supports local deployment, API access, and extensive community ecosystem with thousands of custom models and tools.
Why: Open-source standard with extensive customization options, making it the foundation for many custom image generation workflows.
Free Best for Open Source Visit
Microsoft Research's open image-to-3D model
Added Jul 7, 2026
TRELLIS 2 is an open-source image-to-3D generation model from Microsoft Research, released in 2026. It reconstructs 3D assets from single images or text prompts and is designed for research and experimentation.
Why: TRELLIS 2 is a valuable open research model for image-to-3D. It is ideal for academics, indie developers, and anyone who wants to run 3D generation locally or build on top of open weights.
Free Best for Open 3D Research Visit
FLUX image model family (provider site)
Added Feb 5, 2026
Publishes the FLUX family of state-of-the-art image generation models including FLUX.1, FLUX.1-dev, FLUX.2, and specialized variants. Provides open-source models with exceptional quality and prompt adherence for modern image generation workflows. FLUX models represent cutting-edge diffusion technology with superior text rendering, style control, and image quality. Offers multiple model variants optimized for different use cases including speed, quality, and specialized applications.
Why: Important modern image model family to know and track, representing the cutting edge of open-source image generation.
Best for Images Visit
Open physical-AI omnimodel for robotics and AV
Added Jul 7, 2026
NVIDIA Cosmos 3 is an open physical-AI omnimodel released around May 31 to June 1, 2026. It generates video, 3D, and physical-world simulations to train and evaluate robotics and autonomous vehicle systems without expensive real-world data collection.
Why: Cosmos 3 is a major open contribution to physical AI. By simulating realistic worlds, it can accelerate training for robots and self-driving cars while reducing the need for dangerous or costly real-world trials.
Free Best for Physical AI Simulation Visit
Intelligent routing to optimal models for 3x cost savings on inference
New Added Sep 4, 2026
Snowflake AI Gateway is an enterprise inference routing layer released September 2026. It intelligently routes queries to the most cost-effective model that meets quality requirements. Users define acceptable accuracy/latency thresholds, and the gateway automatically chooses between Claude, GPT, Gemini, or open-source alternatives based on current pricing and performance. Uses machine learning to predict which model is optimal for each query pattern. Integrated directly into Snowflake SQL and Python notebooks.
Why: Cost optimization at scale is major enterprise concern. Automatic model selection removes guesswork and prevents overpaying for high-capability models on simple tasks. Claimed 3x savings shown in practice through intelligent model tiering. Direct Snowflake integration means no architectural changes required.
Paid Best for Cost Optimization Visit
7B quantized model for offline laptop deployment, competes with Meta on-device push
New Added Sep 4, 2026
Alibaba released a 7-billion parameter language model optimized for consumer laptop deployment, released August 2026. Quantized to 4-bit with custom ONNX optimization for CPU/GPU inference. Competitive response to Meta's on-device model strategy, targeting Windows/Mac laptops with 8GB+ RAM. Open weights under OpenMDW-1.1 license. Achieves reasonable performance on everyday tasks (email drafting, code generation) while running entirely offline without cloud dependency.
Why: On-device AI becoming competitive necessity. Alibaba's direct challenge to Meta's laptop focus shows enterprise interest in consumer inference. Open weights under permissive license removes licensing friction for deployment and modification. Practical alternative for users valuing privacy and offline capability.
Free Best for On-Device Inference Visit
Open-weight video/world model, 10s clips from images in 6.8s
New Added Sep 1, 2026
LTX-2.5 is an open-weights video and world model from Lightricks (LTX company spun out of Lightricks), released August 2026. Generates 10-second video clips from images in 6.8 seconds on Nvidia superchips. Features multi-shot support, diffusion decoder for higher quality, new conditioning modes, and autoregressive models for real-time use and robotics. Weights freely available on Hugging Face under OpenMDW-1.1 license. 33 million downloads, most-used open world model line on the market. Free for organizations under $10 million annual revenue; larger companies negotiate licenses.
Why: LTX-2.5 dominates the open video/world model space (33M downloads). Open weights under permissive license, strong feature set (multi-shot, better quality, robotics support). For teams building video generation or world model workflows without proprietary constraints, this is the category leader.
Freemium Best for Open Video Generation Visit
Lightweight models with 1B+ downloads and thriving 100K+ derivative ecosystem
New Added Sep 4, 2026
Google Gemma is a family of open-source language models available in 2B, 7B, and 27B parameter sizes. Trained on 6 trillion tokens of high-quality data, designed for efficient local deployment. Includes standard and instruction-tuned variants. Reached 1 billion total downloads across Hugging Face, Kaggle, and GitHub by August 2026. Has spawned 100,000+ community derivatives (finetuned versions, specialized variants, quantizations). Available under Google DeepMind's permissive license for commercial and research use.
Why: 1B+ downloads demonstrates successful open-source adoption. 100K+ derivatives show strong community extending and adapting the models. Covers efficiency needs from edge (2B) to capabilities (27B). Strong proof that open models achieve massive scale in production deployments.
Free Best for Open Source Development Visit
OpenAI's conditional 3D model generation
Added Feb 5, 2026
Generates 3D objects from text prompts or images using OpenAI's Shap-E model, a conditional generative model for 3D assets. Produces high-quality 3D meshes, point clouds, and neural radiance fields (NeRFs) from natural language descriptions. Supports both text-to-3D and image-to-3D workflows, generating detailed 3D models with realistic geometry and textures suitable for game assets, product visualization, and 3D printing applications. Open-source model with comprehensive documentation and active community support, making it ideal for research, prototyping, and educational use.
Why: OpenAI's open-source 3D generation model with comprehensive documentation and active community, representing state-of-the-art conditional 3D asset generation from text and images.
Free Best for Research Visit
30B on-device AI agent, runs natively on consumer hardware
New Added Sep 1, 2026
Muse Glimmer is Meta's 30 billion parameter open-weight model released August 2026, designed to run autonomous AI agents directly on consumer hardware without cloud dependency. It ships under the permissive Apache 2.0 license (Meta's first fully open release since moving to proprietary Muse Spark in April), quantized to roughly 4 bits with block-level speculative decoding for fast local inference. Drops the 700M monthly user restriction that hampered prior Llama releases.
Why: Open weights on Apache 2.0, runs on consumer hardware, removes user-count restrictions that hampered Llama. For teams building local-first or edge-deployed agents, this eliminates licensing friction and eliminates inference costs.
Free Best for On-Device Agents Visit
118B MoE model beats rivals 10x its size on coding benchmarks
New Added Sep 1, 2026
Laguna S 2.1 is a 118 billion parameter Mixture-of-Experts model from Poolside AI, released August 2026. Activates only 8 billion parameters per token, supports context windows up to 1 million tokens, runs under permissive OpenMDW-1.1 license. Benchmarks: 70.2% on Terminal-Bench 2.1 (beating DeepSeek-V4-Pro-Max, Nvidia Nemotron 3 Ultra), 78.5% on SWE-Bench Multilingual. Trained in under 9 weeks on 4,096 Nvidia H200 GPUs.
Why: Benchmark contender: claims to beat models many times its size on two critical coding benchmarks. Open weights under permissive license. Rapid training timeline (9 weeks) suggests efficient engineering. Strong SWE-Bench showing makes it worth evaluating for coding agent workloads.
Free Best for Coding Performance Visit
Open-source MoE LLM with strong Chinese NLP and multimodal capabilities
Added Jan 1, 2026
Baidu ERNIE 4.5 (Enhanced Representation through Knowledge Integration) is a family of large language models released by Baidu in November 2026. The ERNIE 4.5 model family includes 10 variants ranging from 0.3 billion to 424 billion total parameters, utilizing a Mixture-of-Experts (MoE) architecture for efficient inference. Open-sourced under the Apache 2.0 license in June 2026, ERNIE 4.5 demonstrates strong performance in Chinese natural language processing, multimodal understanding, and various AI benchmarks. The model excels in common-sense reasoning, optical character recognition, and Chinese language tasks. Available through ERNIE Bot (web interface) and Baidu's Qianfan platform (API access), with open-source model weights available for local deployment.
Why: Leading Chinese LLM with strong multilingual capabilities, open-source availability, and cost-efficient MoE architecture.
Freemium Best for Chinese Visit
Advanced multilingual LLM with enhanced reasoning and long-context support
Added Jan 1, 2026
GLM-4.5 (General Language Model) is Zhipu AI's latest large language model in the ChatGLM/GLM series, released in 2026. Building on the success of previous GLM models, GLM-4.5 offers enhanced reasoning capabilities, improved multilingual support (with strong Chinese and English capabilities), and advanced instruction-following. The model is designed for both chat and completion tasks, with support for long context windows and fine-tuned variants for specific use cases. GLM-4.5 maintains Zhipu AI's focus on efficient inference and cost-effective deployment. Available through Zhipu AI's platform (web interface and API) with options for local deployment of open-source variants.
Why: Advanced Chinese LLM with strong multilingual capabilities, efficient inference, and comprehensive deployment options.
Freemium Best for Multilingual Visit
Z.ai's post-trained coding and agentic model on the GLM-5.2 base
New this month Added Aug 14, 2026
GLM-5.3 is Z.ai's flagship coding and agentic model, released 14 August 2026. It uses the same 743B-parameter mixture-of-experts base as GLM-5.2, with Z.ai attributing the gains to scaled post-training rather than a new pre-training run. It targets long-horizon agentic coding, business-process automation, defensive security work and tasks that span many steps. The model supports three reasoning-effort levels and a 1M-token route for coding plans.
Why: The Terminal-Bench 3.0 score moved from 4.6% to 28.3%, DeepSWE v1.1 from 46.2% to 66.9%, and CyberGym from 77.2% to 84.5% — all on the same base as GLM-5.2. That is a real signal about post-training returns, even if most numbers are vendor-run and weights are not yet released. It is also reported as one of the fastest models in its class, at roughly 115 tokens per second.
Paid Best for Post-Training Gains Visit
Open-source text-to-3D motion model with 200+ motion categories and production-ready exports
Added Jan 1, 2026
Hymotion 1.0 (also known as HY-Motion 1.0 or Hunyuan Motion 1.0) is Tencent's open-source, billion-parameter text-to-3D motion generation model released in December 2026. Built on a Diffusion Transformer (DiT) architecture with flow matching, it generates high-fidelity, smooth, and diverse 3D character animations from natural language descriptions. Trained on over 3,000 hours of diverse motion data covering 200+ motion categories including locomotion, sports, fitness, social interactions, and daily activities. The model employs a three-stage training paradigm: large-scale pretraining, high-quality fine-tuning with 400 hours of curated text-motion pairs, and reinforcement learning for physical plausibility. Supports standard 3D formats (FBX, BVH, GLTF) for seamless integration with...
Why: Tencent's cutting-edge open-source text-to-3D motion model with production-ready output and extensive motion category support.
Free Best for 3D Motion Visit
NVIDIA's 550B open-weights reasoning model, built for inference speed
Added Aug 4, 2026
Nemotron 3 Ultra is the largest member of NVIDIA's Nemotron 3 family, released 4 June 2026 at Computex. It is a 550B-parameter hybrid latent mixture of experts with roughly 55B parameters active per token, combining Mamba and Transformer blocks, and trained in NVIDIA's 4-bit NVFP4 format on Blackwell hardware. It ships under the NVIDIA Open Model License, which permits commercial use, and NVIDIA published training data, reinforcement learning environments and post-training recipes alongside the weights rather than the weights alone. Weights are on Hugging Face as nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B in both BF16 and NVFP4, and it is served through OpenRouter, Together AI, Baseten, DeepInfra, Fireworks and NVIDIA NIM.
Why: The architecture is optimised for throughput rather than peak benchmark score, and it shows: over 400 output tokens per second at 550B parameters. It is also the most openly documented release at this scale, because publishing the datasets and post-training recipes lets you actually reproduce and extend the model instead of just running it.
Free Best for Open-Weight Throughput Visit
Multimodal model generating image, video and audio from one set of weights
Added Aug 4, 2026
FLUX 3 is Black Forest Labs' multimodal foundation model, announced 23 July 2026. Unlike the FLUX.1 and FLUX.2 image models before it, FLUX 3 learns jointly across images, video and audio in a single unified architecture: it generates video with native synchronised audio, edits images, renders readable text, and, via a FLUX-mimic variant, predicts robot actions, all from the same weights. Video generation runs up to 20 seconds. At launch, video is available through a gated early-access programme, with image generation stated to follow and an open-weight FLUX 3 Dev backbone planned later.
Why: The first credible attempt to collapse image, video and audio generation into a single model rather than a pipeline of separate ones, from the team behind the most widely self-hosted open image models. Access is the catch: video is gated early-access and the open-weight release has not shipped, so treat availability as limited until FLUX 3 Dev lands.
Paid Best Multimodal Generation Visit
Alibaba's 2.4-trillion-parameter flagship, currently in preview
Added Jul 19, 2026
Qwen 3.8-Max is Alibaba's largest model yet, announced July 19, 2026, at 2.4 trillion total parameters using a sparse Mixture-of-Experts design. It's multimodal (text, images, video, documents) with a context window in the ~1M-token range (983,616 tokens per Qwen Cloud metadata) and a 131,072-token max output. It's live now as qwen3.8-max-preview through Alibaba's Token Plan, Qoder, and QoderWork at 10% of eventual standard pricing, targeting coding, agentic workflows, and long-horizon 'professional cowork' tasks. Alibaba says open weights are coming but hasn't published a date, license, model card, or full benchmark table yet; the independent number available is Artificial Analysis, which places the preview at 53.4 on its Intelligence Index, rank 11.
Why: Qwen 3.8-Max is Alibaba's answer to the current wave of massive open-weight-adjacent models from Chinese labs, and the discounted preview pricing makes it worth evaluating early even before the full release details land.
Paid Best for Long-Horizon Agentic Work (Preview) Visit
753B open-weight MoE coding model with a 1M-token context, MIT licensed
Added Aug 4, 2026
GLM-5.2 is Zhipu AI's (Z.ai) flagship open-weight model, released 13 June 2026 under an MIT licence with weights published on Hugging Face at zai-org/GLM-5.2. It is a 753B-parameter mixture-of-experts model activating roughly 40B parameters per token, with a 1M-token context window and 128K maximum output. The headline architectural change is IndexShare, which reuses the same indexer across every four sparse attention layers; Z.ai reports this cuts per-token compute by about 2.9x at full 1M context. It targets long-horizon agentic coding, multi-file refactors, terminal work, and tasks that run for many steps rather than single completions.
Why: The strongest open-weight coding model published to date: 62.1 on SWE-bench Pro against GPT-5.5's 58.6, and 81.0 on Terminal-Bench 2.1, at roughly a sixth of GPT-5.5's API price. The MIT licence carries no regional restrictions, so the weights can genuinely be self-hosted commercially, which is the reason to choose it over a closed model of similar strength.
Freemium Best Open-Weight Coder Visit
Thinking Machines' 975B Apache-2.0 model that takes text, images and audio natively
Added Aug 4, 2026
Inkling is the first open-weights model from Thinking Machines Lab, released 15 July 2026 under Apache 2.0. It is a 975B-parameter mixture of experts with 41B active per token, trained from scratch on 45 trillion tokens of text, images, audio and video, with a context window up to 1M tokens. The architecture uses 256 routed experts plus 2 shared experts per layer with 6 routed experts active per token, a sigmoid router, and interleaved sliding-window and global attention at a 5:1 ratio. It accepts text, image and audio input natively and returns text. Weights are on Hugging Face in both the original format and an NVFP4 checkpoint for Blackwell hardware. A distilled Inkling-Small followed on 31 July.
Why: It is the first roughly trillion-parameter open-weights model that takes audio and images natively rather than through a bolted-on encoder, and Apache 2.0 means the weights can be used commercially without asking anyone. Thinking Machines is candid that this is not the strongest model available but a base worth customising, which is a more honest pitch than most open releases make.
Free Best Open-Weight Multimodal Base Visit