Added May 19, 2026
AI-powered IDE built as a fork of Visual Studio Code, designed with an 'agent-first' paradigm where autonomous AI agents plan, execute, and validate code. Features two primary views: Editor View (traditional IDE with agent sidebar) and Manager View (control center for orchestrating multiple parallel agents across workspaces). Agents generate verifiable 'Artifacts' including task lists, implementation plans, screenshots, and browser recordings. Supports multiple AI models including Gemini 3 Pro, Gemini 3 Deep Think, Gemini 3 Flash, Claude Sonnet 4.5, and open-source GPT variants. Agents have direct access to editor, terminal, and integrated browser, and learn from previous interactions.
Why: Antigravity 2.0 is Google's most credible bid for the agentic IDE seat. The new CLI and SDK make it competitive with Cursor, Claude Code, and Codex for terminal-first and automation workflows.
Added Jul 7, 2026
Claude Fable 5 is Anthropic's Mythos-class model released on June 9, 2026, focused on creative writing, worldbuilding, and narrative depth. It was suspended from distribution on June 12, 2026, under US export controls, making it a limited-availability release.
Why: Claude Fable 5 is notable as Anthropic's most experimental creative model. Even with limited availability, it represents an interesting direction for AI-assisted fiction and long-form creative work.
Added Sep 4, 2026
Claude Fable 5.1 is Anthropic's Fable-tier creative and reasoning model with integrated text watermarking. Released September 2026 as an update to Claude Fable 5, it features automatic watermarking of all text outputs plus a public detection API to verify AI-generated content. Watermarks are imperceptible to users but machine-verifiable, meeting EU AI Act transparency requirements. Priced competitively as a fast model suitable for high-volume applications.
Why: Watermarking + detection API represents significant move toward AI-generated content provenance. EU AI Act compliance built-in from launch, critical for enterprises. Fast execution speed with transparency features addresses emerging regulatory requirements without sacrificing performance.
New this month
Added Sep 9, 2026
GPT-6 Astra is OpenAI's frontier model released September 2026, claimed by the company to mark the onset of artificial general intelligence (AGI). The model can navigate software autonomously without user input, working across browsers, spreadsheets, applications, and producing finished documents and workflows. Features advanced 'recurrent depth' reasoning that operates outside traditional sequential thinking patterns. Solves decades-old mathematics problems and demonstrates autonomous agent capabilities across domains.
Why: Represents a claimed inflection point in AI autonomy. Demonstrates system-level reasoning and multi-step task completion. First model to achieve 'Critical' level cybersecurity capabilities (restricted access). Raises important questions about AI development velocity and safety infrastructure.
Added Feb 4, 2026
NotebookLM is an AI-first research and study assistant grounded in your own documents. Unlike generic chatbots, it only answers based on the sources you upload (PDFs, Google Docs, Slides, Websites), making it hallucination-resistant. It features 'Audio Overview,' which turns your notes into an engaging, podcast-style discussion between two AI hosts. It allows you to 'chat' with your documents, generate summaries, and find connections across multiple sources instantly.
Why: The Audio Overview feature is a viral sensation for a reason: it transforms dry study material into an engaging podcast. It is arguably the best free AI study tool available today.
Added Aug 4, 2026
Claude Opus 5 is Anthropic's flagship model, released 24 July 2026 with a 1M-token context window and five selectable effort levels (low, medium, high, xhigh, max). Effort is the main control: output token spend runs roughly 8x from low to max, and Artificial Analysis measures a 407-Elo spread in task quality across that range, so the same model behaves like several different price and capability tiers. API pricing is $5 per million input tokens and $25 per million output, with cache writes at $6.25 and cache hits at $0.50. It leads the Intelligence Index at 60.7 and tops the Coding Agent Index, scoring 89 percent on Terminal-Bench v2.1 at max effort and 53 percent on Humanity's Last Exam.
Why: It is the current number one on the independent Artificial Analysis Intelligence Index, and it got there while costing less per task than the model it displaced: $2.03 average per index task against Fable 5's $2.75. The effort dial is the reason to pick it over a fixed-tier model, because one integration covers cheap high-volume calls and expensive long-horizon agent runs.
New this month
Added Sep 22, 2026
Claude Opus 5.5 is Anthropic's September 22, 2026 update to Opus 5, priced at $4 per million input tokens and $20 per million output (down from $5/$25), with cache reads cut from $0.50 to $0.20 per million and cache writes from $6.25 to $5. A faster mode in Claude Code and the Claude Platform runs up to 2.5x quicker at $8/$40 per million tokens. Anthropic's own benchmarks put it ahead of the larger Claude Fable 5.1 on agentic coding (66.4% vs 55.8% on Terminal-Bench 4.0), computer use (81.8% vs 80.7% on OSWorld 2.0), and knowledge work (1846 vs 1735 Elo on GDPval-AA v2.1), while generating output more than 30% faster than Opus 5. Five-hour usage limits on Pro, Max, Team and seat-based Enterprise plans were also increased alongside the release.
Why: It is priced 20% below Opus 5 on the headline rate and roughly 40% cheaper on typical agentic workloads once cache-read savings are counted, while benchmarking ahead of the more expensive Fable 5.1 on the coding and computer-use tasks this directory's readers actually run. That combination, cheaper than its own predecessor and ahead of the tier above it, is unusual enough to be the pick over Opus 5 for anyone already on Claude.
Added Jul 9, 2026
GPT-5.6 Sol is the highest-capability model in OpenAI's GPT-5.6 family, released July 9, 2026 after a limited preview starting June 26. It has a 1.05M-token context window with up to 128K output tokens, and is positioned for complex professional work: advanced coding, scientific and technical reasoning, long-document analysis, computer use, and multistep agent workflows. It sits alongside the cheaper Terra and Luna tiers in the same family: Sol is $5 input and $30 output per million tokens, Terra $2.50 and $15, Luna $1 and $6, and on July 30, 2026 OpenAI cut Luna's price by 80% and Terra's by 20%. Free and Go users of ChatGPT get Terra; paid users choose any of the three and set effort per model.
Why: Sol ranks third overall on the Artificial Analysis Intelligence Index at 58.9, behind Claude Opus 5 and Claude Fable 5, a genuine top-tier frontier model rather than an incremental update, and OpenAI's clear flagship pick for the hardest professional-grade tasks.
New this month
Added Sep 22, 2026
GPT-6 Sol is OpenAI's September 22, 2026 update to GPT-5.6 Sol, priced at $2 per million input tokens and $10 per million output, down from $4/$20, a permanent price cut OpenAI has confirmed is not promotional. It is positioned for complex, repeated developer work: building features, reviewing code, debugging, and analyzing data. On the DeepSWE v1.1 software-engineering benchmark, independent testing found Sol reaches 68.8% at maximum effort for $2.14 per task, against Claude Fable 5's 69.9% at its highest effort setting for $21.63 per task, near-identical accuracy at roughly a tenth of the cost. OpenAI says it makes about half as many mistakes as GPT-5.6 Sol.
Why: The near-parity DeepSWE result against Claude Fable 5 at a tenth of the cost is the reason to look at this model specifically for coding workloads: it is a real price war entry, not a token repricing with nothing behind it, though independent reviewers note overall intelligence scores are flat versus GPT-5.6 rather than improved.
New
Added Oct 1, 2026
GPT-6.1 Sol is OpenAI's September 29, 2026 release at DevDay, one week after GPT-6 Sol, positioned as the price-performance point of the GPT-6 line. OpenAI says it nearly matches GPT-6 Astra on agentic coding, computer use and professional work at one-fifth the standard token prices: $2 per million input and $10 per million output, with cached input cut to $0.10. The company reports the share of responses containing a factual error at low reasoning effort drops from 11.4% to 7.7%, staying within 1.9% of Astra across all reasoning settings, and that the model fails less often at flagging broken search tools, following explicit restrictions and avoiding unauthorized outcomes. It takes text and image input and returns text. It is available to Plus, Pro, Business, Enterprise and Edu users in...
Why: A week after GPT-6 Sol reset the coding price floor, GPT-6.1 Sol claims near-Astra agentic performance at the same $2/$10 rate — if those numbers hold under independent testing, this becomes the default for cost-sensitive agentic work. It enters the leaderboard on OpenAI's own claims, clearly labeled as such, with the scrapped GPT-6.1 Astra noted for context.