Every update
News, leaks & rumors
Everything we're tracking across upcoming AI models, newest first - official announcements alongside the leaks and rumors that precede them. News shapes the watchlist, but only verified public access marks a model released.
One click creates your free account (or signs you into an existing one) and turns on alerts for every major release. There's no password to set, and you can unsubscribe any time.
OpenAI says GPT-5.6 Sol in ChatGPT now gives Plus and Pro users more focused, reliable responses and a thought-level slider, while free users move toward GPT-5.6 Luna.
OpenAI · Aug 6, 2026
Read at openai.com →Read OutYet reportOpenAI outlines its account of GPT-5.6 efficiency across inference and agentic workflows, with a focus on useful work per dollar rather than a single benchmark result.
OpenAI · Jul 29, 2026
Read at openai.com →OpenAI reports that retaining reasoning and enabling compaction in its Responses API harness raised GPT-5.6 Sol's ARC-AGI-3 public-set score from 13.3% to 38.3% while using fewer output tokens.
OpenAI · Jul 29, 2026
Read at openai.com →Read OutYet reportGoogle says the model can plan multi-step tasks, orchestrate lower-level robotics tools, and use continuous video to track progress and adapt to failures.
Google DeepMind · Jul 30, 2026
Read at deepmind.google →OpenAI says GPT-5.6 Sol's API Fast mode replaces Priority Processing, offering up to 2.5 times Standard speed at twice the price while keeping the same model intelligence.
OpenAI · Jul 30, 2026
Read at openai.com →Read OutYet reportAWS published technical guidance for Claude Opus 5 on Amazon Bedrock, focused on agentic systems and production inference workloads. This report does not alter the model's release state.
Amazon Web Services · Jul 24, 2026
Read at aws.amazon.com →AWS published implementation guidance for using GPT-5.6 Sol, Terra, and Luna through Amazon Bedrock’s Responses API, including model selection, prompt caching, Codex integration, quotas, and scaling considerations.
Amazon Web Services · Jul 24, 2026
Read at aws.amazon.com →Read OutYet reportGoogle’s product post introduces Gemini 3.6 Flash alongside 3.5 Flash-Lite and 3.5 Flash Cyber. It describes 3.6 Flash as a workhorse model for coding, knowledge work, and multimodal tasks, with a listed price of $1.50 per million input tokens and $7.50 per million output tokens.
Google · Jul 21, 2026
Read at blog.google →Read OutYet reportGoogle says the cyber-focused model is built on Gemini 3.5 Flash and is being introduced through a limited-access CodeMender pilot.
Google DeepMind · Jul 21, 2026
Read at deepmind.google →Google's API documentation identifies the model code as gemini-3.5-flash-lite and describes its intended high-volume, low-cost workloads.
Google AI for Developers · Jul 21, 2026
Read at ai.google.dev →Google DeepMind announced Gemini 3.6 Flash alongside Gemini 3.5 Flash-Lite and Gemini 3.5 Flash Cyber.
Google DeepMind · Jul 21, 2026
Read at deepmind.google →AWS published implementation guidance for accessing xAI's Grok 4.3 through Amazon Bedrock's Mantle endpoint, including its OpenAI-compatible API path, configurable reasoning, tool calling, and regional constraints.
Amazon · Jul 16, 2026
Read at aws.amazon.com →Read OutYet reportAWS says OpenAI's GPT-5.6 Sol, Terra, and Luna are generally available through Amazon Bedrock, with access through the Responses API and region-specific availability.
AWS · Jul 13, 2026
Read at aws.amazon.com →Read OutYet reportPersonal Intelligence makes the Gemini app feel tailored to you. With your permission, it pulls from Google tools like Gmail, Google Photos, YouTube and Search to provid…
Google DeepMind · Jun 29, 2026
Read at blog.google →Elon Musk called Grok 4.5 'roughly comparable to Opus 4.7, but much faster,' pitching it as a cheaper, more token-efficient alternative at $2/$6 per million tokens. It is the first SpaceXAI model co-trained with Cursor (Anysphere), released July 8 with public availability July 9.
TechCrunch · Jul 8, 2026
Read at techcrunch.com →SpaceXAI launched Grok 4.5, its 'smartest model' for coding, agentic tasks, and knowledge work, trained jointly with Cursor. xAI reports SWE-Bench Pro 64.7%, Terminal-Bench 2.1 83.3%, and #1 on Harvey's Legal Agent Benchmark, served at ~80 TPS with roughly 2x token efficiency, priced at $2/$6 per million input/output tokens. Public access rolls out July 9; EU availability expected mid-July.
xAI · Jul 8, 2026
Read at x.ai →NVIDIA Nemotron 3 Ultra is offering leading performance at lower cost than top closed models with the largest and most widely adopted AI agent orchestration platform. LangChain tuned its Deep Agents harness for NVIDIA Nemotron 3 Ultra, achieving the highest accuracy among open models, while completing more tasks at higher throughput and running at 10x […]
NVIDIA · Jul 8, 2026
Read at blogs.nvidia.com →US lifts restrictions on Anthropic’s most powerful AI models San Juan Daily Star
Anthropic · Jul 8, 2026
Read at news.google.com →China issues backdoor security risk over Anthropic's Claude Code Seeking Alpha
Anthropic · Jul 8, 2026
Read at news.google.com →China warns of security risks in Anthropic’s AI tool, impacting market confidence Crypto Briefing
Anthropic · Jul 8, 2026
Read at news.google.com →China Says It Has Found Security Vulnerabilities in Anthropic’s Claude Code WSJ China warns about AI risks with Anthropic's Claude Code CNBC China issues 'backdoor' security alert over Anthropic's Claude Code Reuters
Anthropic · Jul 8, 2026
Read at news.google.com →Elon Musk confirms Grok 4.5 public release following positive Beta feedback The Eastleigh Voice
xAI · Jul 8, 2026
Read at news.google.com →Elon Musk says Grok 4.5 to launch on July 9, calls it 'Opus-class' The Economic Times
xAI · Jul 8, 2026
Read at news.google.com →New OpenAI Signals data shows how ChatGPT adoption is growing globally, with users increasing usage, exploring more capabilities, and driving growth across regions and languages.
OpenAI · Jun 30, 2026
Read at openai.com →See how Australian Payments Plus uses ChatGPT Enterprise and Codex to move faster through payments complexity. AP+ saves time, improves quality, and keeps human judgment central.
OpenAI · Jul 7, 2026
Read at openai.com →MUFG uses ChatGPT Enterprise to build an AI-native organization, improve workflows, and deliver new AI-powered financial services at scale.
OpenAI · Jul 7, 2026
Read at openai.com →CISA Reportedly Uses Anthropic Mythos to Scan Government Software for Flaws Security Boulevard
Anthropic · Jul 8, 2026
Read at news.google.com →The U.S. Government Is Now Using Anthropic’s AI to Hunt Bugs in Its Own Code 24/7 Wall St.
Anthropic · Jul 8, 2026
Read at news.google.com →CISA Deploys Anthropic’s Mythos AI to Hunt Vulnerabilities in U.S. Government Code Security Affairs
Anthropic · Jul 8, 2026
Read at news.google.com →Mistral Releases Leanstral 1.5, an Open Model That Solved 587 of 672 Putnam Math Problems DevOps.com
Mistral AI · Jul 6, 2026
Read at news.google.com →CISA Reportedly Using Anthropic’s Mythos to Scan Government Software for Flaws SecurityWeek
Anthropic · Jul 7, 2026
Read at news.google.com →US cybersecurity agency using Anthropic’s Mythos to audit govt software: Report Firstpost
Anthropic · Jul 7, 2026
Read at news.google.com →In this post, we present a multi-step pipeline directed by Amazon Nova, which uses its contextual vision reasoning to coordinate complementary tools, including Meta’s open-source Segment Anything Model (SAM 3) deployed on Amazon SageMaker AI for pixel-level segmentation, and Amazon Textract for optical character recognition (OCR). This pipeline is designed to provide comprehensive and compliant PII redaction even for challenging edge cases such as fingerprints, ID cards, or license plates in arbitrary orientations.
Amazon · Jul 6, 2026
Read at aws.amazon.com →In this post, you deploy a two-phase infrastructure for multi-turn RL using Amazon Nova Forge on Amazon SageMaker HyperPod. By the end, you have an event-driven pipeline that starts training when you upload data to Amazon Simple Storage Service (Amazon S3). The training job teaches the model to play Wordle, a placeholder for your own RL task.
Amazon · Jul 6, 2026
Read at aws.amazon.com →In this post, we introduce Reverse Direct Preference Optimization (rDPO), the novel unlearning technique behind Amazon Nova Customizable Content Moderation Settings (CCMS), and show how it reduces over-deflection while preserving model quality. We also provide pointers for customers who want to apply these preference optimization techniques to their own experiments.
Amazon · Jul 6, 2026
Read at aws.amazon.com →Mistral AI Targets Frontier Gap With Open-Weight Model Entering July Early Access Tech Times
Mistral AI · Jul 7, 2026
Read at news.google.com →Mistral Releases Leanstral 1.5 for Math Proof Engineering with Lean WinBuzzer
Mistral AI · Jul 6, 2026
Read at news.google.com →Example spark task "Monitor interior design internships for this summer"
Google DeepMind · Jun 30, 2026
Read at blog.google →EXCLUSIVE: US cyber agency is using Anthropic's Mythos to audit government code, sources say Reuters
Anthropic · Jul 6, 2026
Read at news.google.com →Exclusive-US cyber agency is using Anthropic's Mythos to audit government code, sources say CNA
Anthropic · Jul 7, 2026
Read at news.google.com →Showcasing the importance of open source innovation in American AI, Palantir’s new intelligent engine — introduced today — uses NVIDIA Nemotron open models to serve the needs of U.S. government agencies. Open source software has long been a pillar of U.S. technology leadership. In 1969, DARPA connected four university computers — from UCLA, Stanford, UCSB […]
NVIDIA · Jun 29, 2026
Read at blogs.nvidia.com →Editor’s note: This post is part of the Nemotron Labs blog series, which explores how the latest open models, datasets and training techniques help businesses build specialized AI systems and applications on NVIDIA platforms. Each post highlights practical ways to use an open stack to deliver real value in production — from transparent research copilots […]
NVIDIA · Jun 23, 2026
Read at blogs.nvidia.com →In this post, we show how pairing Amazon Nova 2 Lite with Anthropic’s Claude Sonnet 4.6 delivers an efficient solution for digitizing scanned documents at scale. We built a two-model pipeline on Amazon Bedrock for digitizing scanned yearbook pages. Amazon Nova 2 Lite handles native multimodal extraction in a single call: detecting photos, extracting visible names with coordinates, and returning page-level metadata. Claude Sonnet 4.6 then performs spatial reasoning to match names to faces based on page layout.
Amazon · Jun 29, 2026
Read at aws.amazon.com →In this post, you'll learn how fine-tuning Amazon Nova models using Amazon SageMaker AI addresses these specific issues by teaching the models to recognize your exact data patterns, distinguish between similar fields, and process information more efficiently—achieving up to 94.77% extraction accuracy while reducing costs 50%.
Amazon · Jun 30, 2026
Read at aws.amazon.com →Google Meet's "Take notes for me" feature is available to Google AI Pro and Ultra subscribers in select languages.
Google DeepMind · Jun 29, 2026
Read at blog.google →OpenAI previews GPT-5.6 Sol, a next-generation model with stronger capabilities in coding, science, and cybersecurity, paired with its most advanced safety stack.
OpenAI · Jun 26, 2026
Read at openai.com →Mistral releases Leanstral 1.5 open model for proof engineering TestingCatalog AI News
Mistral AI · Jul 4, 2026
Read at news.google.com →mp4 showing a title card reading "Build with our generative media models"
Google DeepMind · Jun 30, 2026
Read at blog.google →OpenAI and Molecule.one show how a near-autonomous AI chemist using GPT-5.4 improved a key drug-making reaction, advancing medicinal chemistry research.
OpenAI · Jun 17, 2026
Read at openai.com →Learn how GPT-5.5 Instant improves ChatGPT’s health and wellness responses with stronger reasoning, better context, clearer communication, and physician-informed evaluations.
OpenAI · Jun 18, 2026
Read at openai.com →GPT-5 Pro helped solve a 3-year-old immunology mystery, offering insights into T cell behavior. The breakthrough could support cancer and autoimmune research.
OpenAI · Jun 23, 2026
Read at openai.com →Anthropic Says Claude Fable 5 Will Return to Subscriptions Once Capacity Allows Startup Fortune
Anthropic · Jul 4, 2026
Read at news.google.com →Mistral AI Releases Leanstral 1.5: An Apache-2.0 Lean 4 Code Agent Model Solving 587 of 672 PutnamBench Problems MarkTechPost
Mistral AI · Jul 3, 2026
Read at news.google.com →Mistral's open-source Leanstral 1.5 aces formal math benchmarks and catches real bugs in code the-decoder.com
Mistral AI · Jul 4, 2026
Read at news.google.com →Elon Musk's SpaceX showcases Grok-powered smartphone built to reshape AI interaction TweakTown
xAI · Jul 3, 2026
Read at news.google.com →Anthropic Says 'We're Grateful' as Trump Administration Lifts Export Controls on Claude Fable 5 and Mythos 5 After AI Security Standoff Yahoo Finance
Anthropic · Jul 3, 2026
Read at news.google.com →Anthropic Fable 5 'Blipped' Out, Elon & Zuck Cloud 'Plan Bs', & More. AI-RTZ #1137 AI: Reset to Zero
Anthropic · Jul 4, 2026
Read at news.google.com →Anthropic finally brings back Claude Fable 5, but you’ll have to live with a temporary usage limit Yahoo Tech
Anthropic · Jul 4, 2026
Read at news.google.com →Anthropic developer shares prompting tips for Fable 5 that focus on finding your own blind spots first the-decoder.com
Anthropic · Jul 4, 2026
Read at news.google.com →7 Super Dangerous Things to Build With Anthropic’s Fable 5 and Mythos 5 Before They Get Banned Again inc.com
Anthropic · Jul 3, 2026
Read at news.google.com →Anthropic Debuts Claude Sonnet 5 As Agentic AI Push Goes Mainstream Yahoo Tech
Anthropic · Jul 3, 2026
Read at news.google.com →More details on Fable 5’s cyber safeguards and our jailbreak framework Anthropic
Anthropic · Jul 3, 2026
Read at news.google.com →Anthropic’s ‘Fable 5’ Platform Back Online After Export Control Cutoff The National Interest
Anthropic · Jul 3, 2026
Read at news.google.com →Claude Fable 5 isn’t permanently leaving subscriptions, Anthropic says BleepingComputer
Anthropic · Jul 3, 2026
Read at news.google.com →Sonnet 5 narrows the gap: its performance is close to that of Opus 4.8, but at lower prices. It’s a substantial improvement over its predecessor, Sonnet 4.6, on important aspects of agentic performance like reasoning, tool use, coding, and knowledge work.
Anthropic · Jun 30, 2026
Read at anthropic.com →As of June 26 Fable 5 remains suspended (claude-fable-5 API calls error out), but market odds jumped to ~60% for a return 'next week' after Tom Brown replaced Dario Amodei in Commerce Dept talks. Code hints point to a US-first restoration (gov-ID verification goes live July 8), possibly with a weekly usage limit. Important: Fable 5 reappearing in the Azure Foundry catalog and in mobile model-pickers is a stale-UI artifact, not a confirmed relaunch.
explainX · Jun 26, 2026
Read at explainx.ai →Cursor maker Anysphere, recently acquired by SpaceX, says its first fully self-trained model should ship within the next few weeks. It is trained from scratch (no open-source base), is on par with Opus and GPT in size, uses 10-20x the compute of prior Cursor models, and is meant to work beyond coding. The model is not yet named.
The Decoder · Jun 23, 2026
Read at the-decoder.com →With one week left in June, xAI's rumored ~6T-parameter Grok 5 remains in training and unreleased; Musk's Q1 and Q2 targets both slipped. Polymarket's June-30 release contract has fallen to roughly 7–12%. xAI is not yet tracked by OutYet.
Lines.com · Jun 23, 2026
Read at lines.com →Per TestingCatalog, Mistral CEO Arthur Mensch confirmed (June 16) a new model arriving this summer - the start of a 'large yet sparse' mixture-of-experts family - that will ship as open weights, with early access opening in July for research, government, and industry partners. No model name or benchmarks yet.
TestingCatalog · Jun 18, 2026
Read at testingcatalog.com →Z.ai (formerly Zhipu AI) shipped GLM-5.2, a 744B-total / 40B-active open-weight LLM built for long-horizon coding agents, with a 1M-token context and 128K output. Released under MIT on Hugging Face/ModelScope (~June 13-16). Z.ai claims it trails Claude Opus 4.8 by ~1pt on FrontierSWE and beats GPT-5.5 and Opus 4.7 on several long-horizon benchmarks.
TestingCatalog · Jun 18, 2026
Read at testingcatalog.com →As of ~June 19, Gemini 3.5 Pro remains in limited preview for select Vertex AI enterprise customers and has not reached the consumer Gemini app or AI Studio. GA is expected in the final two weeks of June (per Pichai's I/O 'next month' commitment); Polymarket gives ~50-55% odds of a pre-June-30 release. Targets a 2M-token context and Deep Think reasoning.
Codersera · Jun 19, 2026
Read at codersera.com →Fable 5 (and Mythos 5) have been offline since a June 12 US export-control directive forced a global suspension. June 20-21 updates report the directive standing but softening, Fable 5 briefly resurfacing in an Android app, and NSA breach testimony reshaping the ban. Availability remains uncertain.
Tech Times · Jun 21, 2026
Read at techtimes.com →The letter is signed by both democratic and republican congressmen.
Anthropic · Jun 22, 2026
Read at reddit.com →Anthropic released Claude Fable 5, then a U.S. government export directive forced it offline within three days.
InfoQ · Jun 15, 2026
Read at infoq.com →Anthropic introduced Claude Fable 5, a Mythos-class model made safe for general use, alongside the restricted Claude Mythos 5.
Anthropic · Jun 9, 2026
Read at anthropic.com →TechCrunch: Claude Fable 5 is a publicly accessible version of Anthropic's Mythos model, shipping with guardrails on high-risk domains like cybersecurity and biology.
TechCrunch · Jun 9, 2026
Read at techcrunch.com →At Google I/O, Google said Gemini 3.5 Pro is already in internal use and would roll out "next month," targeting a 2M-token context window and Deep Think reasoning.
Google · May 19, 2026
Read at blog.google →