Every update

News, leaks & rumors

Everything we're tracking across upcoming AI models, newest first - official announcements alongside the leaks and rumors that precede them. News shapes the watchlist, but only verified public access marks a model released.

One click creates your free account (or signs you into an existing one) and turns on alerts for every major release. There's no password to set, and you can unsubscribe any time.

Introducing Grok 4.5

SpaceXAI launched Grok 4.5, its 'smartest model' for coding, agentic tasks, and knowledge work, trained jointly with Cursor. xAI reports SWE-Bench Pro 64.7%, Terminal-Bench 2.1 83.3%, and #1 on Harvey's Legal Agent Benchmark, served at ~80 TPS with roughly 2x token efficiency, priced at $2/$6 per million input/output tokens. Public access rolls out July 9; EU availability expected mid-July.

xAI · Jul 8, 2026

Read at x.ai

NVIDIA Nemotron Achieves Benchmark-Leading Performance With LangChain Deep Agents Harness

NVIDIA Nemotron 3 Ultra is offering leading performance at lower cost than top closed models with the largest and most widely adopted AI agent orchestration platform. LangChain tuned its Deep Agents harness for NVIDIA Nemotron 3 Ultra, achieving the highest accuracy among open models, while completing more tasks at higher throughput and running at 10x […]

NVIDIA · Jul 8, 2026

Read at blogs.nvidia.com

Automatically redact PII in images with Amazon Nova

In this post, we present a multi-step pipeline directed by Amazon Nova, which uses its contextual vision reasoning to coordinate complementary tools, including Meta’s open-source Segment Anything Model (SAM 3) deployed on Amazon SageMaker AI for pixel-level segmentation, and Amazon Textract for optical character recognition (OCR). This pipeline is designed to provide comprehensive and compliant PII redaction even for challenging edge cases such as fingerprints, ID cards, or license plates in arbitrary orientations.

Amazon · Jul 6, 2026

Read at aws.amazon.com

Open Models, Closed Environments: Palantir Brings Secure AI to US Agencies With NVIDIA Nemotron

Showcasing the importance of open source innovation in American AI, Palantir’s new intelligent engine — introduced today — uses NVIDIA Nemotron open models to serve the needs of U.S. government agencies. Open source software has long been a pillar of U.S. technology leadership. In 1969, DARPA connected four university computers — from UCLA, Stanford, UCSB […]

NVIDIA · Jun 29, 2026

Read at blogs.nvidia.com

How Businesses Are Building Specialized AI They Can Trust

Editor’s note: This post is part of the Nemotron Labs blog series, which explores how the latest open models, datasets and training techniques help businesses build specialized AI systems and applications on NVIDIA platforms. Each post highlights practical ways to use an open stack to deliver real value in production — from transparent research copilots […]

NVIDIA · Jun 23, 2026

Read at blogs.nvidia.com

Pair Nova 2 Lite with Claude for cost-optimized document processing

In this post, we show how pairing Amazon Nova 2 Lite with Anthropic’s Claude Sonnet 4.6 delivers an efficient solution for digitizing scanned documents at scale. We built a two-model pipeline on Amazon Bedrock for digitizing scanned yearbook pages. Amazon Nova 2 Lite handles native multimodal extraction in a single call: detecting photos, extracting visible names with coordinates, and returning page-level metadata. Claude Sonnet 4.6 then performs spatial reasoning to match names to faces based on page layout.

Amazon · Jun 29, 2026

Read at aws.amazon.com

Fable 5 still offline (Day 14) but restoration odds surge to ~60%; Azure/app 'sightings' are stale UI, not a relaunch

As of June 26 Fable 5 remains suspended (claude-fable-5 API calls error out), but market odds jumped to ~60% for a return 'next week' after Tom Brown replaced Dario Amodei in Commerce Dept talks. Code hints point to a US-first restoration (gov-ID verification goes live July 8), possibly with a weekly usage limit. Important: Fable 5 reappearing in the Azure Foundry catalog and in mobile model-pickers is a stale-UI artifact, not a confirmed relaunch.

explainX · Jun 26, 2026

Read at explainx.ai

Z.ai releases GLM-5.2, an open-weight 744B-param coding model with a 1M-token context

Z.ai (formerly Zhipu AI) shipped GLM-5.2, a 744B-total / 40B-active open-weight LLM built for long-horizon coding agents, with a 1M-token context and 128K output. Released under MIT on Hugging Face/ModelScope (~June 13-16). Z.ai claims it trails Claude Opus 4.8 by ~1pt on FrontierSWE and beats GPT-5.5 and Opus 4.7 on several long-horizon benchmarks.

TestingCatalog · Jun 18, 2026

Read at testingcatalog.com

Gemini 3.5 Pro still in limited Vertex preview; GA expected before end of June

As of ~June 19, Gemini 3.5 Pro remains in limited preview for select Vertex AI enterprise customers and has not reached the consumer Gemini app or AI Studio. GA is expected in the final two weeks of June (per Pichai's I/O 'next month' commitment); Polymarket gives ~50-55% odds of a pre-June-30 release. Targets a 2M-token context and Deep Think reasoning.

Codersera · Jun 19, 2026

Read at codersera.com