AI News.
AI is the fastest-moving beat in tech, and the one most buried in hype. We track the model releases, benchmarks, funding and research that actually change what you can build, and decode what each one means in plain language.
AI
GPT-5.6 Goes Public Thursday as OpenAI Opens the Gate
OpenAI said this morning that GPT-5.6 Sol, Terra and Luna launch publicly this Thursday, with global preview access opening now, ending a two-week government-gated lockout for the top-scoring coding model.
AI
Meta Launches Muse, an AI Image Model Aimed at Firefly
Meta launched Muse on July 7, a native image-generation and editing model inside Meta AI that restyles your own photos, escalating the image war with Adobe Firefly and Google Imagen.
AI
OpenAI Ships gpt-realtime-2.1 With Lower Voice Latency
OpenAI released gpt-realtime-2.1 and a cheaper mini variant for building low-latency voice agents, cutting p95 Realtime latency by at least 25% through improved caching.
AI
Cohere Ships North Mini Code, a 30B Coder for One H100
Cohere released North Mini Code, an open-weight 30B mixture-of-experts coding model that runs on a single H100 GPU with a 256K context and is free to use, aimed at enterprises that need agentic coding without sending source to a cloud API.
AI
Google Ships Gemini 3 Pro Image and 3.1 Flash Image
Google launched two new image models, Gemini 3 Pro Image and Gemini 3.1 Flash Image, on June 30, 2026, splitting its lineup into a cheap high-volume tier and a premium tier while Gemini 3.5 Pro keeps slipping.
AI
Zhipu's GLM-5.2 Tops the Open-Weight Model Rankings
Beijing lab Zhipu AI released GLM-5.2, a 753-billion-parameter open-weight model under the permissive MIT license that now ranks first among open models and fourth overall, meaning the best model you can download and self-host is Chinese.
AI
Anthropic Overtakes OpenAI on Revenue Run-Rate
Anthropic now reports a roughly $47B annualized revenue run-rate, pulling ahead of OpenAI's reported $25-33B, driven mainly by enterprise and coding subscriptions rather than consumer chat.
AI
Claude Sonnet 5: Near-Opus Coding at Half the Price
Anthropic's Claude Sonnet 5, now the default for every free and Pro user, scores 85.2% on SWE-bench Verified while costing $2 per million input tokens. That puts near-flagship coding within a rounding error of Opus 4.8 at roughly a quarter of the price, and makes it the new value benchmark for the whole field.
AI
California Gives Every State Agency Claude at Half Price
Governor Gavin Newsom signed a first-of-its-kind deal giving every California state agency, city and county access to Anthropic's Claude at a 50% discount, days after the federal government designated Anthropic a supply-chain risk, exposing a growing state-versus-federal split on frontier AI.
AI
375ai wants to be the data layer for physical AI
375ai is building a real-world data network for physical AI: fixed 375edge sensor nodes mounted on billboards plus a 375go phone app that pays a crowd to scan the world, feeding labeled multimodal data to autonomous-vehicle, robotics and world-model teams. The company says it has logged 2.7 billion events across a planned 40,000 US locations.
AI
OpenAI Ships GPT-5.6 but the Government Locks the Door
OpenAI previewed GPT-5.6 Sol, Terra, and Luna on June 26, then restricted them to about 20 trusted partners at the US government's request, the first time Washington has gated an American AI model before a public launch.
AI
Anthropic's biology fix: one tool beats a bigger AI
Anthropic found that AI agents fail at basic biology data retrieval not because the models are weak but because scientific databases are a mess. Bolting on a deterministic tool called gget virus lifted Claude Sonnet 4 from 16.9% to 92.8% accuracy, and every model it tested cleared 92%.
AI
Meituan's LongCat-2.0 Is a 1.6T Coder on Chinese Chips
Meituan open-sourced LongCat-2.0 on June 30, 2026: a 1.6-trillion-parameter Mixture-of-Experts agentic coding model, MIT-licensed, trained entirely on a 50,000-card domestic Chinese chip cluster with no Nvidia or AMD hardware.
AI
Claude Fable 5 Returns and Retakes the Coding Crown
Anthropic restored global access to Claude Fable 5 on July 1, 2026, 20 days after a US export-control order pulled it offline, and the Mythos-class flagship immediately retook the SWE-Bench Pro coding lead at 80.3%, the highest score of any generally usable model.
AI
OpenAI's GeneBench-Pro Exposes AI's Genomics Judgment Gap
OpenAI's GeneBench-Pro, released June 30, 2026, is a 129-problem benchmark that tests whether AI agents can make real analytical judgments over messy biology data. Its top model, GPT-5.6 Sol Pro, solved just 31.5%, exposing a gap in what OpenAI calls research taste.
Gemini 3.5 Flash Makes Computer Use a Native Tool
On June 24, 2026, Google made computer use a built-in tool inside Gemini 3.5 Flash, so the same cheap, fast production model that does search grounding can now read a screen and click through a UI without a separate agent model.
Microsoft's MAI Models Signal It Wants to Need OpenAI Less
Microsoft unveiled its own MAI-Code-1-Flash and MAI-Thinking-1 models, its clearest move yet to build in-house AI, cut developer costs, and lean less on OpenAI.
SubQ Claims the First Subquadratic Frontier LLM
Miami startup Subquadratic launched SubQ 1M-Preview, the first frontier LLM built on a fully subquadratic attention design, letting it scale to a 12-million-token context instead of paying the transformer's quadratic tax.
Anthropic Filed Confidentially for an IPO, Ahead of OpenAI
Anthropic confidentially filed a draft S-1 with the SEC on June 1, 2026, days after a $65 billion round valued it at $965 billion. It is the first frontier AI lab to formally start the path to a public listing.
Google Ships Two New Gemini Image Models, and Pricing Is the Story
Google released two new image models on June 18, 2026: Gemini 3 Pro Image and the cheaper Gemini 3.1 Flash Image. The split signals AI image generation is now a tiered, price-competitive market, not a single-flagship race.
OpenAI Previews GPT-5.6 With Sol, Terra, and Luna, and the Real Story Is the Tiering
OpenAI opened a limited preview of GPT-5.6 with three models named Sol, Terra, and Luna. The flagship gets the headlines, but the cheaper tiers tell you where this race is actually going.
The AI Talent War Just Hit a New High With Noam Shazeer's Jump to OpenAI
Noam Shazeer, a co-author of the paper that invented the transformer, is leaving Google for OpenAI to lead architecture research. A single hire this big tells you where the real bottleneck in AI actually sits.
AI
RAG Explained: Why Retrieval Beats a Bigger Model
Retrieval-augmented generation is the unglamorous technique quietly powering most useful AI products. It's also the cheapest way to make a model 'know' your data.
AI
The Real Cost of Running a Large Language Model
Training headlines grab attention, but the bill that never stops arriving is inference, the cost of actually answering each question, forever.
AI
Open Weights vs Open Source: The AI License Fight
When a company says its AI model is 'open,' it's worth asking open in what sense. The word is doing a lot of quiet work.
Why AI Agents Are Harder Than the Demos Suggest
A polished demo of an AI agent booking a trip looks like the future. Shipping one that works reliably for real users is a different sport entirely.
AI
The Quiet Rise of Small Language Models
The race isn't only about who has the biggest model anymore. Increasingly, the interesting work is about how small you can go without losing the magic.
Tokens, Not Words: How an AI Actually Reads Your Prompt
When you type a sentence to an AI, it doesn't see words the way you do. It sees tokens, and that small fact explains a lot of the model's quirks.
Why AI Models Hallucinate, and What Actually Helps
The most frustrating thing about a language model is its habit of stating false things with total confidence. The cause is baked into how these systems work.
The Context Window Is the New RAM
Every conversation with an AI has a memory limit. Understanding that limit, the context window, explains why your long chats start to drift.