Jul 24, 2026

6 Posts

Diagram showing research program, first article collection, then question generation, followed by model evaluation and analysis
Jul 24, 2026

Web Retrieval Flusters LLMs: AI agents searching online can struggle to retrieve correct info, researchers find

Large language models often are called upon to gather news. In this task, researchers found, their ability to find relevant reports is the weakest link.
Cloudflare diagram on "why the web is being crawled" showing 51.8 percent model training, 36.6 percent mixed use, and 8.6 percent search
Jul 24, 2026

Cloudflare’s Web Crawler Flare-Up: Cloudflare moves to block all AI training bots by default

Web publishers using Cloudflare will soon be able to separately control an AI bot’s access based on what it does, allowing search indexing while blocking AI training or agent activity.
Animated GIF of Meta's Muse Spark 1.1's performance on benchmarks, showing strong performance on tool use and roughly second-tier overall intelligence
Jul 24, 2026

Meta Sparks A Price War: Muse Spark 1.1 makes a jump in intelligence at a lower price than peers

With Llama, Meta marked itself as an open alternative to OpenAI. With its new closed models, Meta now positions itself as a low-cost, high-value competitor.
Kimi diagram of open frontier model size over time, with Moonshot AI's Kimi K3 at 2.8 trillion parameters, ahead of DeepSeek at 1.6T parameters
Jul 24, 2026

Kimi K3 Reveals How A Giant Frontier AI Model Works: Moonshot's latest model outshines all but GPT-5.6 Sol and Claude 5 Fable

Moonshot’s latest model leapfrogged the month-old GLM-5.2 and a host of proprietary competitors to finish just behind GPT-5.6 Sol and Claude Fable 5 on many benchmarks.
A rally for open weight / open source AI with leaders bearing a banner that reads "open source, open weights, open future"
Jul 24, 2026

When Guardrails Go Wrong: After a closed model went amok on a key vendor's system, an open weight model helped save the day

A few frontier labs have tried to tell a story of open models being dangerous because they can be used to launch cyberattacks, and of their “safe” proprietary models with strong guardrails being there to defend us.
A rally for open weight / open source AI with leaders bearing a banner that reads "open source, open weights, open future"
Jul 24, 2026

Kimi K3 Redraws the Open Frontier, Muse Spark 1.1 Undercuts Competitors, Cloudflare Moves to Cut Off Crawlers

The Batch News & Insights: A few frontier labs have tried to tell a story of open models being dangerous because they can be used to launch cyberattacks, and of their “safe” proprietary models with strong guardrails being there to defend us.

Subscribe to The Batch

Stay updated with weekly AI News and Insights delivered to your inbox