Top storyStudy Finds AI Safety Benchmarks Can Use Far Fewer Tests
A UK AI Security Institute-affiliated preprint reproduced several safety-benchmark results with 97–99% fewer prompts, while warning that shorter tests do not prove real-world safety.
Updated daily35 verified stories
Hand-picked AI coverage on product launches, policy shifts, safety research, and industry moves, explained in plain English by a nonprofit education team.
Every story points back to the original filing, paper, or announcement.
What happened, why it matters, and what to watch — no jargon tax.
When the signal is thin, we publish nothing rather than padding the feed.
A growing stream of verified perspectives for people who need to understand AI without chasing hype.
Top storyA UK AI Security Institute-affiliated preprint reproduced several safety-benchmark results with 97–99% fewer prompts, while warning that shorter tests do not prove real-world safety.

The Department of Commerce has lifted the emergency export controls that forced Anthropic to disable its most capable models worldwide, ending an unprecedented government intervention in a commercial AI release.

Anthropic's new Claude Sonnet 5 approaches the performance of its flagship Opus 4.8 model at a fraction of the price, and becomes the default model for Free and Pro users.

OpenAI has begun a restricted preview of its GPT-5.6 models with a small group of partner organizations, after sharing the models and release plans with the U.S. government.

Noam Shazeer, co-author of the paper that introduced the Transformer architecture and co-lead of Google's Gemini models, is joining OpenAI to lead AI architecture research.

The Chinese lab has released the weights of GLM-5.2 under a permissive MIT license, giving developers everywhere a frontier-class coding model they can download and run themselves.

The U.S. Department of Defense has finalized landmark agreements with major artificial intelligence companies to deploy frontier commercial models on classified military networks.

Meta is reportedly testing Muse Spark, a highly personalized AI assistant designed to perform autonomous tasks across hardware devices and smart wear.

Google unveiled a major update to its search engine at Google I/O 2026, integrating Gemini 3.5 Flash and adding agentic capabilities that can build custom mini-apps or monitor information in real time.

OpenAI has partnered with Canva to introduce direct visual template generation and editing capabilities inside the ChatGPT Plus workspace.

Anthropic has launched Claude Design, a canvas-like workspace designed for editing visual assets and design prototyping, alongside a direct integration with Canva.