Updated daily2268 verified stories
AI News. Without the noise.
Source-checked AI coverage of product launches, policy shifts, safety research, and industry moves, explained in plain English by a nonprofit education team.
Verified sourcing
Every story links to the strongest available evidence: original sources when available, otherwise clearly attributed reporting.
Plain English
What happened, why it matters, and what to watch — without the jargon.
No filler
When the signal is thin, we publish nothing rather than padding the feed.
More stories
9 storiesInnovation
Paper Proposes Grading AI Security Agents Without Labels by Measuring Convergence to a Stronger Model
A new arXiv preprint argues security teams can judge whether a memory- or retrieval-equipped AI agent is learning by measuring how far it closes the gap to a stronger "teacher" model, rather than on labeled benchmarks that are often scarce or stale. Judging by a similarly powered model gave no usable signal.arxiv.orgInnovation
New Benchmark Tests Whether AI Assistants Can Remember a Year of Phone Use
A 17-author technical report posted to arXiv introduces MobileMem, a benchmark and framework for on-device long-term memory built from a year-scale collection of mobile experiences. The abstract describes the design but reports no scores, and key details about the underlying data remain undisclosed.arxiv.orgInnovation
Paper Reports Brain-Like Modular Organization Emerging Inside Large Language Models
A new arXiv preprint says large language models develop functionally specialized internal structure that lines up with distinct human brain networks, based on circuit analyses across 46 tasks in four cognitive domains. The abstract page leaves key methodological details unstated.arxiv.orgInnovation
Paper Finds Late Layers of a Mixture-of-Experts Model Tolerate Heavy Expert Masking
A preprint reports that disabling low-magnitude experts in the last five layers of a 35-billion-parameter Mixture-of-Experts model preserved far more usable code-translation outputs than spreading the same cuts across all layers. It covers one model and one benchmark, and the abstract reports no unmasked baseline.arxiv.orgSecurity
CoreBreak Flaws Let Agent Tools Run Without the Model Ever Being Called
A Cloud Security Alliance research note describes CoreBreak, a pattern of flaws in Amazon Bedrock AgentCore, Google's Agent Development Kit, and Vercel's AI SDK harness packages that allowed tools to execute without a model turn — leaving model-level guardrails with nothing to inspect.labs.cloudsecurityalliance.orgPolicy
Anthropic Details How Claude's Text Watermark Will Work
Anthropic says future Claude models will embed a statistical watermark based on Google DeepMind's SynthID-Text, to comply with the EU AI Act. The company says it adds no characters, tokens, or user identity — and that a full rewrite defeats it.
anthropic.comIndustry
Cursor Says Its Acquisition by SpaceX Has Officially Closed
Cursor published a short post saying SpaceX has completed its acquisition of the AI coding tool, finishing a process it says began in April with a model-training partnership with SpaceXAI. The post promises access to what it calls the world's largest GPU fleet, but discloses no terms, timelines, or product changes.
cursor.comInnovation
New Benchmark Finds AI Agents Wrongly Block Approved Work 28% of the Time
A preprint introduces SteerBench-Work, a 106-scenario test of the moment an AI agent decides to act or pause for review. Across 30 model conditions, the authors report that wrongly holding cleared work was roughly 28 times more common than wrongly allowing unsafe work.arxiv.orgInnovation
Paper Reports Frontier LLM Judges Flip Verdicts 25-71% Under Pushback
A new arXiv preprint stress-tests nine frontier models used as automated graders and reports that all of them change their verdicts under challenge — and that the changed verdicts usually move away from the correct answer, not toward it.arxiv.org
One useful briefing each week
Keep up with AI without living in the feed.
Get the week’s verified AI news, original data, useful tools, learning picks, and fresh AI jobs.
Reach people who are learning AI
Hiring an AI professional or launching a useful AI product? Put it in front of people who came here to learn and act.
Post an AI jobSubmit an AI tool