Updated daily1982 verified stories
AI News. Without the noise.
Source-checked AI coverage of product launches, policy shifts, safety research, and industry moves, explained in plain English by a nonprofit education team.
Verified sourcing
Every story links to the strongest available evidence: original sources when available, otherwise clearly attributed reporting.
Plain English
What happened, why it matters, and what to watch — without the jargon.
No filler
When the signal is thin, we publish nothing rather than padding the feed.
More stories
9 storiesInnovation
Replication Study Says FLOPs Still Mispredict AI Runtime, and the Proposed Fix Fails on Newer Hardware
A preprint by two researchers reproduces an earlier study on why equal FLOP counts do not mean equal execution time. It confirms the underlying claim but reports that the α-FLOPs correction formula generally underestimates runtime on newer hardware, which shows jumps and oscillations the formula does not capture.arxiv.orgEnterprise
Benchmark Paper Finds Four Ways to Query Enterprise Data With LLMs All Score Under 26%
A new arXiv preprint pits four architectures for natural-language querying of enterprise databases against each other on a synthetic bilingual benchmark. None answered more than about a quarter of cases correctly, and the design that scored highest was not the safest or the cheapest.arxiv.orgInnovation
Paper Proposes Grading AI Security Agents Without Labels by Measuring Convergence to a Stronger Model
A new arXiv preprint argues security teams can judge whether a memory- or retrieval-equipped AI agent is learning by measuring how far it closes the gap to a stronger "teacher" model, rather than on labeled benchmarks that are often scarce or stale. Judging by a similarly powered model gave no usable signal.arxiv.orgInnovation
New Benchmark Tests Whether AI Assistants Can Remember a Year of Phone Use
A 17-author technical report posted to arXiv introduces MobileMem, a benchmark and framework for on-device long-term memory built from a year-scale collection of mobile experiences. The abstract describes the design but reports no scores, and key details about the underlying data remain undisclosed.arxiv.orgInnovation
Paper Reports Brain-Like Modular Organization Emerging Inside Large Language Models
A new arXiv preprint says large language models develop functionally specialized internal structure that lines up with distinct human brain networks, based on circuit analyses across 46 tasks in four cognitive domains. The abstract page leaves key methodological details unstated.arxiv.orgInnovation
Paper Finds Late Layers of a Mixture-of-Experts Model Tolerate Heavy Expert Masking
A preprint reports that disabling low-magnitude experts in the last five layers of a 35-billion-parameter Mixture-of-Experts model preserved far more usable code-translation outputs than spreading the same cuts across all layers. It covers one model and one benchmark, and the abstract reports no unmasked baseline.arxiv.orgSecurity
CoreBreak Flaws Let Agent Tools Run Without the Model Ever Being Called
A Cloud Security Alliance research note describes CoreBreak, a pattern of flaws in Amazon Bedrock AgentCore, Google's Agent Development Kit, and Vercel's AI SDK harness packages that allowed tools to execute without a model turn — leaving model-level guardrails with nothing to inspect.labs.cloudsecurityalliance.orgPolicy
Anthropic Details How Claude's Text Watermark Will Work
Anthropic says future Claude models will embed a statistical watermark based on Google DeepMind's SynthID-Text, to comply with the EU AI Act. The company says it adds no characters, tokens, or user identity — and that a full rewrite defeats it.
anthropic.comIndustry
Cursor Says Its Acquisition by SpaceX Has Officially Closed
Cursor published a short post saying SpaceX has completed its acquisition of the AI coding tool, finishing a process it says began in April with a model-training partnership with SpaceXAI. The post promises access to what it calls the world's largest GPU fleet, but discloses no terms, timelines, or product changes.
cursor.com
One useful briefing each week
Keep up with AI without living in the feed.
Get the week’s verified AI news, original data, useful tools, learning picks, and fresh AI jobs.
Reach people who are learning AI
Hiring an AI professional or launching a useful AI product? Put it in front of people who came here to learn and act.
Post an AI jobSubmit an AI tool