Оновлюється щодня2521 перевірені історії
Новини AI. Без шуму.
Перевірений джерелом штучний інтелект висвітлює випуск продуктів, зміни в політиці, дослідження безпеки та зміни в галузі, пояснені простою англійською мовою некомерційною освітньою командою.
Перевірене джерело
Кожна історія пов’язана з найвагомішими наявними доказами: оригінальними джерелами, якщо вони доступні, в іншому випадку чітко зазначеними повідомленнями.
Звичайна англійська
Що сталося, чому це важливо та що дивитися — без жаргону.
Без наповнювача
Коли сигнал слабкий, ми нічого не публікуємо, а лише доповнюємо канал.
Більше історій
9 історіїІнновація
FlashPrefill V2 paper reports large long-context serving speedups with block-sparse attention
An arXiv paper describes FlashPrefill V2, a block-sparse attention system for long-context LLM serving, and reports speedups of up to 47.26x over FlashAttention-2 at 128K context on NVIDIA H20 GPUs.arxiv.orgІнновація
SWE-bench Science benchmark finds coding agents struggle with scientific software
A new arXiv preprint introduces a 119-task benchmark for testing coding agents on scientific software and reports that the best-performing agent scored below 50% on pass@1.arxiv.orgІнновація
У документі ArXiv пропонується навчання на основі етапів для довгострокових агентів LLM
MileGPO використовує виявлення віх і локальні докази для покращення розподілу кредитів під час навчання агентів мовної моделі виконання тривалих багатоетапних завдань. Автори повідомляють про найсучасніші результати на ALFWorld і WebShop, але претензії залишаються обмеженими експериментами газети.arxiv.orgІнновація
Preprint tests whether LLM agents know when to remember, verify or ask
A new benchmark evaluates whether language-model agents correctly decide when interaction-derived information should be saved, checked, used temporarily or clarified with a user.arxiv.orgІнновація
Аудит міжмовної справедливості водяних знаків мовної моделі
Пропонується нова структура оцінювання схем водяних знаків у великих мовних моделях, зосереджена на міжмовній справедливості.arxiv.orgполітика
Stanford AI Index finds AI policy expanding as sovereignty and investment diverge
Stanford HAI’s 2026 AI Index says national AI strategies are spreading, while data-localization rules, state-backed computing capacity and public investment remain uneven across regions.hai.stanford.eduІнновація
Together AI benchmark: GLM-5.3 trails GPT-5.6 Sol on the first try, wins on retries at half the price
A Together AI analysis of 904 DeepSWE rollouts reports that OpenAI's GPT-5.6 Sol leads on first-attempt coding accuracy while the open-weight GLM-5.3 leads once retries are allowed, at roughly half the cost per attempt.together.aiІнновація
Preprint proposes a locally tokenized AI model for robust time-series watermarking
An arXiv preprint proposes an AI generative model and watermarking method for more reliable multivariate time-series data after editing. Authors report tests across finance, energy and neuroimaging benchmarks, but the abstract gives no numerical results or evidence of deployment.arxiv.orgІнновація
Natural Language Code Retrieval for 1C:Enterprise: An Open Benchmark and Efficient Bi-Encoder
Natural language code retrieval is a rapidly evolving task in computer science. However, the 1C:Enterprise ecosystem combines Russian syntax with highly domain-specific terminology, for which open datasets and specialized models have been virtually non-existent.arxiv.org
Один корисний брифінг щотижня
Будьте в курсі ШІ, не живучи в стрічці.
Отримуйте перевірені новини про штучний інтелект, оригінальні дані, корисні інструменти, підбірки для навчання та нові вакансії зі штучного інтелекту.
Охопіть людей, які вивчають ШІ
Найняти професіонала ШІ чи запустити корисний продукт ШІ? Покажіть це людям, які прийшли сюди вчитися та діяти.
Опублікувати роботу ШІНадішліть інструмент ШІ