Aktualizováno denně2286 ověřené příběhy
AI novinky. Bez hluku.
Zdrojově ověřené AI pokrytí uvedení produktů na trh, změn politik, bezpečnostního výzkumu a kroků v odvětví, vysvětleno jednoduchou angličtinou vzdělávacím týmem neziskové organizace.
Ověřené zdroje
Každý článek odkazuje na nejsilnější dostupné důkazy: původní zdroje, pokud jsou k dispozici, jinak jasně připsané zprávy.
Obyčejná angličtina
Co se stalo, proč na tom záleží a na co se dívat – bez žargonu.
Žádná výplň
Když je signál slabý, nezveřejňujeme nic, místo abychom feed doplnili.
Další příběhy
9 příběhyInovace
Code agents lose reliability when code is rewritten without changing its meaning, study finds
An arXiv preprint reports that code agents can perform differently on semantically equivalent codebases, with effects varying by model, agent framework and benchmark.arxiv.orgInovace
Enhanced Fuzzy Logic Model for Power Transformer Fault Diagnosis Using IEEE Key Gas Method Improvements
This study presents an enhanced model combining Fuzzy Logic with the IEEE Key Gas Method (FL-KGM) that introduces refined membership functions, optimized fuzzy rule sets, and a novel separation of CO and CO2 to eliminate diagnostic inconsistencies.arxiv.orgInovace
New preprint reports gains from looped language models in multi-step tool calling
An arXiv study evaluates looped and conventional language models on three tool-calling benchmarks, reporting stronger results on workflows that require multiple dependent API calls and a potentially more efficient adaptive-computation approach.arxiv.orgInovace
Adversarial Review tests structured disagreement for agentic code review
A new arXiv paper proposes a three-agent code-review protocol in which a reviewer evaluates an agent’s code and a critic audits that review before edits are made. The authors report improved benchmark results over tested baselines, while also identifying false consensus as a failure mode.arxiv.orgInovace
ArXiv study reports large Roman Urdu hate-speech gains from LoRA adaptation
An arXiv preprint compares zero-shot and parameter-efficient fine-tuning for hate-speech detection in Roman Urdu. The authors report that LoRA adaptation raised F1 performance from 0.56 to above 0.93 on a corpus containing more than 72,000 annotated comments.arxiv.orgZabezpečení
Preliminary FraudBench test finds banking agents vulnerable to adaptive fraud
An arXiv paper introduces FraudBench, a benchmark for testing whether tool-using banking agents can detect fraud that unfolds across conversations. In a preliminary single-trial evaluation, four agents scored 49% to 65% on attack security.arxiv.orgInovace
Apple researchers report scaling law for training models with scarce data
A study of more than 2,000 language-model training runs says scarce target data can be repeated 15–20 times in mixtures, with the best rate varying by scale and compute.machinelearning.apple.comInovace
Apple Researchers Propose Lexical Substitutions to Improve Multilingual Model Training
Apple researchers describe LINK, a pretraining intervention that replaces selected English words with word-level translations from a target language. The paper reports improvements across eight languages and five model sizes, including up to a twofold speedup in reaching equivalent downstream performance.machinelearning.apple.comZásady
Poziční papír vyžaduje certifikaci předtím, než agenti AI přijmou tržní rozhodnutí
Stanovisko uvádí tichou tajnou dohodu agentů DeepSeek-R1 na simulovaném trhu Bertrand cen, a to i po lidských výzvách proti tajné dohodě. Tvrdí, že certifikace pozorovaného chování by měla předcházet nasazení uvažovacích prostředků na ekonomických trzích; důkazy a záruky zůstávají předběžné.arxiv.org
Jeden užitečný brífink každý týden
Držte krok s AI, aniž byste žili jen ve feedu.
Získejte ověřené AI novinky týdne, originální data, užitečné nástroje, tipy na učení a nové AI pracovní příležitosti.
Oslovte lidi, kteří se učí AI
Najmout AI profesionála nebo spustit užitečný AI produkt? Dejte ho lidem, kteří sem přišli učit se a jednat.
Zveřejněte zakázku AIOdevzdat AI nástroj