Apple researchers report scaling law for training models with scarce data
A study of more than 2,000 language-model training runs says scarce target data can be repeated 15–20 times in mixtures, with the best rate varying by scale and compute.
Оновлюється щодня1797 перевірені історії
Перевірений джерелом штучний інтелект висвітлює випуск продуктів, зміни в політиці, дослідження безпеки та зміни в галузі, пояснені простою англійською мовою некомерційною освітньою командою.
Кожна історія пов’язана з найвагомішими наявними доказами: оригінальними джерелами, якщо вони доступні, в іншому випадку чітко зазначеними повідомленнями.
Що сталося, чому це важливо та що дивитися — без жаргону.
Коли сигнал слабкий, ми нічого не публікуємо, а лише доповнюємо канал.
Історії про штучний інтелект, перевірені джерелами, спочатку найновіші, для людей, яким потрібно розуміти штучний інтелект, не ганяючись за рекламою.
A study of more than 2,000 language-model training runs says scarce target data can be repeated 15–20 times in mixtures, with the best rate varying by scale and compute.
Apple researchers describe LINK, a pretraining intervention that replaces selected English words with word-level translations from a target language. The paper reports improvements across eight languages and five model sizes, including up to a twofold speedup in reaching equivalent downstream performance.
A position paper reports tacit collusion by DeepSeek-R1 agents in a simulated Bertrand pricing market, even after human prompts against collusion. It argues that observed-behavior certification should precede deployment of reasoning agents in economic markets; the evidence and safeguards remain preliminary.
A systematic review surveys how large language models are being studied for mental-health analysis, risk assessment, therapy support and multimodal monitoring, while stressing unresolved ethical and regulatory challenges.
An ICML 2026 position paper analyzing 500 Hugging Face model cards argues that open-weight foundation models need coordinated model cards, acceptable-use policies, and licenses to address safety and governance gaps.
A position paper proposes the Transition Complexity Profile, a standardized way to describe how unpredictable and long-range the dynamics of game environments are for game-world modeling and reinforcement learning.
A new arXiv benchmark places 15 language-model agents in a 20-year football-management simulation, testing whether they can make consistent decisions when short-term choices affect long-term outcomes.
A new arXiv benchmark reports that changing only the retrieval method raised a fixed model’s exact accuracy on financial reconciliation cases from 2.05% to 72.44%. The study also finds that a correct root-cause label often does not mean the system returned sufficient evidence for an auditable diagnosis.
A new arXiv study presents scaling-law experiments for text-to-image diffusion models across compute budgets from 10^19 to 10^22 FLOPs. Its authors report that image models need substantially more data per parameter than language models to train efficiently.
A new arXiv paper describes a self-supervised EEG method that reports 92.4% accuracy distinguishing Alzheimer’s disease from cognitively normal controls on the ADFTD cohort.
An arXiv paper introduces Agentic ESOpt, a proposed evolution-strategy framework for fine-tuning long-horizon language-model agents with inference-level GPU memory. The authors report gains for Qwen-3.5-27B on WebArena-Lite and improvements in 28 of 36 prompt-optimization settings.
An arXiv paper argues that mean squared error can misjudge irregular time-series forecasts and proposes a continuous-time metric tested across synthetic, semi-synthetic and eight real-world datasets.
Один корисний брифінг щотижня
Отримуйте перевірені новини про штучний інтелект, оригінальні дані, корисні інструменти, підбірки для навчання та нові вакансії зі штучного інтелекту.
Найняти професіонала ШІ чи запустити корисний продукт ШІ? Покажіть це людям, які прийшли сюди вчитися та діяти.
Опублікувати роботу ШІ Надішліть інструмент ШІ