Study finds AI preference measurements depend heavily on the testing instrument
A new preprint reports that conclusions about what AI models prefer may change substantially with the prompt format used to measure them.
Actualizado diariamente1507 historias verificadas
Cobertura verificada por IA sobre lanzamientos de productos, cambios en políticas, investigaciones de seguridad y movimientos del sector, explicada en un lenguaje sencillo por un equipo educativo sin ánimo de lucro.
Cada historia enlaza con la evidencia más sólida disponible: fuentes originales cuando están disponibles, y por lo demás reportajes claramente atribuidos.
Qué sucedió, por qué es importante y qué observar: sin impuestos de jerga.
Cuando la señal es escasa, no publicamos nada en lugar de rellenar el feed.
Un flujo creciente de perspectivas verificadas para personas que necesitan comprender la IA sin perseguir exageraciones.
A new preprint reports that conclusions about what AI models prefer may change substantially with the prompt format used to measure them.
A new preprint proposes FLARE, a framework for estimating whether healthcare AI deployments are economically viable after accounting for patient volume, verification time, infrastructure and workflow design.
BBC News Chinese reports that AI-generated videos featuring fictional doctors and false health or welfare claims are circulating widely among older people in Taiwan. A Taipei class is teaching seniors to question videos shared through YouTube and messaging groups, while the scale, ownership and possible political…
An arXiv study describes an inference-time method that selectively steers language models when they appear likely to hallucinate or change a correct medical answer under user pressure. In 600 pressure trajectories involving a 4-billion-parameter model, the authors say the unsteered model abandoned its answer 570…
An arXiv study reports that an LLM trained on synthetic limit order book data can generate valid event sequences while failing to learn the book’s underlying state, producing biased estimates and spurious predictability in forecasts.
The Straits Times reports that China is expanding AI data centres in Guizhou, using the province’s cooler climate, renewable energy and available land while seeking to develop the rural west. Residents describe better transport and nearby jobs, but researchers cited by the report question whether data centres…
NDTV reports that Harvard Business School has launched an eight-week, $699 startup bootcamp using AI mentorship modeled on faculty expertise, with live expert sessions and a potential investor pitch.
A paper introduces SyPS, a framework for testing whether changes in confidence, emotion, social consensus and validation-seeking language make large language models more likely to agree with users.
A new arXiv paper describes BioCheck Agent, an AI system that searches PubMed and produces evidence-backed biomedical fact-checking reports rather than isolated true-or-false labels. The authors report improved benchmark accuracy and fewer hallucinated citations compared with a Qwen3.5-4B base model, but real-world…
The Guardian reports that China has introduced national restrictions on AI companion services, including a ban on companions for minors, after concerns that chatbots could foster emotional dependence and displace human relationships.
A new arXiv paper compares tensor parallelism with KV-cache compression for memory-bound large language model serving, reporting that compression was 1.20 to 2.00 times cheaper across its tested configurations, while additional GPUs were the only tested approach that reduced latency.
A revised arXiv paper proposes auditing AI explanations for stability and faithfulness, reporting that models with AUC above 0.99 can still generate flat or uninformative explanations.
Un informe útil cada semana
Recibe las noticias verificadas de IA de la semana, datos originales, herramientas útiles, recomendaciones de aprendizaje y nuevos empleos en IA.
¿Contratar a un profesional de IA o lanzar un producto útil de IA? Ponlo delante de personas que han venido aquí a aprender y actuar.
Publicar un trabajo de IA Enviar una herramienta de IA