Diperbarui setiap hari2411 cerita terverifikasi
Berita AI. Tanpa kebisingan.
Liputan AI yang diperiksa sumber tentang peluncuran produk, perubahan kebijakan, riset keselamatan, dan langkah industri, dijelaskan dengan bahasa sederhana oleh tim edukasi nirlaba.
Sumber terverifikasi
Setiap cerita terhubung ke bukti terkuat yang tersedia: sumber asli jika tersedia, dan laporan yang jelas dikaitkan dengan jelas.
Bahasa Inggris Biasa
Apa yang terjadi, mengapa hal itu penting, dan apa yang harus diperhatikan — tanpa jargon.
Tidak ada pengisi
Ketika sinyalnya tipis, kami tidak mempublikasikan apa pun selain mengisi feed.
Lebih banyak cerita
9 ceritaInovasi
Preprint reports higher LLM candidate ratings for prestigious institutions
An arXiv preprint reports that four large language models rated candidate profiles more favorably when linked to higher-prestige institutions and journals. The authors say institutional and geographic cues can shape evaluation, while the tested models and professional domains remain unspecified in the supplied source.arxiv.orgInovasi
Studi menemukan konteks inferensi dapat mengubah keluaran LLM dalam skenario alokasi medis
Makalah arXiv melaporkan bahwa tiga dari empat model bahasa yang diuji mengubah probabilitas alokasi sumber daya secara berbeda ketika skenario klinis yang sama dievaluasi dengan atau tanpa respons model sebelumnya dalam konteks.arxiv.orgInovasi
Preprint reports AI decoding of natural sentences from non-invasive brain recordings
A research team describes Brain2Qwerty v2, a model that uses real-time MEG recordings to decode typed natural sentences, reporting a 39% average word error rate across nine subjects.arxiv.orgInovasi
Study maps 46 language models built for Portuguese
A systematic mapping study catalogs 46 Portuguese language models and compares architectures, training resources, licensing, code, data, and weights. The authors say the field is growing but difficult to assess because information is spread across papers, technical reports, repositories, and project documentation.arxiv.orgInovasi
Interpretability study links Qwen3-4B’s overconfident answers to a certainty-biased mechanism
An arXiv study reports that Qwen3-4B favors certainty over uncertainty in controlled reasoning tasks and identifies model features that may drive the imbalance. The authors say targeted interventions reduced overconfident errors, but the abstract does not disclose effect sizes or testing details.arxiv.orgInovasi
The Deontic Gap: Large Language Models and the Modal Language of Obligation
Modal auxiliaries such as must, should, and have to mark necessity and obligation within the contexts of speaker authority and interpersonal stance.arxiv.orgKeamanan
Paper proposes obscuring refusal signals to resist abliteration attacks
An arXiv preprint introduces a weight-editing method intended to make safety refusals harder to extract and remove. The paper reports stronger post-abliteration refusal scores on two open models, with different tradeoffs in general-purpose performance.arxiv.orgInovasi
Paper reports a temporal method for detecting hallucinations at the token level
An arXiv preprint describes a hallucination detector that combines text statistics, entailment signals and language-model surprisal across sequences instead of judging tokens independently. Its BiGRU model reached an AUC of 0.840 on RAGTruth, according to the paper.arxiv.orgInovasi
Entity tracking emerges at 410 million parameters, exceeds humans across naturalistic narratives
Understanding language requires tracking entities across discourse - i.e., knowing where things are and how they change, even when not explicitly stated.arxiv.org
Satu pengarahan berguna setiap minggu
Ikuti AI tanpa harus hidup di feed.
Dapatkan berita AI terverifikasi minggu ini, data asli, alat berguna, pilihan pembelajaran, dan pekerjaan AI segar.
Jangkau orang-orang yang sedang mempelajari AI
Menyewa profesional AI atau meluncurkan produk AI yang berguna? Tunjukkan kepada orang-orang yang datang ke sini untuk belajar dan bertindak.
Posting pekerjaan AIKirim alat AI