검증된 출처
모든 기사는 가장 강력한 증거와 연결됩니다: 원가 자료가 있을 경우 그렇고, 그 외에는 명확히 출처가 명확히 기재된 보도입니다.
평범한 영어
무슨 일이 일어났는지, 왜 중요한지, 무엇을 보아야 하는지 등을 전문 용어 없이 설명합니다.
채우기 없는 이야기
신호가 얇을 때는 피드를 채우기보다는 아무것도 발행하지 않습니다.
더 많은 이야기
9 이야기들제품
Qwen presents Qwen3.8-Flash-Next as an open-weight preview of its next architecture
Qwen describes Qwen3.8-Flash-Next as a multimodal mixture-of-experts model with 125 billion total tokens and 6 billion active parameters, and says it previews architecture planned for Qwen4.qwen.ai혁신
LUCAID reports 93% concordance in prospective lung-cancer pathology validation
A preprint describes LUCAID, an agentic multimodal AI system that integrates nine lung-cancer pathology tasks and reports 93.0% concordance with an expert-panel reference standard in prospective clinical validation.arxiv.org혁신
DRRG Uses Discrete Diffusion to Iteratively Refine Radiology Reports
An arXiv paper introduces DRRG, a discrete-diffusion framework that generates radiology reports through iterative masked-token denoising. The authors report stronger results than comparison methods on MIMIC-CXR and CheXpert Plus across several metrics, while using a smaller language-model decoder.arxiv.org산업
Forbes Australia reports Blackbird raises record $1.05 billion fund and shifts AI focus toward deep tech
Blackbird has raised Australia’s largest venture-capital fund, while general partner Samantha Wong says the firm is becoming more cautious about crowded AI applications and expects to invest more in infrastructure, semiconductors and other deep technologies.forbes.com.au혁신
VisCache Reports Faster Vision-Language Model Inference With Selective Visual Cache Pruning
A new arXiv paper proposes VisCache, a no-training framework that reduces visual KV-cache storage for vision-language models while retaining 19% to 28% of the cache. The authors report speedups of up to 2.35 times, but the results have not been independently validated here.arxiv.org혁신
WeMM-Embedding Report Describes Open Multimodal Models for Search and Recommendation
A new technical report introduces WeMM-Embedding, a family of 2B, 4B and 9B models for representing text, images, videos and visual documents in a shared space. The report says the models improve internal WeChat benchmarks and have been deployed across search and recommendation services.arxiv.org혁신
연구에 따르면 비전 언어 모델은 Few-Shot 적응에서 모델 수준 한도에 도달했습니다.
새로운 arXiv 연구에서는 텍스트와 이미지 프로토타입 간의 혼합을 조정하는 것이 소수의 비전 언어 모델 정확도에 대한 주요 제한이 아니라고 보고합니다. 4,800개의 평가 셀을 대상으로 한 실험에서 검증이 필요 없는 선형 프로브는 테스트 세트 정보를 사용하여 선택된 혼합물보다 성능이 뛰어났습니다.arxiv.org혁신
연구에서는 보다 강력한 시각적 AI를 향한 경로로서 점진적인 모션 통합을 확인합니다.
영장류의 시각과 다중 신경망 설계를 비교한 새로운 arXiv 연구에서는 예측 세계 모델이 물체의 외양이 변할 때 가장 강력하다고 보고하며, 이는 동적 AI의 누락된 원리로서 모션의 점진적인 통합을 지적합니다.arxiv.org혁신
NeuralParker는 강화 학습을 사용하여 불규칙한 환경에서 주차 계획을 세웁니다.
새로운 사전 인쇄물에서는 배달 및 서비스 차량을 불규칙한 경계 영역에서 지정된 자세로 안내하도록 설계된 강화 학습 플래너인 NeuralParker를 선보이며 실제 차량 평가에서 성공했다고 보고되었습니다.arxiv.org
매주 한 번씩 유용한 브리핑을 듣습니다
피드에 살지 않고도 AI를 따라가세요.
이번 주의 검증된 AI 뉴스, 원본 데이터, 유용한 도구, 학습 추천, 그리고 새로운 AI 일자리를 받아보세요.
AI를 배우는 사람들에게 다가가세요
AI 전문가를 고용하거나 유용한 AI 제품을 출시하는 것? 배우고 행동하기 위해 이곳에 온 사람들 앞에 그것을 보여주세요.
AI 공고를 올리세요AI 도구를 제출하세요