返回新闻
工业AI Understanding 简报

Positron 为 LPDDR5X 推理芯片筹集了 8.75 亿美元资金

据 Tech Times 报道,Positron AI 以 50 亿美元的估值筹集了 8.75 亿美元,用于开发 Asimov,这是一种围绕商品 LPDDR5X 内存而不是 HBM 设计的推理 ASIC。该芯片仍是未经验证的硅片,计划于 2026 年末流片,并计划于 2027 年下半年投入生产。

4 min readRead the linked source
Source-page capture accompanying Positron raises $875 million for LPDDR5X inference chip
来源参考来源记录
出版商
techtimes.com
来源链接
techtimes.comhttps://www.techtimes.com/articles/327400/20260912/positron-ai-raises-875m-prove-commodity-memory-can-beat-hbm-inference.htm
来源类型
链接来源——主要来源状态尚未确定。
背景60 秒内了解这一点

从这里开始

关键术语

推理
经过训练的模型生成预测或输出的运行时阶段。
内存(代理内存)
AI 代理跨步骤或会话使用存储的上下文来提高连续性。
测试一下自己AI 模型解释测验

发生了什么

Tech Times reports that Reno-based Positron AI closed an $875 million Series C financing round at a $5 billion post-money valuation. The funding is intended to support Asimov, a custom ASIC targeted for tapeout by the end of 2026 and production in the second half of 2027. The design uses LPDDR5X memory rather than the HBM used in leading Nvidia accelerators.

Tech Times reports that Positron’s financing came in two tranches: a $375 million Series C at a reported $3.5 billion pre-money valuation and a Series C-1 of up to $500 million. The outlet says investors included NEA, Andra Capital, Atreides Management, Valor Equity Partners, SemiAnalysis Capital, the Qatar Investment Authority, Cisco Investments and others. The report says Positron plans to use the capital for Asimov and related engineering infrastructure.

According to Tech Times, Asimov is designed to use LPDDR5X, a commodity memory type used in smartphones and laptops, instead of HBM. The company claims its architecture could achieve more than 90% memory-bandwidth utilization, while the report says Positron characterizes typical GPU decode utilization as below 30%. Those figures are attributed to Positron and are not independently confirmed. Tech Times also reports that Positron’s Atlas system is deployed in more than 50 Oracle Cloud Infrastructure racks and is used by the Parasail inference service, with Jump Trading and i3d.net identified as production customers. The source does not establish general availability or pricing.

来源详情: techtimes.com ↗

为什么这很重要

The financing highlights a consequential hardware dispute over how AI should be optimized. Positron argues that sequential model decoding often uses only a fraction of an accelerator’s theoretical memory bandwidth, making effective utilization, memory capacity and supply-chain availability more important than peak bandwidth alone. If the company’s claims hold, its approach could offer lower-cost and more widely deployable inference infrastructure. However, Tech Times’ account makes clear that Asimov has not taped out and that its performance figures remain company design targets, not independent measurements.

Tech Times frames the round as a bet on -specific hardware at a time when AI systems are increasingly run continuously for users rather than only trained. The report says Positron’s design targets between 288 gigabytes and 2,304 gigabytes of memory per Asimov chip, with additional capacity possible through CXL expansion, and says Titan is intended to serve models exceeding 16 trillion parameters. These are planned specifications, not demonstrated capabilities.

The supply-chain argument is also significant. The source says HBM depends on specialized manufacturing and advanced packaging, while LPDDR5X is produced by multiple suppliers at larger commodity scale. That could matter to buyers constrained by memory availability, rack power or cooling requirements. Tech Times reports that Titan is being designed for both air- and liquid-cooled facilities, but no independent total-cost, throughput or efficiency result is provided.

The central caveat is validation. The source cites earlier third-party commentary that Positron’s Atlas comparisons with Nvidia systems require verification, and says Asimov has not yet taped out. Tech Times also reports a substantially lower realizable-bandwidth estimate for Asimov than the peak bandwidth of Nvidia’s future Rubin architecture. The outcome therefore remains unresolved.

Interactive Mechanism

互动机制:它实际上是如何运作的

以交互方式探索这一发展背后的基础技术。

Document Size:128K tokens
Needle Placement Depth (Location in document):50% into text
Attention Context Buffer Map:
Target Fact (50%)
Equivalent Pages~320Standard book pages
Retrieval Accuracy99.9%Needle recall score
RAM / KV Cache5.1 GBMemory overhead
Prompt CachingActive~80% discount on reuse
Core takeaway: Million-token context windows allow querying whole codebases or legal archives in one prompt. However, KV cache memory scales with context length, making prompt caching crucial for real-time production.
交互式概念检查+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

接下来看什么

The key tests are Asimov’s tapeout, independent benchmarks, production availability and the economics of Titan systems built from Asimov chips. Access terms, customer eligibility and pricing for Asimov and Titan are not documented in the source. Positron’s current Atlas deployments provide operating experience, but they do not independently validate Asimov’s future performance claims.

Watch for Asimov’s reported late-2026 tapeout, working silicon and independent tests across different model sizes, context lengths and decode workloads. The most useful measures will be tokens per dollar, tokens per watt, usable memory capacity and sustained throughput rather than peak specifications alone.

Watch whether the reported Atlas deployments expand beyond the current customers and whether Positron offers a clear procurement path for enterprises. The source does not document public purchase terms, cloud access, service pricing or a general release date for Asimov or Titan.

Also watch whether LPDDR5X supply and packaging advantages translate into reliable production at scale. The source presents Positron’s roadmap and investor support as evidence of confidence, but neither establishes that the future ASIC will match HBM-based systems in real-world performance or economics.

相关指南和测验

人工智能模型解释变形金刚AI 的未来测试你所知道的——尝试免费的人工智能测验在我们的词汇表中查找人工智能术语关注 AI 资金追踪器
觉得这有用吗?