뉴스로 돌아가기
산업AI Understanding 브리핑

Positron은 LPDDR5X 추론 칩을 위해 8억 7,500만 달러를 모금했습니다.

Tech Times는 Positron AI가 HBM 대신 상용 LPDDR5X 메모리를 중심으로 설계된 추론 ASIC인 Asimov를 개발하기 위해 50억 달러 가치로 8억 7,500만 달러를 모금했다고 보도했습니다. 이 칩은 2026년 후반에 테이프아웃을 목표로 하고 2027년 하반기에 생산을 계획하는 등 검증되지 않은 실리콘으로 남아 있습니다.

4 min readRead the linked source
Source-page capture accompanying Positron raises $875 million for LPDDR5X inference chip
소스 참조녹음된 소스
출판사
techtimes.com
소스 링크
techtimes.comhttps://www.techtimes.com/articles/327400/20260912/positron-ai-raises-875m-prove-commodity-memory-can-beat-hbm-inference.htm
소스 유형
연결된 소스 — 기본 소스 상태가 설정되지 않았습니다.
맥락60초 안에 이해하세요

여기서 시작하세요

주요 용어

추론
훈련된 모델이 예측 또는 출력을 생성하는 런타임 단계입니다.
메모리(에이전트 메모리)
AI 에이전트는 연속성을 향상하기 위해 여러 단계 또는 세션에서 사용하는 저장된 컨텍스트입니다.
자신을 테스트해 보세요AI 모델 설명 퀴즈

무슨 일이 일어났나요?

Tech Times reports that Reno-based Positron AI closed an $875 million Series C financing round at a $5 billion post-money valuation. The funding is intended to support Asimov, a custom ASIC targeted for tapeout by the end of 2026 and production in the second half of 2027. The design uses LPDDR5X memory rather than the HBM used in leading Nvidia accelerators.

Tech Times reports that Positron’s financing came in two tranches: a $375 million Series C at a reported $3.5 billion pre-money valuation and a Series C-1 of up to $500 million. The outlet says investors included NEA, Andra Capital, Atreides Management, Valor Equity Partners, SemiAnalysis Capital, the Qatar Investment Authority, Cisco Investments and others. The report says Positron plans to use the capital for Asimov and related engineering infrastructure.

According to Tech Times, Asimov is designed to use LPDDR5X, a commodity memory type used in smartphones and laptops, instead of HBM. The company claims its architecture could achieve more than 90% memory-bandwidth utilization, while the report says Positron characterizes typical GPU decode utilization as below 30%. Those figures are attributed to Positron and are not independently confirmed. Tech Times also reports that Positron’s Atlas system is deployed in more than 50 Oracle Cloud Infrastructure racks and is used by the Parasail inference service, with Jump Trading and i3d.net identified as production customers. The source does not establish general availability or pricing.

소스 세부정보: techtimes.com ↗

왜 중요한가요?

The financing highlights a consequential hardware dispute over how AI should be optimized. Positron argues that sequential model decoding often uses only a fraction of an accelerator’s theoretical memory bandwidth, making effective utilization, memory capacity and supply-chain availability more important than peak bandwidth alone. If the company’s claims hold, its approach could offer lower-cost and more widely deployable inference infrastructure. However, Tech Times’ account makes clear that Asimov has not taped out and that its performance figures remain company design targets, not independent measurements.

Tech Times frames the round as a bet on -specific hardware at a time when AI systems are increasingly run continuously for users rather than only trained. The report says Positron’s design targets between 288 gigabytes and 2,304 gigabytes of memory per Asimov chip, with additional capacity possible through CXL expansion, and says Titan is intended to serve models exceeding 16 trillion parameters. These are planned specifications, not demonstrated capabilities.

The supply-chain argument is also significant. The source says HBM depends on specialized manufacturing and advanced packaging, while LPDDR5X is produced by multiple suppliers at larger commodity scale. That could matter to buyers constrained by memory availability, rack power or cooling requirements. Tech Times reports that Titan is being designed for both air- and liquid-cooled facilities, but no independent total-cost, throughput or efficiency result is provided.

The central caveat is validation. The source cites earlier third-party commentary that Positron’s Atlas comparisons with Nvidia systems require verification, and says Asimov has not yet taped out. Tech Times also reports a substantially lower realizable-bandwidth estimate for Asimov than the peak bandwidth of Nvidia’s future Rubin architecture. The outcome therefore remains unresolved.

Interactive Mechanism

대화형 메커니즘: 실제로 작동하는 방식

이 개발의 이면에 있는 기본 기술을 대화식으로 살펴보세요.

Document Size:128K tokens
Needle Placement Depth (Location in document):50% into text
Attention Context Buffer Map:
Target Fact (50%)
Equivalent Pages~320Standard book pages
Retrieval Accuracy99.9%Needle recall score
RAM / KV Cache5.1 GBMemory overhead
Prompt CachingActive~80% discount on reuse
Core takeaway: Million-token context windows allow querying whole codebases or legal archives in one prompt. However, KV cache memory scales with context length, making prompt caching crucial for real-time production.
대화형 개념 확인+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

다음에 무엇을 볼 것인가

The key tests are Asimov’s tapeout, independent benchmarks, production availability and the economics of Titan systems built from Asimov chips. Access terms, customer eligibility and pricing for Asimov and Titan are not documented in the source. Positron’s current Atlas deployments provide operating experience, but they do not independently validate Asimov’s future performance claims.

Watch for Asimov’s reported late-2026 tapeout, working silicon and independent tests across different model sizes, context lengths and decode workloads. The most useful measures will be tokens per dollar, tokens per watt, usable memory capacity and sustained throughput rather than peak specifications alone.

Watch whether the reported Atlas deployments expand beyond the current customers and whether Positron offers a clear procurement path for enterprises. The source does not document public purchase terms, cloud access, service pricing or a general release date for Asimov or Titan.

Also watch whether LPDDR5X supply and packaging advantages translate into reliable production at scale. The source presents Positron’s roadmap and investor support as evidence of confidence, but neither establishes that the future ASIC will match HBM-based systems in real-world performance or economics.

관련 가이드 및 퀴즈

AI 모델 설명트랜스포머AI의 미래알고 있는 내용을 테스트해 보세요. 무료 AI 퀴즈를 시도해 보세요.용어집에서 AI 용어를 찾아보세요.AI 자금 추적기를 팔로우하세요
이것이 유용하다고 생각하시나요?