返回新聞
企業AI Understanding 簡報

Semi Five 與美國無晶圓廠公司簽訂 5,200 萬美元人工智慧加速器合約

Semi Five 宣布簽訂價值 5,200 萬美元的合同,為一家美國無晶圓廠公司設計下一代人工智慧推理加速器,這標誌著其在北美的首次 Spec Hand-off 活動。

4 min readRead the linked source
Source-page capture accompanying Semifive secures $52 million AI accelerator contract with US fabless company
來源參考來源記錄
出版商
prnewswire.com
來源連結
prnewswire.comhttps://www.prnewswire.com/news-releases/semifive-secures-usd-52-million-ai-accelerator-contract-with-us-fabless-company-validating-end-to-end-asic-model-in-north-america-302892224.html
來源類型
連結來源-主要來源狀態尚未確定。
背景60 秒內了解這一點

從這裡開始

關鍵術語

大語言模型(LLM)
在海量文本語料庫上訓練來產生和分析文本的語言模型。
記憶體(代理記憶體)
AI 代理程式跨步驟或會話使用儲存的上下文來提高連續性。
基準測試
用於測量和比較模型性能的標準化測試或資料集。
測試一下自己人工智慧測驗的未來

發生了什麼事

Semifive, a global provider of custom AI ASIC solutions, signed a contract worth roughly KRW 70.3 billion (US $52 million) with an unnamed U.S.‑based AI fabless company to develop a next‑generation AI inference accelerator. The deal represents more than 40 % of Semifive’s total orders for 2025 and about 60 % of new orders booked in the first half of 2026, making it the company’s largest single contract to date. Under Semifive’s “Spec Hand‑off” model, the customer supplies performance specifications while Semifive handles full‑stack development, including design, software, packaging, testing, and mass production. The accelerator will use LPDDR6 memory and PCIe Gen5 interfaces, target high memory bandwidth, low power consumption, and reduced total cost of ownership. Tape‑out is planned for the first half of 2027, with mass production slated for 2028 aimed at hyperscalers and cloud service providers.

Semifive’s press release states the contract is valued at KRW 70.3 billion (approximately US $52 million) and is the company’s first "Spec Hand‑off" engagement in the North American market. The agreement accounts for over 40 % of the firm’s total order book for 2025 and roughly 60 % of the new orders recorded in the first half of 2026.

The Spec Hand‑off model differs from traditional turnkey engagements by allowing the customer to provide only high‑level performance specifications, while Semifive assumes responsibility for detailed chip design, software development, packaging, testing, and eventual mass production. This approach aims to reduce time‑to‑market and development risk for the fabless partner.

Technical highlights include the integration of LPDDR6 memory to lower power consumption associated with memory accesses, and PCIe Gen5 interfaces to alleviate data‑transfer bottlenecks. Semifive also plans to leverage its experience with large‑die (up to 800 mm²) projects to deliver high compute throughput and operational stability.

The development timeline targets a tape‑out in the first half of 2027, followed by global mass production in 2028. The accelerator is intended for deployment by hyperscale cloud service providers and other large‑scale AI inference workloads.

來源詳情: prnewswire.com ↗

為什麼這很重要

The contract underscores the growing demand for custom AI silicon optimized for large‑scale model inference, a market segment traditionally dominated by in‑house designs at firms like Google and Amazon. By securing a sizable North American deal, Semifive expands its footprint beyond Asia and positions itself as a viable partner for U.S. hyperscale cloud operators seeking to lower inference costs and power usage. The use of LPDDR6 and PCIe Gen5 reflects industry trends toward higher bandwidth and energy‑efficient memory solutions, which could influence future ASIC design standards. If the accelerator meets its performance targets, it may provide a cost‑effective alternative to existing offerings from established ASIC vendors, potentially reshaping the competitive dynamics of the AI hardware supply chain.

Custom AI ASICs are increasingly critical for reducing the cost and energy footprint of large language model inference. By securing a major contract with a U.S. fabless firm, Semifive demonstrates its ability to compete for high‑value design work outside its traditional Asian market.

The contract’s size relative to Semifive’s overall order book highlights the strategic importance of the North American AI hardware market and suggests that customers are seeking alternatives to the dominant players that have historically controlled AI silicon supply chains.

Adoption of LPDDR6 and PCIe Gen5 aligns with industry moves toward higher bandwidth, lower‑power memory subsystems, potentially setting a new for future AI inference chips.

If successful, the accelerator could provide cloud providers with a more cost‑effective solution for serving AI workloads, influencing pricing and performance expectations across the AI inference ecosystem.

Interactive Mechanism

互動機制:它實際上是如何運作的

以互動方式探索這項發展背後的基礎技術。

Document Size:128K tokens
Needle Placement Depth (Location in document):50% into text
Attention Context Buffer Map:
Target Fact (50%)
Equivalent Pages~320Standard book pages
Retrieval Accuracy99.9%Needle recall score
RAM / KV Cache5.1 GBMemory overhead
Prompt CachingActive~80% discount on reuse
Core takeaway: Million-token context windows allow querying whole codebases or legal archives in one prompt. However, KV cache memory scales with context length, making prompt caching crucial for real-time production.
互動式概念檢查+10 Points
Future of AI Quiz

What did neural scaling law research (e.g. Kaplan et al., 2020) observe?

接下來看什麼

Key milestones to monitor include the scheduled tape‑out in early 2027 and the commencement of mass production in 2028. Observers should watch for announcements of the unnamed U.S. fabless partner, which could reveal which cloud providers or AI workloads the accelerator will serve. Additionally, any performance benchmarks or power‑efficiency data released by Semifive will indicate whether the chip can compete with incumbent solutions from companies like Nvidia, Broadcom, or Marvell. Finally, the broader market response—such as interest from other hyperscalers or potential follow‑on contracts—will signal the commercial viability of Semifive’s Spec Hand‑off model in North America.

The scheduled tape‑out in early 2027 will be a key technical milestone; any delays or design changes could affect the projected 2028 mass‑production timeline.

Identification of the U.S. fabless partner will clarify which hyperscalers or AI service providers are likely to adopt the accelerator, offering insight into market demand.

Performance and power‑efficiency benchmarks released after tape‑out will be essential for assessing competitiveness against existing solutions from Nvidia, Broadcom, Marvell, and other ASIC vendors.

Potential follow‑on contracts or additional orders from other North American customers will indicate whether Semifive’s Spec Hand‑off model gains broader acceptance in the region.

相關指引和測驗

AI 的未來人工智慧模型解釋人工智慧培訓測試你所知道的—嘗試免費的人工智慧測驗在我們的詞彙表中尋找人工智慧術語關注 AI 資金追蹤器
覺得有用嗎?