返回新聞
產業AI Understanding 簡報

TechCrunch:Nvidia 的 AI 優勢正在擴展到 GPU 之外

TechCrunch 報導,Nvidia 透過 CPU、網路、儲存和編排系統日益激烈地競爭,這些系統可幫助大型 AI 資料中心圍繞 GPU 高效地移動資料。

5 min readRead the original reporting
Source-provided image accompanying TechCrunch: Nvidia’s AI advantage is expanding beyond GPUs
歸因報告來源記錄
出版商
techcrunch.com
來源連結
techcrunch.comhttps://techcrunch.com/2026/08/29/nvidias-ai-advantage-is-moving-beyond-the-gpu/
來源類型
新聞媒體的報道-不是第一方文件。

我們無法獨立確認的內容: 此聲明歸因於指定的商店。我們沒有根據第一方文件對其進行驗證。 (techcrunch.com)

背景60 秒內了解這一點

從這裡開始

關鍵術語

記憶體(代理記憶體)
AI 代理程式跨步驟或會話使用儲存的上下文來提高連續性。
基準測試
用於測量和比較模型性能的標準化測試或資料集。
推理
經過訓練的模型產生預測或輸出的運行時階段。
測試一下自己AI 模型解釋測驗

發生了什麼事

TechCrunch reports that Nvidia’s competitive advantage in AI infrastructure is shifting beyond its graphics processors. As AI data centers scale toward gigawatt-level power consumption, the company is selling integrated systems intended to coordinate memory, storage, networking and compute around the GPU. The report focuses on Nvidia’s Vera Rubin architecture, which combines the Rubin GPU with the Vera CPU, Groq 3 LPX accelerators and related storage and networking racks. Nvidia executive Jason Hardy told TechCrunch that the Vera CPU is designed to help direct data to the GPU at the right time. Hardy said Nvidia saw up to a threefold improvement in certain operations, though TechCrunch did not independently verify that claim. TechCrunch also compares Nvidia’s approach with OpenAI’s Jalapeño chip, which was designed to reduce data movement by keeping an entire workload within one connected system. The approaches differ, but both address the same infrastructure problem: moving data efficiently can matter as much as adding processor capacity.

TechCrunch reports that Nvidia’s recent AI advantage is increasingly tied to infrastructure surrounding the GPU. The outlet says AI computing is growing toward gigawatt scale, making it more difficult to operate large data centers efficiently. In that environment, the article argues, coordinating data movement and system components becomes a central engineering challenge rather than a secondary concern.

The report identifies Nvidia’s Vera Rubin architecture as an example of this strategy. According to TechCrunch, the architecture pairs the Rubin GPU with the Vera CPU, Groq 3 LPX accelerators and comparable racks for storage and networking. The outlet describes these components as specialized systems intended to improve the operation of everything around the GPU, rather than simply processing tokens themselves.

Jason Hardy, Nvidia’s vice president of storage technology, told TechCrunch that the Vera CPU addresses the limited memory capacity of a single server or compute platform and helps orchestrate data. TechCrunch reports Hardy as saying Nvidia observed up to a threefold improvement in certain operations and could use flash storage more fully without creating a bottleneck. That performance claim comes from Nvidia and was not independently confirmed in the source.

TechCrunch contrasts Nvidia’s system-level approach with OpenAI’s Jalapeño chip. The outlet quotes OpenAI as saying Jalapeño was designed to minimize data movement and communication delays by keeping an entire workload within one connected system. Nvidia and OpenAI are using different designs, but TechCrunch says both seek efficiency through smarter control of data movement rather than only through additional processor cycles.

來源詳情: techcrunch.com ↗

為什麼這很重要

The report describes a shift in how AI infrastructure may be evaluated. GPU performance remains important, but the practical output of a large AI system can also depend on memory access, storage, networking and the software and hardware used to coordinate those components. That could make it harder for competitors to challenge Nvidia simply by offering an alternative accelerator. A rival would need to match the performance of a broader system, including the connections among chips and the mechanisms that keep expensive processors supplied with data. The implications remain uncertain. TechCrunch’s account is based partly on Nvidia’s own explanation of its systems and an executive’s performance claim. The source does not provide independent results, pricing, deployment figures or evidence that Nvidia’s full-system advantage will persist as hyperscalers develop their own infrastructure.

The report matters because it broadens the definition of AI computing capacity. A powerful accelerator can be underused if data, model parameters or intermediate results do not reach it quickly enough. The article’s central point is that system design can determine how effectively expensive AI processors are used.

This creates a possible barrier to challengers. TechCrunch reports that hyperscalers such as Amazon and Google have developed their own chips, reducing Nvidia’s status as the only provider of advanced AI GPUs. But an alternative GPU or accelerator may not be sufficient if it lacks comparable memory, storage, networking and orchestration capabilities.

The shift could also affect infrastructure spending. If data movement and coordination are major constraints, buyers may need to evaluate complete racks and system architectures rather than compare accelerator specifications alone. That could strengthen the position of suppliers that can provide tightly integrated hardware and software, while potentially increasing the complexity and cost of switching vendors.

The evidence has important limits. TechCrunch’s reporting includes conversations with Nvidia personnel and a performance figure supplied by a Nvidia executive, but the source provides no independent testing, customer deployment data, prices, power measurements or comparison with competing systems. The report supports the existence of Nvidia’s broader system strategy, but it does not prove that the company will maintain a durable lead.

Interactive Mechanism

互動機制:它實際上是如何運作的

以互動方式探索這項發展背後的基礎技術。

Document Size:128K tokens
Needle Placement Depth (Location in document):50% into text
Attention Context Buffer Map:
Target Fact (50%)
Equivalent Pages~320Standard book pages
Retrieval Accuracy99.9%Needle recall score
RAM / KV Cache5.1 GBMemory overhead
Prompt CachingActive~80% discount on reuse
Core takeaway: Million-token context windows allow querying whole codebases or legal archives in one prompt. However, KV cache memory scales with context length, making prompt caching crucial for real-time production.
互動式概念檢查+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

接下來看什麼

Watch whether Nvidia publishes independent, reproducible measurements for the Vera Rubin system, including end-to-end throughput, energy use, memory and storage performance, and the conditions behind the reported threefold improvement. Watch how Amazon, Google and other large cloud operators respond. TechCrunch reports that hyperscalers have been developing their own chips, but the more consequential competition may involve complete systems for coordinating compute, memory, storage and networking. Also watch whether customers buy Nvidia’s integrated infrastructure as a package or substitute components from multiple vendors. The report suggests that orchestration is becoming a major competitive layer, but it does not establish how widely Nvidia’s systems are available, what they cost, or how they perform in production.

The most useful next evidence would be transparent, third-party testing of Vera Rubin systems. Such testing should separate GPU performance from gains attributable to the Vera CPU, flash storage, networking and orchestration, and should explain the workloads and baseline used for any claimed improvement.

Availability and purchasing terms will also matter. The source describes Nvidia’s architecture and current rollout but does not say which customers can obtain the systems, in what quantities, at what price, or whether individual components can be mixed with hardware from other suppliers. Those details will determine whether the strategy is broadly practical or mainly an integrated Nvidia offering.

Competitors’ responses will show whether the market is moving toward system-level competition. Cloud providers may design alternative architectures around their own chips, while other chipmakers may focus on memory, networking, storage or interconnects rather than compete directly on GPU performance. The source does not identify specific rival systems or provide evidence about their current capabilities.

OpenAI’s Jalapeño design is a useful comparison, but it should not be treated as proof that one approach is superior. OpenAI’s statement, quoted by TechCrunch, describes an effort to reduce data movement inside a connected chip. More information is needed to compare that design with Nvidia’s distributed orchestration model across real workloads, energy use, reliability and total operating cost.

相關指引和測驗

人工智慧模型解釋變形金剛人工智慧培訓AI 的未來測試你所知道的—嘗試免費的人工智慧測驗在我們的詞彙表中尋找人工智慧術語關注 AI 資金追蹤器
覺得有用嗎?