Powrót do Wiadomości
ProduktAI Understanding odprawa

Nvidia ogłasza ogólną dostępność bibliotek cuObject i SDK serwera SCADA

Nvidia wypuścił biblioteki klienckie i serwerowe cuObject oraz nowy pakiet SDK serwera SCADA, umożliwiający przyspieszony dostęp do pamięci obiektowej oparty na RDMA dla obciążeń AI i otwierający ścieżkę dla interoperacyjnych rozwiązań pamięci masowej.

4 min readRead the primary source
Source-provided image accompanying Nvidia announces general availability of cuObject libraries and SCADA Server SDK
Dokument źródłowyŹródło zapisane
Wydawca
developer.nvidia.com
Link źródłowy
developer.nvidia.comhttps://developer.nvidia.com/blog/expanding-ai-storage-access-with-nvidia-cuobject-and-the-nvidia-scada-server-sdk/
Typ źródła
Dokument podstawowy — oficjalne ogłoszenie, dokument, zgłoszenie lub strona własna, którą czytamy bezpośrednio.
KontekstZrozum to w 60 sekund

Zacznij tutaj

Kluczowe terminy

Pamięć (pamięć agenta)
Przechowywany kontekst, którego agent AI używa na różnych etapach lub sesjach, aby poprawić ciągłość.
Wnioskowanie
Faza środowiska uruchomieniowego, w której przeszkolony model generuje prognozy lub dane wyjściowe.
Opóźnienie
Czas między wysłaniem żądania a otrzymaniem danych wyjściowych modelu.
Sprawdź sięQuiz objaśniający modele AI

Co się stało

Nvidia announced that its cuObject libraries are now generally available, expanding the xio‑sig effort that previously included cuFile. The cuObject client and server APIs expose an RDMA wire protocol for accelerated object‑storage operations, allowing GPUs to read and write data without routing through the host CPU. In parallel, Nvidia released a SCADA Server SDK that lets storage providers build servers capable of handling GPU‑initiated requests from SCADA clients. IBM demonstrated a prototype integration of the SDK with its Storage Scale product, showing early interoperability. Nvidia cites partnerships with Google Cloud and Microsoft, both of which are evaluating participation in the cuObject effort, and positions the releases as part of its broader Storage‑Next initiative involving more than 40 industry participants.

Nvidia’s developer blog confirmed that the cuObject client and server libraries have reached general availability. The libraries expose a set of APIs that allow AI accelerators to perform object‑storage reads and writes over RDMA, bypassing the host CPU’s memory path. This mirrors the earlier cuFile offering for file‑based storage, extending the accelerated access model to object stores commonly used in cloud environments.

Alongside cuObject, Nvidia introduced the SCADA Server SDK, a software kit for storage vendors to build servers that can receive and fulfill GPU‑initiated storage requests. The SDK includes a command‑line utility, a storage‑lender service, and reference implementations for both client and server sides. IBM’s prototype, which integrates the SDK with its Storage Scale platform, demonstrates that the approach can work across heterogeneous storage back‑ends.

The announcement highlights ongoing collaborations with Google Cloud and Microsoft, both of which are evaluating deeper involvement in the cuObject effort. Nvidia also notes that the repository structure for cuFile and cuObject, including headers and wire‑protocol definitions, will be made publicly available once the stack passes conformance tests, indicating a commitment to open‑source transparency.

Szczegóły źródła: developer.nvidia.com ↗

Dlaczego to ma znaczenie

Accelerated storage access is a growing bottleneck for large‑scale AI training and , where datasets often reside in remote file or object stores. By enabling zero‑copy, RDMA‑based transfers, cuObject reduces and CPU overhead, potentially increasing throughput for data‑intensive models such as large language models or recommendation systems. The open‑source nature of the libraries and the shared wire protocol aim to standardize GPU‑driven storage access across cloud providers and on‑premise solutions, lowering integration effort for developers and fostering a broader ecosystem of compatible storage products. This could translate into faster model iteration cycles and lower operational costs for enterprises that rely on massive data pipelines.

AI workloads increasingly depend on rapid access to large datasets stored in remote object stores. Traditional storage paths involve copying data through the server’s CPU, incurring and consuming CPU cycles that could otherwise be allocated to model computation. cuObject’s RDMA‑based zero‑copy pathway directly addresses this inefficiency, offering higher throughput and lower latency for data‑intensive training and tasks.

Standardizing the wire protocol through xio‑sig and providing open‑source client/server libraries reduces the engineering burden for developers who previously had to write provider‑specific integrations. This interoperability can accelerate the deployment of AI pipelines across multiple cloud and on‑premise environments, fostering a more competitive storage market.

The broader Storage‑Next initiative, which includes over 40 vendors, aims to codify best practices for GPU‑driven storage access. By contributing to an industry‑wide standard, Nvidia positions itself as a key infrastructure provider, potentially influencing future hardware and software designs that prioritize high‑speed data movement.

Interactive Mechanism

Mechanizm interaktywny: jak to faktycznie działa

Poznaj interaktywnie technologię leżącą u podstaw tego rozwoju.

Document Size:128K tokens
Needle Placement Depth (Location in document):50% into text
Attention Context Buffer Map:
Target Fact (50%)
Equivalent Pages~320Standard book pages
Retrieval Accuracy99.9%Needle recall score
RAM / KV Cache5.1 GBMemory overhead
Prompt CachingActive~80% discount on reuse
Core takeaway: Million-token context windows allow querying whole codebases or legal archives in one prompt. However, KV cache memory scales with context length, making prompt caching crucial for real-time production.
Interaktywna kontrola koncepcji+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

Co obejrzeć dalej

Key indicators to monitor include adoption of cuObject by major cloud providers, the outcome of Nvidia’s conformance testing for the production‑ready stack, and further prototype demonstrations from storage vendors beyond IBM. The evolution of the xio‑sig governance board and the pace at which additional partners join will signal how quickly the industry moves toward a common standard. Finally, any pricing or licensing details released for the libraries or SDK will affect accessibility for startups and research institutions.

The rate at which cloud providers such as Google Cloud and Microsoft Azure adopt cuObject in their storage services will be a primary gauge of market impact.

Completion of Nvidia’s conformance testing and the public release of the cuObject repository will determine how quickly third‑party developers can integrate the libraries into their applications.

Additional prototype demonstrations from other storage vendors, as well as any announced pricing or licensing models for the SDK, will clarify the accessibility of the technology for startups and research labs.

Powiązane przewodniki i quizy

Wyjaśnienie modeli AISzkolenie AIPrzyszłość AISprawdź swoją wiedzę — wypróbuj darmowy quiz dotyczący sztucznej inteligencjiWyszukaj termin związany ze sztuczną inteligencją w naszym glosariuszuPostępuj zgodnie z modułem śledzenia wydań modeli AI
Uznałeś to za przydatne?