Retour aux Actualités
ProduitBriefing AI Understanding

Nvidia annonce la disponibilité générale des bibliothèques cuObject et du SDK du serveur SCADA

Nvidia a publié les bibliothèques client et serveur cuObject ainsi qu'un nouveau SDK de serveur SCADA, permettant un accès accéléré au stockage d'objets basé sur RDMA pour les charges de travail d'IA et ouvrant la voie à des solutions de stockage interopérables.

4 min readRead the primary source
Source-provided image accompanying Nvidia announces general availability of cuObject libraries and SCADA Server SDK
Document de source principaleSource enregistrée
Éditeur
developer.nvidia.com
Lien source
developer.nvidia.comhttps://developer.nvidia.com/blog/expanding-ai-storage-access-with-nvidia-cuobject-and-the-nvidia-scada-server-sdk/
Type de source
Document principal : une annonce officielle, un document, un dépôt ou une page de première partie que nous lisons directement.
ContexteComprenez cela en 60 secondes

Commencez ici

Termes clés

Mémoire (mémoire de l'agent)
Contexte stocké qu'un agent IA utilise au fil des étapes ou des sessions pour améliorer la continuité.
Inférence
Phase d'exécution au cours de laquelle un modèle entraîné génère des prédictions ou des sorties.
Latence
Le temps entre l'envoi d'une requête et la réception de la sortie du modèle.
Testez-vousQuiz sur les modèles d'IA expliqués

Que s'est-il passé

Nvidia announced that its cuObject libraries are now generally available, expanding the xio‑sig effort that previously included cuFile. The cuObject client and server APIs expose an RDMA wire protocol for accelerated object‑storage operations, allowing GPUs to read and write data without routing through the host CPU. In parallel, Nvidia released a SCADA Server SDK that lets storage providers build servers capable of handling GPU‑initiated requests from SCADA clients. IBM demonstrated a prototype integration of the SDK with its Storage Scale product, showing early interoperability. Nvidia cites partnerships with Google Cloud and Microsoft, both of which are evaluating participation in the cuObject effort, and positions the releases as part of its broader Storage‑Next initiative involving more than 40 industry participants.

Nvidia’s developer blog confirmed that the cuObject client and server libraries have reached general availability. The libraries expose a set of APIs that allow AI accelerators to perform object‑storage reads and writes over RDMA, bypassing the host CPU’s memory path. This mirrors the earlier cuFile offering for file‑based storage, extending the accelerated access model to object stores commonly used in cloud environments.

Alongside cuObject, Nvidia introduced the SCADA Server SDK, a software kit for storage vendors to build servers that can receive and fulfill GPU‑initiated storage requests. The SDK includes a command‑line utility, a storage‑lender service, and reference implementations for both client and server sides. IBM’s prototype, which integrates the SDK with its Storage Scale platform, demonstrates that the approach can work across heterogeneous storage back‑ends.

The announcement highlights ongoing collaborations with Google Cloud and Microsoft, both of which are evaluating deeper involvement in the cuObject effort. Nvidia also notes that the repository structure for cuFile and cuObject, including headers and wire‑protocol definitions, will be made publicly available once the stack passes conformance tests, indicating a commitment to open‑source transparency.

Détails de la source: developer.nvidia.com ↗

Pourquoi c'est important

Accelerated storage access is a growing bottleneck for large‑scale AI training and , where datasets often reside in remote file or object stores. By enabling zero‑copy, RDMA‑based transfers, cuObject reduces and CPU overhead, potentially increasing throughput for data‑intensive models such as large language models or recommendation systems. The open‑source nature of the libraries and the shared wire protocol aim to standardize GPU‑driven storage access across cloud providers and on‑premise solutions, lowering integration effort for developers and fostering a broader ecosystem of compatible storage products. This could translate into faster model iteration cycles and lower operational costs for enterprises that rely on massive data pipelines.

AI workloads increasingly depend on rapid access to large datasets stored in remote object stores. Traditional storage paths involve copying data through the server’s CPU, incurring and consuming CPU cycles that could otherwise be allocated to model computation. cuObject’s RDMA‑based zero‑copy pathway directly addresses this inefficiency, offering higher throughput and lower latency for data‑intensive training and tasks.

Standardizing the wire protocol through xio‑sig and providing open‑source client/server libraries reduces the engineering burden for developers who previously had to write provider‑specific integrations. This interoperability can accelerate the deployment of AI pipelines across multiple cloud and on‑premise environments, fostering a more competitive storage market.

The broader Storage‑Next initiative, which includes over 40 vendors, aims to codify best practices for GPU‑driven storage access. By contributing to an industry‑wide standard, Nvidia positions itself as a key infrastructure provider, potentially influencing future hardware and software designs that prioritize high‑speed data movement.

Interactive Mechanism

Mécanisme interactif : comment cela fonctionne réellement

Explorez de manière interactive la technologie sous-jacente à ce développement.

Document Size:128K tokens
Needle Placement Depth (Location in document):50% into text
Attention Context Buffer Map:
Target Fact (50%)
Equivalent Pages~320Standard book pages
Retrieval Accuracy99.9%Needle recall score
RAM / KV Cache5.1 GBMemory overhead
Prompt CachingActive~80% discount on reuse
Core takeaway: Million-token context windows allow querying whole codebases or legal archives in one prompt. However, KV cache memory scales with context length, making prompt caching crucial for real-time production.
Vérification de concept interactive+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

Que regarder ensuite

Key indicators to monitor include adoption of cuObject by major cloud providers, the outcome of Nvidia’s conformance testing for the production‑ready stack, and further prototype demonstrations from storage vendors beyond IBM. The evolution of the xio‑sig governance board and the pace at which additional partners join will signal how quickly the industry moves toward a common standard. Finally, any pricing or licensing details released for the libraries or SDK will affect accessibility for startups and research institutions.

The rate at which cloud providers such as Google Cloud and Microsoft Azure adopt cuObject in their storage services will be a primary gauge of market impact.

Completion of Nvidia’s conformance testing and the public release of the cuObject repository will determine how quickly third‑party developers can integrate the libraries into their applications.

Additional prototype demonstrations from other storage vendors, as well as any announced pricing or licensing models for the SDK, will clarify the accessibility of the technology for startups and research labs.

Guides et quiz associés

Modèles d'IA expliquésFormation IAAvenir de l'IATestez ce que vous savez : essayez un quiz gratuit sur l'IARecherchez un terme d'IA dans notre glossaireSuivez le suivi des versions du modèle AI
Vous avez trouvé cela utile ?