뉴스로 돌아가기
제품AI Understanding 브리핑

Ascend AI 칩용 DeepSeek 및 Huawei 오픈 소스 도구는 Nvidia 의존도를 줄이는 것을 목표로 합니다.

DeepSeek와 Huawei는 Nvidia의 CUDA 생태계를 피하려는 개발자를 대상으로 Huawei의 Ascend AI 프로세서를 위한 컴퓨팅 및 통신 라이브러리와 TileLang 지원을 포함한 오픈 소스 소프트웨어 스택을 출시했습니다.

4 min readRead the original reporting
Source-provided image accompanying DeepSeek and Huawei open‑source tools for Ascend AI chips aim to cut Nvidia reliance
기여 보고녹음된 소스
출판사
tomshardware.com
소스 링크
tomshardware.comhttps://www.tomshardware.com/tech-industry/artificial-intelligence/deepseek-and-huawei-release-open-source-ascend-ai-programming-tools-to-reduce-reliance-on-nvidia-ecosystem-tools-include-compute-and-communication-libraries-as-well-as-ascend-support-for-tilelang
소스 유형
자사 문서가 아닌 뉴스 매체를 통한 보도입니다.

자체적으로는 확인할 수 없었던 내용: 이 소유권 주장은 해당 매장에 귀속됩니다. 당사는 자사 문서와 비교하여 이를 확인하지 않았습니다. (tomshardware.com)

맥락60초 안에 이해하세요

여기서 시작하세요

주요 용어

대형 언어 모델(LLM)
텍스트를 생성하고 분석하기 위해 대규모 텍스트 말뭉치를 학습한 언어 모델입니다.
벤치마크
모델 성능을 측정하고 비교하는 데 사용되는 표준화된 테스트 또는 데이터 세트입니다.
추론
훈련된 모델이 예측 또는 출력을 생성하는 런타임 단계입니다.
자신을 테스트해 보세요AI 모델 설명 퀴즈

무슨 일이 일어났나요?

DeepSeek announced the release of a suite of open‑source programming tools for Huawei’s Ascend AI chips. The bundle includes libraries that handle AI‑specific computation and chip‑to‑chip communication, as well as support for the high‑level language TileLang on Ascend hardware. According to a Reuters report cited by Tom’s Hardware, the tools were co‑developed with Huawei, which provided full engineering support. The companies also optimized the stack for a super‑node configuration built around 128 Ascend 950 accelerators, addressing both efficient per‑chip calculation and fast inter‑chip data movement required for large‑scale AI workloads.

DeepSeek, a Chinese AI startup, released an open‑source software stack for Huawei’s Ascend AI processors. The stack comprises two primary libraries: one for AI‑specific computation kernels and another for high‑throughput chip‑to‑chip communication. Both libraries are intended to run efficiently on Ascend hardware without requiring Nvidia’s CUDA drivers.

In addition to the low‑level libraries, the release adds Ascend support for TileLang, a high‑level programming language designed to simplify AI model development. TileLang abstracts hardware details, allowing developers to write code that can be compiled for multiple accelerator architectures.

The collaboration between DeepSeek and Huawei also involved performance tuning for a super‑node system that links 128 Ascend 950 chips. The optimization focuses on two critical challenges for large AI models: maximizing per‑chip compute utilization and minimizing data‑transfer latency across the node.

The tools are hosted publicly, with source code and build instructions available for developers. DeepSeek states that Huawei provided full engineering support throughout development, but no pricing or commercial licensing details were disclosed. The release is positioned as a community‑driven effort to reduce reliance on Nvidia’s proprietary software stack.

소스 세부정보: tomshardware.com ↗

왜 중요한가요?

The release directly challenges the dominance of Nvidia’s CUDA ecosystem by giving developers a viable, open‑source alternative for high‑performance AI workloads on Ascend silicon. By lowering the software barrier, the stack could broaden the adoption of Huawei’s AI hardware, especially in regions or organizations that are seeking to diversify away from Nvidia for cost, geopolitical, or supply‑chain reasons. If the tools deliver the promised performance, they may shift market dynamics, encourage more competition in AI‑accelerator software, and spur further open‑source contributions to non‑CUDA ecosystems. However, the actual impact will depend on community uptake, documentation quality, and real‑world results, none of which have been independently verified yet.

Reducing dependence on Nvidia’s CUDA ecosystem can lower costs for organizations that currently pay licensing fees or face supply constraints for Nvidia GPUs. An open‑source alternative also mitigates geopolitical risks for companies operating in regions where Nvidia hardware may be restricted.

By providing a ready‑to‑use programming model (TileLang) and performance‑critical libraries, the stack lowers the technical barrier for developers to experiment with Ascend chips. This could accelerate the growth of a software ecosystem around Huawei’s hardware, which has historically lagged behind Nvidia’s extensive tooling.

If the stack delivers comparable or superior performance on large models, it may encourage cloud providers and enterprises to consider Ascend‑based offerings, diversifying the AI‑accelerator market. Such diversification can foster innovation, as hardware vendors compete not only on raw performance but also on the quality and openness of their software ecosystems.

The open‑source nature invites community contributions, potentially leading to rapid bug fixes, feature additions, and broader compatibility with AI frameworks like PyTorch or TensorFlow. However, the lack of independent data means the actual performance gains remain unverified.

Interactive Mechanism

대화형 메커니즘: 실제로 작동하는 방식

이 개발의 이면에 있는 기본 기술을 대화식으로 살펴보세요.

Model Parameter Size:8B Parameters
VRAM Required5.5 GBGPU memory footprint
Target HardwareMacBook / Single GPUDeployment tier
Privacy100% Air-GappedLocal device capability
Core takeaway: Small, quantized models (3B–8B) now run directly inside smartphones and laptops with complete data privacy, while mammoth 400B+ models remain the domain of datacenter clusters.
대화형 개념 확인+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

다음에 무엇을 볼 것인가

Key indicators to monitor include developer adoption rates, performance benchmarks comparing Ascend + TileLang against Nvidia CUDA on comparable models, and any subsequent updates from DeepSeek or Huawei expanding the toolset. Industry reactions—particularly from cloud providers and AI‑chip manufacturers—will reveal whether the stack can erode Nvidia’s market share. Watch for announcements of additional hardware support, integration with popular AI frameworks, and any licensing or support policies that could affect enterprise deployment.

Developer uptake: number of GitHub stars, forks, and contributions to the repository over the next few months.

releases: independent performance comparisons of Ascend + TileLang versus Nvidia CUDA on standard AI workloads (e.g., large language model ).

Enterprise announcements: cloud providers or AI service firms stating support for Ascend chips using the new stack.

Further tooling: any follow‑up releases from DeepSeek or Huawei that add support for additional frameworks, model formats, or hardware generations.

Policy and supply‑chain shifts: reactions from governments or industry groups that may influence hardware procurement decisions away from Nvidia.

관련 가이드 및 퀴즈

AI 모델 설명AI 트레이닝AI의 미래알고 있는 내용을 테스트해 보세요. 무료 AI 퀴즈를 시도해 보세요.용어집에서 AI 용어를 찾아보세요.AI 모델 출시 추적기를 따르세요.
이것이 유용하다고 생각하시나요?