Quay lại Tin tức
sản phẩmAI Understanding tóm tắt

GIGAZINE báo cáo Z.ai phát hành GLM-5.3 dưới dạng mẫu trọng lượng mở

GIGAZINE báo cáo rằng công ty AI Trung Quốc Z.ai đã phát hành GLM-5.3 dưới dạng mẫu trọng lượng mở, cung cấp trọng lượng của nó thông qua Hugging Face và ModelScope sau các đánh giá an toàn và bảo mật đầy hứa hẹn trước đó.

6 min readRead the linked source
Source-provided image accompanying GIGAZINE reports Z.ai releases GLM-5.3 as an open-weight model
Nguồn tham khảoNguồn đã ghi
Nhà xuất bản
gigazine.net
Liên kết nguồn
gigazine.nethttps://gigazine.net/gsc_news/en/20260829-glm-5-3-open/
Loại nguồn
Nguồn được liên kết - trạng thái nguồn chính chưa được thiết lập.
Cũng được trích dẫn

Câu chuyện được sửa đổi lần cuối

Bối cảnhHiểu điều này trong 60 giây

Bắt đầu ở đây

Thuật ngữ chính

cân nặng
Một giá trị số đã học để chia tỷ lệ các tín hiệu truyền qua mạng nơ-ron.
API (Giao diện lập trình ứng dụng)
Một cách có cấu trúc để một hệ thống phần mềm gửi yêu cầu và nhận phản hồi từ hệ thống khác.
Bộ nhớ (Bộ nhớ tác nhân)
Bối cảnh được lưu trữ mà tác nhân AI sử dụng qua các bước hoặc phiên để cải thiện tính liên tục.
Tự kiểm traCâu đố giải thích về mô hình AI

Điều gì đã thay đổi kể từ khi xuất bản

  1. Xuất bản lần đầu
  2. GIGAZINE materially advances the existing GLM-5.3 release update by reporting the August 29 release timing, Hugging Face and ModelScope availability, the reported 744-billion-total/40-billion-active parameter configuration, the GLM-5.3 License security-review clause for organizations above $10 billion in annual revenue, quantized deployment claims, and Artificial Analysis index comparisons.

Chuyện gì đã xảy ra

GIGAZINE reports that Z.ai released GLM-5.3 as an open- model at 00:00 Japan time on August 29, 2026. The model is available to download, run and customize through Hugging Face and ModelScope. GIGAZINE says Z.ai had previously announced that it would open the model after completing security enhancements and safety evaluations.

GIGAZINE reports that Z.ai released GLM-5.3 as an open model at 00:00 Japan time on August 29. The report describes the release as the fulfillment of an earlier promise: GLM-5.3 began offering subscription and API services on August 14, and Z.ai had said the weights would become available after security enhancements and safety evaluations were completed. The report quotes Z.ai’s announcement that the model can be downloaded, run and customized. The timing and release announcement are reported by GIGAZINE; they are not independently confirmed here.

According to GIGAZINE, the weights are available through Hugging Face and ModelScope. The report says the model uses a mixture-of-experts architecture with 744 billion total parameters and 40 billion active parameters. Those figures describe the model’s overall and per-inference scale, but the source does not explain the exact architecture, training data, training compute, context length, or inference settings. GIGAZINE also reports that the GLM-5.3 License requires organizations with annual revenue above $10 billion to undergo a security review by Z.ai. The article does not provide the full license text or explain how the review would be administered or enforced.

GIGAZINE reports that the full model requires data-center-level AI infrastructure. It also says Unsloth has released a quantized version and that its UD-IQ2_M version is stated to run on a Mac with 256GB of unified memory or on a PC with 24GB of VRAM and 256GB of RAM. These are reported compatibility claims, not an independent test by GIGAZINE. The source does not specify performance losses from quantization, expected generation speed, power use, software requirements, or whether the configuration is practical for sustained workloads.

The article further reports that Artificial Analysis evaluated GLM-5.3 with an Intelligence Index score of 60 and an Agentic Index score of 59. GIGAZINE says the first score was above Claude Opus 4.8 and the second surpassed GPT-5.6 Sol and Grok 4.6. The report does not reproduce the underlying tests, sample sizes, confidence intervals, model settings, or complete comparison table. These comparisons should therefore be treated as reported benchmark results rather than independently established rankings. GIGAZINE also notes that Z.ai released the smaller GLM-5.3-Flash as an open model on August 26, linking it to the earlier identification of Ox Alpha.

Chi tiết nguồn: gigazine.net ↗

Tại sao nó quan trọng

The release gives developers access to a very large model that GIGAZINE describes as competitive with leading proprietary systems in some evaluations. That could expand experimentation with locally operated and customized AI, although the reported hardware requirements make the full model impractical for most individuals and smaller organizations.

Open- access changes who can inspect, adapt and deploy a model. Developers may be able to run GLM-5.3 outside a hosted API, customize it for particular applications, or evaluate it under conditions that are difficult to study with a closed service. That can improve research access and create another option for organizations concerned about provider dependence. The practical value depends on the completeness of the release, the license terms, the quality of tooling and the cost of operating the model.

The reported scale also puts limits on that openness. A 744-billion-parameter mixture-of-experts model is not equivalent to a lightweight local application simply because its weights are downloadable. GIGAZINE’s description of data-center-level infrastructure suggests that the full model will remain accessible mainly to well-funded developers and institutions. Quantized versions may broaden access, but the source does not independently establish their speed, quality, reliability or total operating cost. Hardware access, memory capacity and software compatibility will determine whether the release is useful beyond specialized users.

The performance claims matter because they frame the release as a potential competitor to proprietary frontier models, especially for coding and agentic tasks. If independently reproduced, such results could strengthen the case for open- systems in software development, automated workflows and security research. But a composite index is not a guarantee of reliable performance in real deployments. The source gives no evidence about factual accuracy, failure rates, cybersecurity behavior, refusal consistency, multilingual quality, privacy protections or performance on users’ own workloads.

The license clause introduces a less familiar constraint into an otherwise open- release. GIGAZINE reports that organizations above a specified revenue threshold must pass a security review by Z.ai. That could affect large commercial adopters’ procurement and compliance decisions, while leaving smaller users under different obligations. The source does not clarify whether the review is mandatory before use, what information applicants must provide, how long it takes, what standards apply, or what happens if an organization does not pass. Those unknowns are central to assessing how open the release is in practice.

Interactive Mechanism

Cơ chế tương tác: Nó thực sự hoạt động như thế nào

Khám phá công nghệ cơ bản đằng sau sự phát triển này một cách tương tác.

Thinking Budget (Test-Time Tokens):1,024 tokens
Complex Accuracy79%Math & Code Logic
Latency3.2sTime to first full output
Inference Cost$0.0092Per query estimated
Reasoning StyleStep VerificationInternal chain depth
Active Thinking Trace:
1Deconstruct user problem into formal constraints
2Propose candidate hypotheses & step-by-step calculation
3Self-correction: Backtrack and refute subtle edge cases
4Exhaustive consistency check & final output synthesis
Core takeaway: Test-time compute fundamentally changes AI economics. Instead of only scaling during pre-training, giving reasoning models more tokens at inference time allows them to systematically solve PhD-level STEM problems.
Kiểm tra khái niệm tương tác+10 Points
AI Models Explained Quiz

Which component of an AI application is the machine-learning model itself?

Xem gì tiếp theo

The main questions are whether the reported performance can be reproduced independently, how usable the license is in practice, and what safety evidence accompanied the release. GIGAZINE does not provide the underlying evaluation methodology, detailed safety findings, or independent confirmation of the model’s claimed capabilities.

First, watch for independent evaluations that publish task definitions, prompts, model settings, hardware configurations and error analysis. GIGAZINE’s account gives index scores and comparisons but not the evidence needed to determine whether GLM-5.3’s reported advantages are broad, statistically meaningful or sensitive to benchmark selection. Reproducible testing across coding, reasoning, tool use and safety tasks will be more informative than a single aggregate score.

Second, watch how developers handle the full model and the quantized release. Useful reporting should establish actual memory use, throughput, latency, power requirements, software support and quality changes after quantization. The source reports stated hardware configurations but does not independently confirm them. It also does not say whether the downloadable files include all components needed for deployment or whether users must obtain additional proprietary services.

Third, watch the implementation of the GLM-5.3 License. The reported security-review requirement for organizations with annual revenue above $10 billion could become a significant practical barrier or compliance obligation. The full terms, definitions, review process and enforcement mechanisms need to be examined before large organizations treat the model as a conventional open- dependency. Smaller users should also check whether other restrictions apply.

Finally, watch for safety documentation and incident reports. GIGAZINE says the release followed security enhancements and safety evaluations, but the article does not describe those evaluations or publish their results. Important unknowns include the tested misuse categories, red-team coverage, safeguards in the downloadable weights, handling of cyber-related capabilities, and the process for reporting vulnerabilities. Until that information is available, the release demonstrates access to a powerful model, not proof that it is safe or dependable for high-impact use.

Hướng dẫn và câu hỏi liên quan

Giải thích về mô hình AIChatGPT & LLMĐào tạo AIKiểm tra những gì bạn biết — thử một bài kiểm tra AI miễn phíTra cứu một thuật ngữ AI trong bảng thuật ngữ của chúng tôiTheo dõi trình theo dõi phát hành mô hình AI

Cập nhật và sửa chữa

Câu chuyện kinh điển này được cập nhật tại chỗ khi sự kiện đang phát triển có thay đổi cơ bản. URL và ngày xuất bản ban đầu của nó không bao giờ thay đổi.

  • GIGAZINE materially advances the existing GLM-5.3 release update by reporting the August 29 release timing, Hugging Face and ModelScope availability, the reported 744-billion-total/40-billion-active parameter configuration, the GLM-5.3 License security-review clause for organizations above $10 billion in annual revenue, quantized deployment claims, and Artificial Analysis index comparisons.
Xem nhật ký chỉnh sửa công khai
Tìm thấy điều này hữu ích?