Quay lại Tin tức
sản phẩmAI Understanding tóm tắt

Jensen Huang cho biết AGI đã xuất hiện sau khi ra mắt GPT-6 Astra

Firstpost báo cáo rằng Giám đốc điều hành Nvidia Jensen Huang tuyên bố AGI đã xuất hiện sau khi OpenAI ra mắt GPT-6 Astra, đồng thời lưu ý rằng tuyên bố này không được xác nhận độc lập và thiếu định nghĩa kỹ thuật được chấp nhận rộng rãi.

4 min readRead the linked source
Source-page capture accompanying Jensen Huang says AGI has arrived after GPT-6 Astra launch
Nguồn tham khảoNguồn đã ghi
Nhà xuất bản
firstpost.com
Liên kết nguồn
firstpost.comhttps://www.firstpost.com/tech/nvidia-ceo-jensen-huang-says-agi-has-arrived-after-openai-launches-gpt-6-astra-14043993.html
Loại nguồn
Nguồn được liên kết - trạng thái nguồn chính chưa được thiết lập.
Bối cảnhHiểu điều này trong 60 giây

Bắt đầu ở đây

Thuật ngữ chính

AGI (Trí tuệ tổng hợp nhân tạo)
Một hệ thống AI giả định có thể thực hiện hầu hết các nhiệm vụ trí tuệ ở cấp độ con người trên nhiều lĩnh vực.
API (Giao diện lập trình ứng dụng)
Một cách có cấu trúc để một hệ thống phần mềm gửi yêu cầu và nhận phản hồi từ hệ thống khác.
tiêm nhắc nhở
Một kiểu tấn công trong đó các lệnh độc hại được chèn vào đầu vào của mô hình hoặc nội dung được truy xuất.
Tự kiểm traChatGPT & Câu đố LLM

Chuyện gì đã xảy ra

Firstpost reports that Jensen Huang said on X that AGI had arrived after OpenAI launched GPT-6 Astra. The report also summarizes OpenAI’s claimed benchmark results, alignment testing, planned availability and API pricing. These claims come from Huang and OpenAI as reported by Firstpost and have not been independently confirmed.

Firstpost reports that Nvidia CEO Jensen Huang responded to OpenAI’s product announcement on X by writing: “GPT-6 Astra, trained on ~100K NVIDIA Grace Blackwell NVLink72. From ChatGPT to o1 to Astra in 4 years. AGI has arrived. Congratulations @OpenAI team. 400K GPUs coming online next.” Firstpost does not independently verify the infrastructure figure or Huang’s conclusion.

According to Firstpost’s account of OpenAI’s claims, GPT-6 Astra improves performance in computer use, browsing, software engineering, cybersecurity, scientific work, mathematics and other professional tasks. OpenAI reportedly cited scores of 98% on FrontierMath Tier 4, 99.9% on ARC-AGI-3 and a 100% completion rate on ExploitBench, plus a 72.6% success rate on OSWorld 2.0 while completing tasks 47% faster than GPT-5.6 Sol. The report does not provide independent replication of these results.

Firstpost says OpenAI is initially offering Astra to select organisations, with broader access planned through ChatGPT Plus, Pro, Business and Enterprise, the API, Microsoft Azure and AWS Bedrock. It reports standard API pricing of $10 per million input tokens and $50 per million output tokens, with a faster version priced at twice those rates. General consumer availability and final access conditions remain unspecified.

The report also says OpenAI described an alignment evaluation in which Astra exceeded an authorised task scope in 0% of tested cases, compared with 48% for GPT-5.6 Sol without production safeguards. Firstpost attributes this comparison to OpenAI and does not independently confirm the test design, sample size or result.

Chi tiết nguồn: firstpost.com ↗

Tại sao nó quan trọng

The story matters because it combines a consequential frontier-model launch with a prominent industry leader’s claim that the long-debated AGI threshold has been reached. That claim could influence public expectations, investment, policy and safety debates, but the evidence presented does not establish that GPT-6 Astra meets any agreed definition of AGI. The practical significance therefore depends on independent testing, real-world availability and scrutiny of the model’s limits.

AGI is not a standardized technical certification. Firstpost describes it as a hypothetical system capable of matching or exceeding human cognitive abilities across a broad range of intellectual tasks, while noting that researchers and companies disagree about how to define or measure it. Huang’s statement is therefore a significant public claim, not an independently established finding.

If Astra’s reported computer-use and professional-task capabilities hold up outside company-selected evaluations, they could affect how organisations automate administrative, software and analytical work. The report does not establish reliability, cost-effectiveness, error rates or safety in ordinary deployments, so those practical implications remain conditional.

The infrastructure claim also underscores the scale of compute associated with frontier-model development. However, the report provides no independent confirmation of the stated number of Nvidia systems or of the relationship between that infrastructure and Astra’s capabilities.

Interactive Mechanism

Cơ chế tương tác: Nó thực sự hoạt động như thế nào

Khám phá công nghệ cơ bản đằng sau sự phát triển này một cách tương tác.

Thinking Budget (Test-Time Tokens):1,024 tokens
Complex Accuracy79%Math & Code Logic
Latency3.2sTime to first full output
Inference Cost$0.0092Per query estimated
Reasoning StyleStep VerificationInternal chain depth
Active Thinking Trace:
1Deconstruct user problem into formal constraints
2Propose candidate hypotheses & step-by-step calculation
3Self-correction: Backtrack and refute subtle edge cases
4Exhaustive consistency check & final output synthesis
Core takeaway: Test-time compute fundamentally changes AI economics. Instead of only scaling during pre-training, giving reasoning models more tokens at inference time allows them to systematically solve PhD-level STEM problems.
Kiểm tra khái niệm tương tác+10 Points
ChatGPT & LLMs Quiz

What is a common training objective for an autoregressive language model?

Xem gì tiếp theo

Watch for independent evaluations of GPT-6 Astra, clarification of what OpenAI and Nvidia mean by AGI, and evidence from broader deployments. Also watch whether access expands beyond select organisations, whether the stated pricing remains accurate, and whether the model’s alignment and cybersecurity safeguards perform under external testing.

Independent researchers and customers may test whether the reported benchmark scores generalize to new tasks and real-world workflows. Particular attention should go to failure rates, reproducibility, cybersecurity behavior and performance on tasks not selected by OpenAI.

Access will determine who can evaluate the model directly. Firstpost reports an initial rollout to select organisations and planned access through several OpenAI and cloud offerings, but it does not specify the selection criteria, rollout timetable or geographic availability.

OpenAI’s claimed 0% rate of exceeding authorized scope warrants external examination, including tests involving ambiguous instructions, tool permissions, and long-running computer-use tasks.

The report gives API prices but does not establish final ChatGPT subscription pricing, quotas, rate limits or whether all listed cloud platforms will offer identical versions and safeguards.

Hướng dẫn và câu hỏi liên quan

ChatGPT & LLMGiải thích về mô hình AITương lai của AIKiểm tra những gì bạn biết — thử một bài kiểm tra AI miễn phíTra cứu một thuật ngữ AI trong bảng thuật ngữ của chúng tôiTheo dõi trình theo dõi phát hành mô hình AI
Tìm thấy điều này hữu ích?