Kembali ke Berita
produkAI Understanding taklimat

OpenAI memperkenalkan GPT-6 Astra dengan pengekodan yang dilaporkan dan keuntungan penggunaan komputer

RTTNews melaporkan bahawa OpenAI telah memperkenalkan GPT-6 Astra, memetik keuntungan penanda aras utama, akses awal terhad, ketersediaan lebih luas yang dirancang dan harga API bermula pada $10 setiap juta token input.

4 min readRead the linked source
Source-provided image accompanying OpenAI introduces GPT-6 Astra with reported coding and computer-use gains
Rujukan sumberSumber direkodkan
Penerbit
rttnews.com
Pautan sumber
rttnews.comhttps://www.rttnews.com/amp/3688689/openai-introduces-gpt-6-astra-with-major-advances-in-ai-and-coding.aspx
Jenis sumber
Sumber terpaut — status sumber primer belum ditetapkan.
Juga dipetik

Cerita terakhir disemak

KonteksFahami perkara ini dalam masa 60 saat

Mulakan di sini

Istilah utama

API (Antara Muka Pengaturcaraan Aplikasi)
Cara berstruktur untuk satu sistem perisian menghantar permintaan dan menerima respons daripada sistem lain.
AGI (Kecerdasan Am Buatan)
Sistem AI hipotetikal yang boleh melaksanakan kebanyakan tugas intelektual pada peringkat manusia merentas banyak domain.
Penanda aras
Ujian piawai atau set data yang digunakan untuk mengukur dan membandingkan prestasi model.
Uji diri andaChatGPT & Kuiz LLMs

Apa yang berubah sejak penerbitan

  1. Pertama kali diterbitkan
  2. RTTNews materially advances the existing GPT-6 Astra launch update with additional reported benchmark scores, comparisons with other models, API pricing, planned distribution through Azure and Amazon Bedrock, and details about the model’s intended professional tasks and cybersecurity safeguards. These details remain attributed to RTTNews and OpenAI and are not independently confirmed in the supplied source.

Apa yang berlaku

RTTNews reports that OpenAI introduced GPT-6 Astra, describing it as the company’s most intelligent and aligned model to date. The report attributes results, capabilities, rollout plans and pricing to OpenAI; none of those claims are independently confirmed in the supplied source.

RTTNews reports that GPT-6 Astra is designed for computer use, software engineering, cybersecurity, scientific research and professional work. According to the report, OpenAI says Astra scored 98 percent on FrontierMath Tier 4, 99.9 percent on ARC-AGI-3 and 100 percent on ExploitBench. RTTNews also reports scores of 64.6 percent on Terminal-Bench Science 0.1 and 57.9 percent on Terminal-Bench 4.0, compared with results attributed to Claude Fable 5.1 and GPT-5.6 Sol.

The report says Astra scored 59.3 percent on Agents’ Last Exam, ahead of the cited results for Claude Opus 5 and GPT-5.6 Sol, and that OpenAI highlighted a 96 percent score on GPQA Diamond. RTTNews says OpenAI presented Astra as capable of creating websites, analyzing data, producing presentations and spreadsheets, testing software and completing online forms. The report also says the launch version includes safeguards against advanced offensive cybersecurity tasks.

Butiran sumber: rttnews.com ↗

Mengapa ia penting

If the reported results hold, Astra could materially affect how organizations use AI for software engineering, computer-based work, scientific research and cybersecurity. Its reported ability to complete multi-step tasks, combined with access through major cloud platforms, would make evaluation, oversight and cost comparisons important for prospective users.

The reported results span coding, scientific reasoning, computer-use and cybersecurity-related evaluations, making Astra’s significance broader than a routine model refresh. For organizations, the practical question is whether these claimed gains translate into reliable performance on real workflows such as software testing, data analysis and document production. The source provides no independent replication or methodology beyond the named tests and reported scores.

RTTNews reports that API pricing starts at $10 per million input tokens and $50 per million output tokens, with the tested configuration described as about 31 percent cheaper than the cited comparison on Terminal-Bench Science 0.1. If accurate, cost could influence model selection for high-volume workloads. The source does not establish total operating costs, performance outside the reported tests or whether the pricing applies to every Astra configuration.

Interactive Mechanism

Mekanisme Interaktif: Bagaimana Ia Berfungsi Sebenarnya

Terokai teknologi asas di sebalik pembangunan ini secara interaktif.

Thinking Budget (Test-Time Tokens):1,024 tokens
Complex Accuracy79%Math & Code Logic
Latency3.2sTime to first full output
Inference Cost$0.0092Per query estimated
Reasoning StyleStep VerificationInternal chain depth
Active Thinking Trace:
1Deconstruct user problem into formal constraints
2Propose candidate hypotheses & step-by-step calculation
3Self-correction: Backtrack and refute subtle edge cases
4Exhaustive consistency check & final output synthesis
Core takeaway: Test-time compute fundamentally changes AI economics. Instead of only scaling during pre-training, giving reasoning models more tokens at inference time allows them to systematically solve PhD-level STEM problems.
Semakan Konsep Interaktif+10 Points
ChatGPT & LLMs Quiz

What is a common training objective for an autoregressive language model?

Apa yang perlu ditonton seterusnya

The key questions are when broader access begins, which organizations receive the initial rollout, how the model performs in independent testing and how OpenAI limits high-risk cybersecurity use. The report does not provide a date for general availability or details of the launch safeguards.

RTTNews says Astra is initially rolling out to a limited group of organizations before becoming available to ChatGPT Plus, Pro, Business and Enterprise users. It also reports planned availability through the OpenAI API, Microsoft Azure and Amazon Bedrock. The report does not identify the initial organizations, specify regional restrictions, give a broader-release date or clarify whether all listed ChatGPT plans will receive identical capabilities.

Further scrutiny should focus on independent results, real-world reliability, failure rates in computer-use tasks and the scope of the cybersecurity safeguards. The source does not explain how offensive cyber capabilities are restricted, what monitoring is used or how users can verify the model’s claimed advantages.

Panduan & kuiz berkaitan

ChatGPT & LLMModel AI DiterangkanEjen AIUji apa yang anda tahu — cuba kuiz AI percumaCari istilah AI dalam glosari kamiIkuti penjejak keluaran model AI

Kemas kini dan pembetulan

Kisah kanonik ini dikemas kini apabila peristiwa yang sedang berkembang berubah secara material. URL dan tarikh penerbitan asalnya tidak pernah berubah.

  • RTTNews materially advances the existing GPT-6 Astra launch update with additional reported benchmark scores, comparisons with other models, API pricing, planned distribution through Azure and Amazon Bedrock, and details about the model’s intended professional tasks and cybersecurity safeguards. These details remain attributed to RTTNews and OpenAI and are not independently confirmed in the supplied source.
Lihat log pembetulan awam
Adakah ini berguna?