Quay lại Tin tức
sản phẩmAI Understanding tóm tắt

Cognition phát hành giải pháp khai thác Fusion để mang lại hiệu suất cho tác nhân AI với chi phí thấp hơn

GIGAZINE báo cáo rằng Cognition đã phát hành Fusion, một bộ khai thác dành cho GPT-6 Astra và Claude Fable phù hợp với các hệ thống đại lý cạnh tranh về điểm chuẩn đồng thời giảm chi phí được báo cáo lên tới 39%.

4 min readRead the linked source
Source-provided image accompanying Cognition releases Fusion harness for lower-cost AI agent performance
Nguồn tham khảoNguồn đã ghi
Nhà xuất bản
gigazine.net
Liên kết nguồn
gigazine.nethttps://gigazine.net/gsc_news/en/20260914-cognition-fusion/
Loại nguồn
Nguồn được liên kết - trạng thái nguồn chính chưa được thiết lập.
Bối cảnhHiểu điều này trong 60 giây

Bắt đầu ở đây

Thuật ngữ chính

Đặc vụ AI
Một hệ thống phần mềm có thể quan sát, suy luận và thực hiện các hành động để đạt được mục tiêu, thường sử dụng các công cụ và bộ nhớ.
Điểm chuẩn
Một bài kiểm tra hoặc tập dữ liệu được tiêu chuẩn hóa dùng để đo lường và so sánh hiệu suất của mô hình.
suy luận
Giai đoạn chạy trong đó mô hình được đào tạo tạo ra dự đoán hoặc kết quả đầu ra.
Tự kiểm traCâu đố về đại lý AI

Chuyện gì đã xảy ra

GIGAZINE reports that Cognition, the company behind Devin, released Fusion, a harness designed to improve how advanced AI models plan, execute, review, and recover from stalled tasks. The system pairs a high-end model for planning and review with a less expensive model for execution, while running two agents in parallel with separate contexts and tools.

GIGAZINE reports that Cognition developed and released Fusion as a harness for OpenAI’s GPT-6 Astra and Anthropic’s Claude Fable. The article describes a harness as the layer that supplies models with operating instructions, tool-use rules, and conditional branching, potentially affecting whether an agent completes a task or becomes stuck in a loop.

According to GIGAZINE, Fusion combines a state-of-the-art model for planning and review with a cost-effective model for execution. Cognition recommended GPT-6 Astra and Claude Fable as the advanced models, and identified Cognition’s SWE-2 coding model as the cheaper execution option. The article says the combination performed best with Claude Fable 5.1.

GIGAZINE reports a score of 62.2 for Claude Fable 5.1 with Claude Code, compared with 61.7 for Claude Fable 5.1 combined with the lower-cost model and Fusion. The latter combination was reported as 36% cheaper. For GPT-6 Astra, the article says Fusion achieved performance equivalent to Codex while reducing costs by 39%.

The article attributes evaluations to AI analytics companies Artificial Analysis and Vals AI. It says Fusion runs two agents in parallel, each with independent context and tools, and exchanges only summaries, results, and feedback rather than the full conversation. GIGAZINE says this approach makes greater use of prompt caching. The source does not provide the full methodology, absolute prices, or availability terms.

Chi tiết nguồn: gigazine.net ↗

Tại sao nó quan trọng

If the reported results hold, Fusion could reduce the cost of using frontier AI agents without a comparable loss in coding performance. That matters for developers and organizations running tool-using systems at scale, where model and repeated agent attempts can become major expenses. The results are reported by GIGAZINE and evaluated by Artificial Analysis and Vals AI, but the source does not provide enough methodology to independently assess the comparisons.

The reported results suggest that agent architecture can materially affect the economics of frontier models. A system that delegates execution to a cheaper model while retaining a stronger model for planning and review could lower the cost of software agents without requiring users to abandon higher-capability models.

Cost reductions are particularly relevant for coding agents, which may make many tool calls, repeat failed attempts, or maintain long contexts. If independently reproducible, the reported 36% and 39% reductions could influence how companies design production agent workflows and choose between proprietary harnesses.

The evidence remains limited in the supplied report. GIGAZINE provides selected scores and percentage savings but not the task set, run count, pricing assumptions, latency, failure rates, or comparison conditions. The claimed equivalence to Codex and the stated efficiency gains are therefore not independently confirmed here.

Interactive Mechanism

Cơ chế tương tác: Nó thực sự hoạt động như thế nào

Khám phá công nghệ cơ bản đằng sau sự phát triển này một cách tương tác.

Agent Lifecycle Stage:
1
User Intent & Planning: "Audit customer refund request #4092 and settle payment."
2
Tool Calling: Emits structured JSON call crm_get_transaction(id='4092').
3
Guardrail & Verification:🛡️ Paused: High-value action requires human operator sign-off.
4
Final Settlement: Refund recorded, email receipt dispatched, and audit log stored.
Core takeaway: An AI agent is not just a language model—it is a closed loop of planning, tool invocation, and environment feedback. Production systems require self-healing retries and strict human approval guardrails.
Kiểm tra khái niệm tương tác+10 Points
AI Agents Quiz

An agent must create a draft calendar event for Tuesday at 2 p.m. Which evidence would establish the requested result?

Xem gì tiếp theo

The key questions are who can access Fusion, which models and tools it supports, whether it is generally available, and how its costs are calculated. Further reporting should clarify the tasks, baselines, sample sizes, prompt-cache assumptions, and whether the results transfer beyond coding. The source also uses both “SWE-2” and “SWR-2” for the lower-cost model, an inconsistency that should be resolved.

Access conditions are not documented in the source. It is unknown whether Fusion is publicly downloadable, available through Cognition’s products, restricted to selected users, or offered through another commercial arrangement. Pricing and any required model subscriptions are also unknown.

Follow-up reporting should establish whether Fusion supports models beyond GPT-6 Astra and Claude Fable, whether users can supply their own tools and prompts, and how much of the savings depends on prompt caching or specific provider pricing.

The naming contains a source-level inconsistency: the article identifies the cheaper coding model as SWE-2 but later refers to “SWR-2.” That designation should be verified before treating the comparison as a precise product claim.

Independent testing should examine coding success rates, latency, reliability, context handling, and performance on tasks other than the reported . Parallel agents may also increase orchestration complexity or total tool activity even when model costs fall.

Hướng dẫn và câu hỏi liên quan

Đại lý AIGiải thích về mô hình AIPrompt EngineeringKiểm tra những gì bạn biết — thử một bài kiểm tra AI miễn phíTra cứu một thuật ngữ AI trong bảng thuật ngữ của chúng tôiTheo dõi trình theo dõi phát hành mô hình AI
Tìm thấy điều này hữu ích?