跳到主要內容
AI News HubLIVE
站內改寫2 分鐘閱讀

待翻譯:Basis completes a tax workbook 2x faster with GPT-6 Astra

文章摘要

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:GPT-6 Astra completed a 50-tab tax workbook twice as fast as GPT-5.6 Sol, and its stronger understanding of user intent gives Basis more confidence in real-world use.

待翻譯:Basis completes a tax workbook 2x faster with GPT-6 Astra
回報錯誤

更正管道尚未開通,可先複製下方文章資訊留存。

查看更正說明
直接讀正文

AI 服務暫時不可用,以下為來源正文,待恢復後補全翻譯。

OpenAI September 28, 2026 Basis completes a tax workbook 2x faster with GPT‑6 Astra GPT‑6 Astra took 50% less time to finish complex tasks vs. GPT‑5.6 Sol and showed a deeper understanding of user intent. Start building with OpenAI Company size: Startup Region: North America Industry: Technology Products: API Workbook test 50% Less time to complete a 50-tab tax workbook vs. Sol in Basis’s test. Internal evaluations 20% Approximate improvement in Basis’s internal evaluation scores. Loading… Basis⁠(opens in a new window) builds AI agents to automate much of the manual work that accountants do each day, helping them shift their time from repetitive tasks to strategic work. The company’s research focuses on agents that can reliably complete long tasks, and with GPT‑6 Astra, it’s seeing a stronger understanding of what accountants want to accomplish. “GPT-6 Astra does a better job of really understanding the intent of the user and the problem.” —Mitch Troyanovsky, Co-founder, Basis Basis compared GPT‑6 Astra and GPT‑5.6 Sol on a complicated tax workbook with 50 tabs. The task was to complete the workbook accurately and reliably, and GPT‑6 Astra was markedly faster. “GPT-6 Astra is able to complete that workbook in half the time that GPT-5.6 Sol is able to.” —Mitch Troyanovsky, Co-founder, Basis Basis also noted that GPT‑6 Astra makes better decisions at the start of a task, helping Basis’s agents take a more direct path through the work with less time spent correcting mistakes. Troyanovsky says that also makes the model more efficient in its use of tokens. Basis also has GPT‑6 Astra adjust how much reasoning it uses as a task progresses, dialing up computation when a step is difficult, and using less when a step is easier. The model can make these adjustments while keeping its cache intact. Troyanovsky says this helps reduce cost and response time, making long-running tasks more economical for Basis and its customers. Basis saw about a 20% improvement in its internal evaluation scores with GPT‑6 Astra, driven by better understanding of user intent, including when to ask questions, flag assumptions, and follow instructions. Basis evaluates how its agents work and their final answers, including whether they follow templates, consult primary sources for tax questions, and check their own work. GPT‑6 Astra can infer these expectations from a broader context, with fewer explicit instructions. That reduces the need for Basis to write rules for individual situations and gives the team more confidence that its agents can handle situations beyond those covered in internal tests. OpenAI <3 startups Join the communityStart building(opens in a new window) Keep reading Lenfest grows landmark program with OpenAI support CompanySep 28, 2026 Are you a Codex Original? Sep 28, 2026 Proaction boosts sales 60% and saves 75+ hours with Codex Sep 25, 2026

展開要點與分析

文章情報

工程師進階

要點

  • AI 服務暫時不可用,系統已先保留來源內容與降級後設資料。
  • GPT-6 Astra completed a 50-tab tax workbook twice as fast as GPT-5.6 Sol, and its stronger understanding of user intent gives Basis more confidence in real-world use.

技術影響

可能影響 Agent 架構、工具呼叫、工作流自動化和產品整合。

要點與分析由自動化流程生成,可能有誤,請結合原始來源核實。