Basis completes a tax workbook 2x faster with GPT-6 Astra
Basis, a company that builds AI agents to automate accounting tasks, utilized GPT-6 Astra to complete a complex tax workbook with 50 tabs. Compared to GPT-5.6 Sol, GPT-6 Astra performed the task twice as fast and demonstrated a stronger understanding of accounting objectives. This improved speed and reliability reduces the need for Basis to create specific rules for individual situations, enhancing confidence in their agents' ability to handle diverse scenarios beyond internal testing.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
发布当时偏移:UTC+02026年9月28日 00:00 UTC
收录当时偏移:UTC+02026年9月28日 22:00 UTC
- 发布
- 2026年9月28日 00:00
- 收录
- 2026年9月28日 22:00
- 来源类型
- 官方发布
- 档位
- 当事方
- 信源状态
- 正常
档位是按信源手工设定的编辑判断,不是逐条打分。
讨论趋势
百分比基于采集到的讨论信号,不代表新增评论数或独立参与人数。曲线仅用于同一话题在不同时段的比较。
Basis (opens in a new window) builds AI agents to automate much of the manual work that accountants do each day, helping them shift their time from repetitive tasks to strategic work. The company’s research focuses on agents that can reliably complete long tasks, and with GPT‑6 Astra, it’s seeing a stronger understanding of what accountants want to accomplish.
“GPT-6 Astra does a better job of really understanding the intent of the user and the problem.”
—Mitch Troyanovsky, Co-founder, Basis
Completing a 50-tab tax workbook in half the time
Basis compared GPT‑6 Astra and GPT‑5.6 Sol on a complicated tax workbook with 50 tabs. The task was to complete the workbook accurately and reliably, and GPT‑6 Astra was markedly faster.
“GPT-6 Astra is able to complete that workbook in half the time that GPT-5.6 Sol is able to.”
Basis also noted that GPT‑6 Astra makes better decisions at the start of a task, helping Basis’s agents take a more direct path through the work with less time spent correcting mistakes. Troyanovsky says that also makes the model more efficient in its use of tokens.
Matching reasoning effort to the task
Basis also has GPT‑6 Astra adjust how much reasoning it uses as a task progresses, dialing up computation when a step is difficult, and using less when a step is easier. The model can make these adjustments while keeping its cache intact. Troyanovsky says this helps reduce cost and response time, making long-running tasks more economical for Basis and its customers.
Building confidence in real-world use
Basis saw about a 20% improvement in its internal evaluation scores with GPT‑6 Astra, driven by better understanding of user intent, including when to ask questions, flag assumptions, and follow instructions.
Basis evaluates how its agents work and their final answers, including whether they follow templates, consult primary sources for tax questions, and check their own work. GPT‑6 Astra can infer these expectations from a broader context, with fewer explicit instructions.
That reduces the need for Basis to write rules for individual situations and gives the team more confidence that its agents can handle situations beyond those covered in internal tests.