01 What happened

Basis, a provider of AI agents for accounting, released findings from an internal performance test between GPT-6 Astra and GPT-5.6 Sol. According to Basis, the Astra model completed a 50-tab tax workbook in half the time of GPT-5.6 Sol.

02 Key details

  • Basis reports that Astra completed the tax workbook in 50% of the time required by GPT-5.6 Sol.
  • The firm claims a 20% improvement in internal evaluation scores by better identifying intent and flagging assumptions.
  • Basis adjusts reasoning effort throughout task execution, assigning higher computation to difficult steps.
  • Evaluations measure agent adherence to document templates and the accuracy of primary source references.

03 Why it matters

The case shows why reasoning effort can be adjusted during a long accounting task and why output quality needs a separate check. These performance figures come from Basis's own tests.

04 Who it matters to

Accounting professionals, AI tool developers and financial analysts.

Original sourceOpenAI