Astra shows a significant improvement over the GPT-5 family on long-array extraction, one of the hardest document extraction tasks.
here's what we're seeing:
> 77.7% bank, 72.8% legal, and 61.9% clinical mean per-document accuracy. this is +24–36 percentage points over GPT-5.6 Sol.
> Astra maintains high accuracy on larger arrays before performance drops. on clinical docs for example, astra maintains 100% accuracy through roughly 6,500 expected extracted values. Sol drops from 84% at roughly 3,500 values to 14% at 5,000.
> but it still struggles with the largest docs. Astra gets only 9–28% of expected values correct on the largest docs, compared with roughly 0–3% for earlier GPT-5 models.