In the last round of the Copilot LLM challenge Fable 5.1 came out on top. You can find that here:
https://blog.ciaops.com/2026/09/02/we-have-a-new-llm-in-copilot-winner/
That was then and this is now. Less than a month later a new model from OpenAI, GPT 6 Astra has become available in Copilot (Cowork specifically). I therefore pitted Astra 6 against the reigning champion Fable 5.1. The result is, unsurprisingly, that we have a new winner:
GPT 6 Astra
and you can see the results for yourself here:
https://github.com/directorcia/general/blob/master/Copilot/Comparisons/20260917-Cowork-Grok-eval.md
with all the outputs here:
https://github.com/directorcia/general/tree/master/Copilot/Comparisons/20260816
Interestingly, there was only a small improvement last time when Fable 5.1 pipped Opus 5 by 8.7 to 8.5 in the overall score. This time however Astra won by a significant margin of 9.03 to 8.30.
Again, not unsurprisingly, where Astra lost out was on cost, almost doubling the cost of Fable 5.1 as you can see:
Astra 6 = 6,067 credits
Fable 5.1 = 3,453 credits
Fable 5 (Preview) = 3,387 credits [Model no longer shown]
Fable 5 (Copilot)(Preview) = 3,373.8 credits [Model no longer shown]
GPT 5.6 Sol = 2,800.50 credits
Sonnet 5 = 2,509.3 credits
Opus 4.8 = 2,487 credits [Model no longer shown]
Opus 5 = 2,240 credits
GPT 5.5 = 1,200 credits
GPT 5.6 Terra = 260 credits
So the latest model is the best (unsurprising). The latest model is also the most expensive (unsurprising).
I have also updated the summary report for all the documents created by the different models here:
https://github.com/directorcia/general/blob/master/Copilot/Comparisons/20260816-File-paramters.md
I’m sure there will be more model releases coming soon but my rudimentary testing certainly indicates they are improving, if somewhat more expensive each time.
One thought on “Astra takes the trophy”