← Back to KHAO

OpenAI · GPT · Sam Altman ·

OpenAI quietly boosts some of Astra’s evaluation metrics, and continues to change others post-launch

2 min read

Compiled by KHAO Editorial — aggregated from 1 source. See llms.txt for citation guidance.

◌ Single Source

Emily Forlini.

OpenAI has changed several evaluation benchmarks for its GPT-6 Astra model since first publishing a blog post announcement mid-afternoon on Sept. 3.

Key facts

Summary

The changes occurred amid an unusual rollout of the blog post. When OpenAI’s X account tweeted out the blog post at 3:32 p.m., the link was not loading properly, returning an error message. It turns out OpenaAI published the blog shortly after 2pm but retracted it for reason the company said it could not disclose, but which it said were unrelated to the benchmark performance figures. (OpenAI first told them it was a bug in the content management system, and then an internet outage.) Upon republishing the blog, it had different evaluation metrics that seemed to favor Astra—and some figures have continued to change even since then.

Read full article at Fortune Technology →

#OpenAI #GPT #Sam Altman