Wireva

OpenAI adjusts Astra benchmark figures after launch blog post

OpenAI changed several evaluation metrics for its GPT-6 Astra model after publishing its launch announcement, with some scores favoring Astra and others fluctuating for rival models. The company says the adjustments reflect fixes to ensure accurate comparisons.

Monitoring item. The full text is not distributed. Extract and source below.

OpenAI changed several evaluation metrics for its GPT-6 Astra model after publishing its launch announcement, with some scores favoring Astra and others fluctuating for rival models. The company says the adjustments reflect fixes to ensure accurate comparisons.

Same event, other desks

Story file →