OpenAI has made headlines with the launch of its latest AI model, GPT-6 Sol, which has reportedly outperformed its rivals in key benchmark tests. The report highlights positive developments indicating that this development could significantly influence the choices businesses make regarding AI solutions.
Benchmark Test Results
In a recent series of benchmark tests, GPT-6 Sol demonstrated superior performance compared to Anthropic's Claude Opus 5. Notably, in the AutomationBench test, GPT-6 Sol achieved an impressive score of 332 at its highest reasoning setting, while Claude Opus 5 lagged behind with a score of 269, despite costing over 11 times more per task.
Performance in Agents Last Exam
Furthermore, GPT-6 Sol excelled in the Agents Last Exam, scoring 564, which is a substantial improvement over Claude Opus 5's best score. This performance, coupled with the lower cost of using GPT-6 Sol, positions OpenAI's model as a compelling option for businesses looking to leverage advanced AI technology.
Implications for the AI Landscape
These results not only underscore the capabilities of GPT-6 Sol but also suggest a potential shift in the competitive landscape of AI solutions.
On September 22, 2026, OpenAI launched two new models, GPT-6 Sol and Luna, which are designed to be cost-effective alternatives for everyday tasks, complementing the recently announced GPT-6 Sol. For more details, see read more.














