IA10 MIN

GPT-5.6 is here, and it wants to do the whole job—not just answer questions

OpenAI is launching a model built to research, use tools and complete long assignments. It could save hours, but tracing a mistake across all those steps will be harder.

Editorial illustration showing repeated OpenAI logos on a blue background
Image: Aïda Amer / Axios
01

OpenAI is launching three models, not one

OpenAI's GPT-5.6 family consists of flagship Sol, everyday model Terra and lower-cost Luna. The company says Sol improves coding, knowledge work, cybersecurity and science. Its new ultra setting coordinates multiple agents across parallel workstreams rather than merely spending longer on one answer.

That wording marks the real shift: OpenAI is selling end-to-end delegation, not just a smarter chatbot. The target is a usable document, application or analysis produced across several stages.

02

The benchmark claims need context

OpenAI reports 53.6 on Agents' Last Exam and 92.2% on BrowseComp for Sol, alongside gains in computer use, documents, spreadsheets and presentations. These are material claims, but they come from the vendor's launch. Independent testing must account for reasoning settings, price and the human effort required to review a complete job.

03

The opportunity is end-to-end work

Independent studios and small publications can research, prototype and localise at a depth that used to require larger departments. That can unlock abandoned ideas. It can also flood the web with interchangeable work if saved time is spent on volume instead of reporting, design and judgement.

04

The biggest risk is a polished wrong answer

Obvious mistakes are easy to spot. An incomplete report that looks finished is harder. Teams need original sources, decision logs and a named human owner. They should also avoid placing their entire operating process inside one vendor whose prices and limits can change.

Test GPT-5.6 on work you already understand and measure total time, including verification. The model proves its value through a deliverable you can defend—not through one impressive response. Our AI section will follow independent results and real-world failures.

SOURCES

Where the information comes from

Original announcements, documents and reporting used to prepare this article.

01OpenAI — anuncio y evaluaciones de GPT-5.602OpenAI — GPT-5.6 System Card03Artificial Analysis — índice independiente de modelos
00

The conversation starts here

Sign in with a supporter account to comment. Sign in

Nobody has commented yet. Want to go first?

KEEP READING

You may also like

FRONT PAGE