Get the app →
BriefTea logoBriefTea

Every story in sixty words

General · 6 March 2026

GPT-5.4 scores 75% on computer benchmark

GPT-5.4 scores 75% on computer benchmark

OpenAI released its new GPT-5.4 model for ChatGPT, Codex, and developer API users. It starts rolling out today for Plus, Team, and Pro customers. The model natively controls computers, making it the first general-purpose OpenAI tool with this ability. It achieved a 75.0% score on the OSWorld-Verified benchmark, exceeding the 72.4% human baseline. It is 33% less likely to produce false factual claims than GPT-5.2, and financial spreadsheet modelling scores improved to 87.5% from 68.4%.

Reported by Tom's Guide · How we write briefs

More briefs

GeneralGoogle paying £260m to settle UK Play Store developer lawsuit PoliticsJames Cleverly steps down from the shadow cabinet to campaign for London mayor GeneralThe UK Navy just followed four Russian ships again GeneralSarah Ferguson is moving back to the UK following the Epstein scandal GeneralIcelanders vote on restarting EU membership talks after 13 years GeneralGoogle now auto-expands AI answers, pushing links down Browse all stories →