OpenAI released GPT-5.4, the first of its general-purpose models to have computer use built in — reading the screen as images and driving mouse and keyboard. It scored 75.0% on OSWorld-Verified, a benchmark of desktop tasks, well up from GPT-5.2's 47.3%. OpenAI described this as "surpassing human performance at 72.4%"; that reference figure comes from the OSWorld paper, which had people attempt its 369 tasks, and a specialised computer-use agent had already passed it the previous December. GPT-5.4 scored 83% on GDPval, which measures realistic knowledge work. The API context window reached one million tokens, the company's longest, and a new "Tool Search" cut the tokens spent on tool calling. OpenAI said individual factual claims were wrong 33% less often than in GPT-5.2, and whole responses 18% less often. It shipped in ChatGPT as GPT-5.4 Thinking, with a GPT-5.4 Pro for more compute on hard problems. On the 17th, smaller GPT-5.4 mini and GPT-5.4 nano followed; mini reached free and Go accounts as well as the API and Codex, while nano was API-only.