Skip to content
GOPPO

News · AI summarised to understand what matters

Back to news

Models

Published on

OpenAI Launches GPT-5.3 Instant: Fewer Refusals, Less "Cringe," 26.8% Fewer Hallucinations

OpenAI updated ChatGPT's most-used model with a focus on more natural tone, fewer unnecessary refusals, and more reliable answers — especially when using web search.

  • openai, gpt-5, chatgpt, alucinacoes, tom-conversacional

Summary

OpenAI launched GPT-5.3 Instant on March 3, 2026, as an update to the most widely used model in ChatGPT. Unlike previous releases centered on advanced reasoning capabilities, this update focuses on the everyday experience: tone, conversational flow, and reliability. The model is immediately available to all ChatGPT users and to developers via the API under the identifier `gpt-5.3-chat-latest`. GPT-5.2 Instant remains available in the legacy models section for paid users until June 3, 2026, when it will be retired. Updates to the Thinking and Pro modes are described as coming soon.

In practice

OpenAI identified three main areas of improvement in GPT-5.3 Instant.

The first is the reduction of unnecessary refusals and overly cautious tone. According to the company, GPT-5.2 Instant would frequently refuse questions it could safely answer, or precede responses with lengthy disclaimers about what it couldn't do. GPT-5.3 Instant moves directly to the useful answer without defensive or moralizing preambles. OpenAI gave a concrete example: when asked about trajectory calculations for long-range archery, the previous model opened with a paragraph explaining what it couldn't help with; the new model goes straight to what it can do.

The second improvement is tone. The company acknowledged that GPT-5.2 Instant could come across as "cringe" — overly dramatic or condescending, with phrasing like "Stop. Take a breath." GPT-5.3 Instant has a more focused and natural conversational style, cutting back on unnecessary proclamations and assumptions about the user's emotional state. OpenAI also said it is working to keep ChatGPT's personality more consistent across updates.

The third area is response quality when using web search. The model has improved how it balances what it finds online with its own knowledge and reasoning. Rather than returning lists of links or mechanically summarizing search results, GPT-5.3 Instant uses existing understanding to contextualize recent information — better identifying the intent behind a question and surfacing the most relevant information upfront.

On factual reliability, OpenAI conducted two internal evaluations. The first, focused on higher-stakes domains — medicine, law, and finance — showed the model reduced hallucinations by 26.8% when using web search and by 19.7% when relying only on internal knowledge, compared to prior models. The second evaluation, based on real ChatGPT conversations flagged by users for factual errors, showed hallucinations decreased by 22.5% with web use and 9.6% without.

The context window was expanded from 128,000 to 400,000 tokens.

OpenAI also published the model's safety card, acknowledging that GPT-5.3 Instant shows regressions compared to GPT-5.2 Instant in some safety categories — specifically disallowed sexual content and self-harm content. The company noted these results may change after launch and that the regressions for graphic violence and violent illicit behavior have low statistical significance.

Within hours of the GPT-5.3 Instant announcement, OpenAI posted on X: "5.4 sooner than you think," suggesting an accelerating iteration cycle. Subsequent reports confirmed the GPT-5.4 launch in the days that followed.

What we still don't know

GPT-5.3 Instant is described by OpenAI as a product improvement rather than a capability advance. The company did not disclose details about the model's architecture, the alignment techniques used to reduce refusals, or the impact on inference cost. The expanded 400,000-token context window was not accompanied by specific benchmarks for long-context task performance.

The safety regressions acknowledged by OpenAI raise questions about the relationship between reducing refusals and maintaining safety guardrails — a tension the company does not address in detail.

Why it matters

  • The bet on a "less annoying" model signals a shift in OpenAI's priorities: perceived quality in everyday interaction has become as relevant to user retention as technical benchmark results.
  • The 26.8% reduction in hallucinations when using web search is meaningful for enterprise use cases in medicine, law, and finance, where factual reliability has practical consequences.
  • Expanding the context window to 400,000 tokens — more than triple the previous version — opens new possibilities for long-document analysis and complex workflows.
  • Admitting safety regressions in the model's own safety card sets a transparency precedent, but also raises the question of whether reducing refusals may compromise existing safety guardrails.