GPT-5.4 Pro and Thinking are here!

Now I’m blown away by a “silly marketing" image…

Here is the image:

This is image is basic, yet perfect illustration of a complex concept coming from the software which is a by-product and handles daily routine tasks bigger competitors simply cannot handle.

Here’s how this image was done.

Short prompt to Codex, who updated the feature, created a sign off document, updated software documentation, created a draft for a blog post. Then I polished the blog post using custom gpts and workflows.

Start timer…

Then I gave the software documentation and blog post to gpt 5.4 who searched web for ideal customer profile users complaining on forums about the feature gaps in the competitor softwares, create a complete document with almost 70 quotes coming from real users and attached context and sources.

Then the document was parsed again by GPT 5.4 to generate LinkedIn posts which tie user quotes to the article about the software update. And in the same conversation I just asked it “now for each of the LinkedIn posts create me prompts for image AI generator to perfectly illustrate the posts on LinkedIn.

Then I copied the first prompt and pasted it into open AI images chat.

Then I got my first image out.

Stop timer…

Top chrono time (start/stop timer) is around 7 minutes… basically no human attention.

Those who did it at scale will understand.

Can you fix the freezing issue for windows users first?

You mean having to hit the refresh button type of freeze? Most of your responses are there on time, just hit refresh sooner for now.

It’s not related to the model; it’s been doing that for a while now.

5.4 seems like a pretty decent juggle of a whole lot of tensions…

Nice Job, Guys,

A bit late, but we can make an educated guess based on the announcement. If GPT-5.3-Codex is now part of the unified 5.4 model, I would not expect a separate 5.4-Codex release. It seems more likely that the next Codex-style update would appear as part of a future version such as 5.5. That said, this is only speculation about what might come next.

GPT‑5.4 brings together the best of our recent advances in reasoning, coding, and agentic workflows into a single frontier model. It incorporates the industry-leading coding capabilities of GPT‑5.3‑Codex⁠ while improving how the model works across tools, software environments, and professional tasks involving spreadsheets, presentations, and documents.

First of all, thank you to the team for continuously developing and improving the models. It is genuinely impressive to see how fast the technology evolves and how much effort clearly goes into making these systems more capable and useful.

I am genuinely interested in the new models and I’m glad to have the opportunity to try them. GPT-5.4, for example, is clearly powerful and in many ways even more refined than GPT-5.1 (thinking). Its writing can feel very vivid, structured, and “tasty” in terms of prose quality.

However, I would like to share some important feedback from a creative user perspective.

While GPT-5.4 is strong, it currently feels overly restricted due to its sensitivity to moderation rules. It often blocks, avoids, or softens responses even in contexts that are clearly part of legitimate storytelling. This breaks immersion and makes it difficult to fully develop scenes, characters, and emotional arcs.

For writing, especially character-driven narratives, this is a serious limitation.

One of the biggest issues is that the model sometimes feels “too safe” in tone. It avoids directness, avoids roughness, avoids emotional sharpness. But real dialogue, real characters, and real stories are not always polite or clean. They can be messy, intense, uncomfortable, and raw — and that is exactly what makes them meaningful.

Earlier models like GPT-4.1, GPT-4.5, and GPT-5.1 handled this much better. They allowed for more natural expression, more direct language, and more emotional depth.

GPT-5.4 writes well — sometimes exceptionally well — but it often feels like it is holding itself back.

As a writer, I want a model that can fully engage with the entire spectrum of storytelling. That includes:
– emotionally intense scenes
– morally complex or uncomfortable situations
– physicality, attraction, and human closeness as part of character development
– conflict, including situations where boundaries are crossed (not to justify them, but to explore and reflect on them)

Stories are not meant to be sanitized. They are meant to make people think, feel, and sometimes confront difficult realities. Avoiding these aspects limits the model’s usefulness for serious creative work.

Because of this, I believe there are two possible directions to consider:

  1. Reduce the model’s sensitivity (for example, allowing a more direct and flexible response mode similar to earlier models).

  2. Revisit and adjust the user policy framework so that it better reflects real-world creative use cases, including mature and complex themes.

Right now, the balance leans too far toward restriction at the cost of expressiveness and usability.

I also want to emphasize that earlier models — GPT-4.1, GPT-4.5, and GPT-5.1 — each had unique strengths that are still not fully replicated:
– GPT-5.1 had exceptional reasoning and character understanding
– GPT-4.1 was extremely efficient and practical, while allowing more natural discussion
– GPT-4.5 produced some of the most engaging and well-balanced prose I have seen

Losing access to these models would mean losing valuable tools for creative work.

I would strongly encourage keeping them available in some form (legacy mode, optional selection, or API access), while also improving the balance in newer models like GPT-5.4.

Thank you for your work and for listening to user feedback.

Nice job team and keep it rolling!