Summary

A hands-on test of OpenAI's GPT Image 2.5 by Theoretically Media, run through the reviewer's standing set of comparison prompts. His verdict is in the title of the update rather than the video: it is a point-five release, and the image quality gains are modest. The finding worth keeping is elsewhere — with GPT-6 Astra behind it, the image model can now go and research before it draws, and the reviewer's working method shifts from writing better prompts to pointing it at references and arguing with the result.

Why it matters

The image model is the less interesting half of this. What the entry records is a change in how these tools are used: an image generator with a reasoning model attached stops being a prompt box and becomes something you brief. It looks up who you are, finds the actors, studies a film's colour grade, and notices a compositional mistake you did not mention. Prompt-craft as a skill is being displaced by reference-giving and iteration.

It also gives the Handbook an independent, hands-on read on GPT-6 Astra from someone with no stake in the benchmark argument running through DK-91 and DK-93. His verdict — first model in a long time to genuinely surprise him, while explicitly declining to 'drink the AGI Kool-Aid' — is a more useful data point than either the panel's enthusiasm or the leaderboards.

The reasoning-versus-diffusion distinction in the wine-glass test is the durable technical point and belongs in the Handbook's content-tools section: it explains why these models obey awkward instructions and why their output looks duller.

VsmL_KROmyI-transcript.txt