OpenAI's next-generation image model, tentatively dubbed GPT-Image-2, appears to be running blind tests on Chatbot Arena. Blake Robbins flagged the speculation on X, noting that users are encountering anonymous models producing outputs that are photorealistic to a degree that makes GPT-Image-1 look like a rough draft. Three specific model codenames have surfaced in the testing pool:
maskingtape-alpha
gaffertape-alpha
packingtape-alpha
Multiple users have described the circulating sample images from these models as a massive leap in quality over existing generators.
What Chatbot Arena Tells Us
For anyone unfamiliar with the platform: Chatbot Arena runs blind A/B comparisons where users submit prompts and vote on which output is better without knowing which model produced it. It has become one of the more reliable signals in AI benchmarking precisely because it removes marketing from the equation. Users judge outputs on merit alone.



