ChatGPT Images 2.5
Posted by vertigoruntime 3 hours ago
Comments
Comment by conradludgate 3 hours ago
Most depressing thing I have read today
Comment by bluegatty 2 hours ago
People are having fun making images.
Like they do all sorts of things on the web.
Making an image is as simple as visiting a web-page.
That's it.
Comment by DrammBA 1 hour ago
Comment by nozzlegear 1 hour ago
Comment by jddj 1 hour ago
Comment by pixl97 18 minutes ago
Well, that and infographics on what to do when the AI's start to take over.
Comment by swader999 1 hour ago
Comment by nozzlegear 54 minutes ago
Comment by CamperBob2 1 hour ago
Comment by crab_galaxy 2 hours ago
Yeah ignoring the massive environmental, social, and moral implications I suppose it’s not a big deal at all.
Comment by bluegatty 2 hours ago
That's going to be an issue of regulation and cost - not some arbitrary decision over 'whether I should make an image or just use text'.
Second - there are really not much in the way 'social or moral implications' for the most part.
This is a dystopian delusion that I'm hinting at here.
Folks are blowing this way out of proportion, and cynically so.
People now have 'automated Photoshop' and a slightly more power to express themselves. It's not going to change that much, and it's not a bad thing, it's a good thing. Mildly.
Comment by mat0 31 minutes ago
Comment by minimaxir 21 minutes ago
Checks and balances in trust still exist, believe it or not.
Comment by sphinxterai 14 minutes ago
Comment by hiharryhere 2 minutes ago
Comment by cindyllm 2 hours ago
Comment by cchance 1 hour ago
Comment by cobolcomesback 1 hour ago
Nevermind the fact that the examples in OpenAI’s blog post are examples of creating images that have no plausible purpose other than to fake stuff that didn’t happen.
But yea sure, it’s all just “for fun”. Sure.
Comment by nozzlegear 1 hour ago
Comment by swader999 1 hour ago
Comment by hightrix 1 hour ago
A hammer is great for pushing nails into wood. It also can crack someone’s skull with efficiency. I’m assuming you do not use hammers then, right?
Comment by seventhtiger 1 hour ago
Comment by ai_critic 1 hour ago
Some of us find opportunities for the small joys that help offset the background radiation of despair.
Comment by throawayonthe 2 hours ago
Comment by well_ackshually 1 hour ago
3 billion images per week isn't people, it's automated farms, dogshit marketing companies and other attacks on your brain.
Comment by dinfinity 1 hour ago
What percentage is benign?
Do you have a source?
Comment by well_ackshually 1 hour ago
Comment by BalinKing 48 minutes ago
Comment by mat0 34 minutes ago
Comment by villish 2 hours ago
Comment by hmstx 2 hours ago
That game was developed before generating images became widespread, and sunset as that "art form" started really taking off (game-as-a-service... probably partly reverse engineered and retooled with AI too). So during its lifetime, it was pretty much AI free.
I remember having been seduced by that game's art style at the time. It wasn't "the usual guys with futuristic guns and armor", it had some stylization as well, nice color...
Watching the stream, it occurred to me that it was in fact artwork from "the days before", in a way I hadn't been hit with until that point, even as a returning artist looking at pre-AI works.
"People came up with the concepts and developed and rendered them themselves", I thought to myself. Not in a "oh look, floppy disks. How quaint!" way, either.
Weird feeling.
Comment by cbovis 1 hour ago
Comment by Rover222 2 hours ago
Comment by blep-arsh 1 hour ago
Comment by rustystump 2 hours ago
generating an ai image is vastly more costly than loading a webpage…but still not that costly compared to most other things.
this is depressing in the same way content slop or fyp feeds are fun.
if you want a silly pic with ur kids, just pay someone on fiver. it will be only slightly more expensive and likely far more exciting for kids as they have to wait for the dopamine hit.
Comment by dinfinity 1 hour ago
Incredible alternative. Talk about depressing.
Comment by bluegatty 2 hours ago
???
You mean to imply there's something 'more authentic' about paying $5 for a random Philippino guy to use Photoshop to mix-up some family photos with sticker overlays?
What are you even talking about?
On what level is that either more authentic, better for the environment, more novel ... of frankly better in any way?
That is 'not even an argument' - and so you're making my point for me:
People can now make images.
It's mostly for fun, some of it will be used professional, some Tweets will be a bit more visual - but that's it.
It's not going to change that much, and it's generally better on the whole for people to be able to create what they want.
People are putting this out of proportion.
Maybe I'm too old, but I've been around - this is not a big deal in the grand scheme.
And it's 'net slightly positive'.
Comment by cholantesh 1 hour ago
Or, y'know, learn how to draw, or accept that it's not exactly going to be Chuck Jones caliber stuff.
Comment by 0x70run 2 hours ago
Comment by emsign 2 hours ago
Comment by bluegatty 2 hours ago
It echos sentiments of painters, intellectuals and odd religious moralists and other rabble rousers who spent decades raving against 'tHe eVIL pHoToGRaPH' in late 19th and early 20t century.
People can now make more expressive images than they could before, which is generally a good thing.
Of course, people could always use new technology for negative purposes - and - people have been using Photoshop for decades already to make 'fake nudes' - this is not a new phenomenon.
None of this is even a very big deal. We can make images now much more easily, that's it.
Comment by xp84 2 hours ago
If anything, the fact that I know for a fact that I can, in 45 seconds, generate an image of any arbitrary person holding Solo cups at a party arm in arm with Vladimir Putin (to crib an idea from one of the demos) should make everyone less and less susceptible to blackmail in general - in fact, even if the source material is real! There was a time 40 years ago when a photograph's existence was considered proof of a fact - the existence of Photoshop-like tools made that into a weak proof, and the existence of genAI has made "a photograph" into a format that is known to be as malleable as a .txt file.
Whether we like that or not, no one will stop bad actors from attempting to abuse this technology, so democratizing its use does help people to understand how they work and make potential victims much less credulous.
Comment by recitedropper 3 hours ago
I challenge OpenAI to put a "your carbon footprint" field next to each generation. If you have nothing to hide, more information is surely better, right?
Comment by purpleflame1257 2 hours ago
Comment by recitedropper 2 hours ago
I await OpenAI's carbon footprint field.
Comment by raincole 2 hours ago
Comment by vlyan 2 hours ago
Comment by recitedropper 2 hours ago
Separately--yes, I would love carbon receipts for everything. Such a system would be fascinating and do real good in this world.
Comment by csallen 2 hours ago
The point isn't to justify anything.
It's to put things into perspective, and give people a more informed and holistic point of view, and then to ask them to reconsider their opinions in light of more information. There's nothing bad about that. In fact it's a good thing.
Especially in a world where people are so easily misled by others who cherrypick unexceptional statistics and then intentionally present them as exceptional in order to generate outrage, retweets, clicks, whatever.
Comment by rcoveson 2 hours ago
Comment by kraftman 2 hours ago
Comment by vlyan 2 hours ago
Comment by visarga 2 hours ago
Comment by Alifatisk 2 hours ago
Whataboutism?
Comment by EmiDub 2 hours ago
If someone says “I think x should stop because of y”. It is a valid argumentative response to say “y also occurs with z, so should z also stop?”. It is argumentatively revealing either hypocrisy, or that the original imperative relies on additional, unstated arguments/assumptions/biases.
Comment by pfannkuchen 2 hours ago
Maybe I’m confused, but I thought whataboutism was more like “you are saying I am doing X bad thing? What about Y bad thing you’re doing, let’s discuss that instead”.
It’s a deflection technique, not a check on logical consistency.
Comment by teravor 1 hour ago
> “you are saying I am doing X bad thing? What about Y bad thing you’re doing, let’s discuss that instead”.
no one is generally expecting or even asking for "let’s discuss that instead". what they are doing is trying to point out that your lack of of care about Y implies that you never actually cared about X in the first place but rather you only care about me doing X.often times X is actually bad but human nature is such that no one actually cares about it while wanting to appear to when locked down on it. for this reason arguing about X directly is bad optics.
whataboustism is basically an effort to draw attention to the selective enforcement of norms, laws, morality, whatever. hypocrisy is implied.
Comment by david-gpu 1 hour ago
Comment by LinXitoW 2 hours ago
AI haters have somehow discovered personal responsibility. Awesome. They've never ever applied that to anything else they PERSONALLY do, though. It's only things others do. For any personal choices, suddenly it's all the governments fault or the corporations fault. No personal responsibility to be seen.
My 5-ish years of veganism have easily made up for a life time of prompting. Yet, I still would never generate video or images in part because of their impact.
How many of you non-vegan AI haters are going to be vegan now?
Comment by recitedropper 2 hours ago
I have gone lacto-ovo vegetarian in the past year, so I'll take your 5 years of veganism as inspiration. Otherwise, I welcome people waking up to the harm they cause, regardless of their past hypocrisy.
Comment by shigawire 2 hours ago
People are flawed.
Comment by Zambyte 2 hours ago
Comment by magicalist 2 hours ago
Comment by Zambyte 2 hours ago
Comment by UltraSane 2 hours ago
Comment by LinXitoW 2 hours ago
It's really hard to believe that people care about the personal impact of their decisions, when it's only ever other peoples decisions that get talked about.
Comment by WarmWash 2 hours ago
But datacenters using water is what the public needs to fear...
Comment by masswerk 2 hours ago
Comment by woodgala 2 hours ago
Comment by gf000 2 hours ago
No one was talking about the ethical part (which you may or may not agree with). Carbon footprint is an actually measurable quantity that we can compare - what's your problem?
Comment by apitman 2 hours ago
Comment by numlock86 2 hours ago
Comment by handbanana_ 2 hours ago
Comment by lukeify 2 hours ago
Comment by gretch 50 minutes ago
It’s be like if your 2001 BMW car was super fuel efficient, and then for whatever reason the 2028 model guzzled gas and you chose to buy it anyway
Just because the thing you did once upon a time was efficient, does not mean you have a lifetime pass to engage with it no matter the updated circumstances
Comment by xmprt 2 hours ago
Comment by embedding-shape 2 hours ago
Comment by arrowleaf 2 hours ago
Comment by uludag 2 hours ago
Thankfully web browsing is no where near that.
Comment by pigpop 2 hours ago
Comment by tredre3 2 hours ago
Comment by cinntaile 2 hours ago
Comment by willchis 1 hour ago
Comment by sho_hn 2 hours ago
Comment by ndarray 2 hours ago
Comment by BoredomIsFun 2 hours ago
Comment by Betelbuddy 2 hours ago
Comment by villish 2 hours ago
Comment by spider-mario 2 hours ago
Comment by ModernMech 2 hours ago
Comment by UltraSane 2 hours ago
Comment by taurath 2 hours ago
Theres going to be a reckoning.
Comment by pizzafeelsright 57 minutes ago
Most are closed systems. Water gets recycled. Water gets used like bathtubs and water bottles, right?
You cannot believe water is destroyed.
Comment by nichohel 40 minutes ago
Comment by eboy 2 hours ago
Comment by kranke155 2 hours ago
Comment by ndarray 2 hours ago
Comment by UltraSane 2 hours ago
Comment by gpt5 2 hours ago
Comment by serf 1 hour ago
(hint : it's not all regex and checksums.)
Comment by dimitri-vs 45 minutes ago
Comment by zamadatix 12 minutes ago
Comment by dimitri-vs 47 minutes ago
Comment by busymom0 2 hours ago
Comment by Black616Angel 2 hours ago
Comment by Citizen_Lame 2 hours ago
Comment by kwanbix 1 hour ago
From my kids schools to banners for food places.
Comment by algoth1 1 hour ago
Comment by bahmboo 2 hours ago
Comment by MonkeyIsNull 2 hours ago
Golf: ~531 billion gal/year
Data centers: ~17 billion gal/year
Golf ≈ 31× as much
Indirect Water footprint via Electricity Usage: Golf: ~500–531B gallons
Data centers including electricity: ~228B gallonsComment by squeegmeister 1 hour ago
Comment by Revanche1367 2 hours ago
Comment by wolfy1993 2 hours ago
Or do you literally mean the entire sport of golf vs 1 data center?
Comment by alexk307 2 hours ago
Comment by xp84 2 hours ago
The fact is, at this point there is so much demand for DC capacity that the prices are super high, and thus it's worth building them. Many people believe that it's a bubble or whatever. If they're right, well then a lot of people building DCs right now will lose their shirts. If they're wrong and the demand is sustained, DCs are exactly what we need to be building.
Comment by djeastm 18 minutes ago
It's not really arbitrary though, is it? It's decided by duly-elected representatives who answer to the citizens of the municipality. A government "of the people", etc.
I feel like what you're arguing against is precisely why we have local governments in the first place.
Comment by Black616Angel 2 hours ago
Yes, some people have fun, but most of the "fun" nowadays is just endless mindless consumption.
Comment by moomoo11 1 hour ago
Comment by raincole 2 hours ago
Comment by Scrapemist 2 hours ago
Comment by rs_rs_rs_rs_rs 3 hours ago
Comment by davidweatherall 3 hours ago
Comment by handbanana_ 2 hours ago
Comment by fsloth 2 hours ago
How is the person supposed to know your opposing point of view if you don't tell it?
Comment by freedomben 2 hours ago
Comment by Banditoz 28 minutes ago
Comment by simianwords 2 hours ago
Comment by ksd482 3 hours ago
But I see your point too. I too like to generate cute and funny images and share it with my friends and family.
But this being done at scale can lead to overall degradation of the online experience.
Comment by conradludgate 2 hours ago
Some people feel the need to attach an image to almost everything the post - even in chat rooms (including slack at work). These images serve no purpose, but the poster feels it is useful - though it was often reaction gifs in the past I'm seeing a lot more AI content than I ever saw reaction gifs, maybe because of novelty or maybe because it's ultra personalised.
Sometimes images help to visualise something or to get a point across, but I see so many people who think it's necessary to reply to a discord message with a cat with human limbs doing a dance, or a photo of "themselves" climbing a mountain with the Rust logo to show them mastering Rust... Ok?
Image models are useful and I'm thankful for much better visual reasoning but it really really frustrates me the constant need to burn money for all of this slop.
Comment by paimapi 2 hours ago
>Generating 1,000 images with a powerful AI model, such as Stable Diffusion XL, is responsible for roughly as much carbon dioxide as driving the equivalent of 4.1 miles in an average gasoline-powered car. In contrast, the least carbon-intensive text generation model they examined was responsible for as much CO2 as driving 0.0006 miles in a similar vehicle
Comment by XenophileJKO 2 hours ago
If anything now I feel LESS guilty about usage.
Comment by Black616Angel 2 hours ago
* if we trust the numbers in the other comment
Comment by XenophileJKO 1 hour ago
It equals around 0.004% of actual driven miles if my estimates are right.
That is with an estimated ~270 billion miles vehicle miles every week (I put commercial in there.. about ~200 billion miles if you only include passenger vehicles.)
Comment by Kranar 2 hours ago
For context that means a days worth of image generation emits about the same amount of carbon dioxide as a single transatlantic flight (New York to London). There are approximately 1500 transatlantic flights per day.
Comment by arjie 2 hours ago
Comment by paimapi 1 hour ago
Comment by arjie 1 hour ago
Comment by paimapi 28 minutes ago
Comment by whatisthiseven 2 hours ago
Ours cars are atrocious. We can have the most powerful artificial minds imagine 1000 images from simple prompts, and that takes as much energy as moving a human being 4 miles.
And that will get more energy efficient. Gas cars have barely budged.
Gas cars are the problem, not AI.
Comment by applfanboysbgon 2 hours ago
(Note, however, that OpenAI's image model is much much larger than SDXL; however, we don't have precise numbers for it. Nonetheless, misinformation is misinformation.)
Comment by gloxkiqcza 2 hours ago
Comment by lopatin 2 hours ago
Comment by sensanaty 1 hour ago
Comment by tlogan 2 hours ago
Concerns about misinformation? Compute/energy use? Creative work Replaced by AI?
Comment by andai 2 hours ago
Comment by Rover222 2 hours ago
Comment by spike021 1 hour ago
Comment by embedding-shape 1 hour ago
Comment by mattbettinson 3 hours ago
Comment by icedrift 2 hours ago
Comment by Macuyiko 2 hours ago
Comment by 100percentjake 2 hours ago
Comment by sanid 2 hours ago
Comment by tokioyoyo 4 minutes ago
Comment by embedding-shape 2 hours ago
Comment by visarga 2 hours ago
Comment by magicalist 2 hours ago
If it doesn't look like the food you're about to purchase, yes? Why is this even a question in this context.
Comment by strangecasts 1 hour ago
(Have also heard about people - not just scammers/spam operations - using generated pictures and descriptions for dating profiles and real estate ads, and I cannot understand what outcome they expect when someone follows up and immediately realizes what's up)
Comment by WarmWash 1 hour ago
Comment by strangecasts 1 hour ago
Comment by lambdas 2 hours ago
Comment by Betelbuddy 3 hours ago
Comment by steve_adams_86 2 hours ago
Comment by dude250711 2 hours ago
Comment by bko 2 hours ago
Comment by ks2048 2 hours ago
Comment by bko 2 hours ago
Comment by embedding-shape 1 hour ago
Comment by Betelbuddy 2 hours ago
Comment by yoz-y 2 hours ago
Comment by krelian 2 hours ago
Comment by ebbi 42 minutes ago
Comment by coffeebeqn 3 hours ago
Comment by tyjen 3 hours ago
Comment by nxc18 1 hour ago
Comment by Insanity 2 hours ago
Comment by arrowleaf 3 hours ago
Comment by infecto 3 hours ago
Comment by rexskimmer 3 hours ago
Comment by emsign 2 hours ago
Comment by throwaway613746 3 hours ago
Comment by Nition 30 minutes ago
In order, there's:
- Imagine you had better decorations for your fish.
- Imagine you had a tattoo of your pet.
- Imagine you had a unique candle holder.
- Imagine you had a worse haircut.
- Imagine your flowers were arranged.
- Imagine your child had a suit.
- Imagine your dog had a costume.
- Imagine you owned a scanner.
- Imagine you could hang out with friends.
- Imagine your bed was made.
Comment by jjcm 2 hours ago
It's wild how much of a difference this is - images are coming in at around 35-40s. Very noticable, and makes a difference when you're iterating quickly: https://jjcm.org/gpt-image-2.5-speed.mp4
Comment by jjcm 1 hour ago
Warcraft 3 style agentic dev interface: https://image.non.io/cd9ea5cd-8ed7-44e0-ad3f-480ff0e51875.we...
Overall it used the reference images I gave it a bit better than gpt-image-2. I noticed 2 had issues getting the blue button just right. 2.5 nailed it.
A "John Politics" meme site: https://image.non.io/d2922164-fa96-4d07-a141-2febadb02939.we...
Did very well modifying the pose while keeping the appearance of Glenn Powell. gpt-image-2 had a lot of the "fried" look for some of his skin in prior designs I did for johnpolitics.com
A cyberpunk inspired ramen website: https://image.non.io/8d5d8f10-0f0f-4d91-b5ea-33af7538b150.we...
Dark mode sites surfaced the fried look quite a bit in prior models, but this definitely looks better on that front. One thing that looks perhaps worse though is the microglyphs - note the teal lines to the bottom right of the ramen, they're kinda blurry / not straight.
Overall fixed some of the main issues / gripes I had with gpt-image-2
Comment by raincole 1 hour ago
Comment by jjcm 1 hour ago
I'll also be adding in microsoft's mai-image-2.6 soon, but I need to update my SOC2 to add MS as a provider before I turn that on for other users. The full list of models in rotation is here: https://image.non.io/d53a9760-8b74-4386-b032-d59da2cd5319.we...
Comment by raincole 1 hour ago
Comment by jjcm 1 hour ago
Edit: confirmed that the download should be fixed now - should be properly a png now.
Comment by meerita 2 hours ago
Comment by ActionHank 2 hours ago
Comment by Escapado 1 hour ago
I just tried out the new version of ChatGPT to generate images for the main characters in a concept art style and for me it makes the story come to life a bit more.
I also run a Pathfinder Campaign every other week and have found great pleasure visualizing scenes in this way for myself. Sometimes I share them with the players and so far the feedback on that has been very positive.
I wish I could picture things in my mind but at least now I have something that can help me with my handicap!
Comment by meerita 7 minutes ago
Comment by throwup238 2 hours ago
Comment by Starlevel004 2 hours ago
Comment by RobinL 2 hours ago
Comment by andai 2 hours ago
https://www.reddit.com/r/HelloInternet/comments/f0wuej/are_y...
Comment by alstonite 2 hours ago
Comment by SirMaster 1 hour ago
Comment by dinfinity 1 hour ago
I have a friend who only reads fiction for which he's seen the accompanying show first because he otherwise has no visual reference for the characters and environments.
Comment by meerita 12 minutes ago
Comment by fsloth 2 hours ago
Comment by AyanamiKaine 3 hours ago
The sad part is my mother would love "remixing" my old child photos of me.
Comment by cobolcomesback 2 hours ago
What is the point of having a fake picture of your dog in a costume? What’s the point of having a fake picture about being at a party?
The only use case I can think of for this is for someone who likes to make up stories and lie about what they’ve done. Is that really the target market?
Comment by xp84 1 hour ago
Yes, it's a major first-world problem to not get a photo, but to the right person, it could mean a lot to them to "fix" the photo they took with A and B to add C to their 'rightful place.'
Comment by AyanamiKaine 2 hours ago
Mark where you at the party today? Yes of course look at these pictures I took (╥﹏╥)
Comment by embedding-shape 1 hour ago
As far as I can tell, the biggest uses of these sort of models is to pump out lots of content on multimedia-based social media, so yeah, it's basically for the sort of person who doesn't shy away from exaggerating, omitting or outright lying in order to get more views.
Comment by onion2k 3 hours ago
Except you need to have taken a photo of the unmade bed, uploaded it to ChatGPT, prompted for the 'fake made bed' version, and downloaded the image.
Surely it's less effort to just make the bed?
Comment by andai 2 hours ago
Comment by numlock86 2 hours ago
Well, that's like the entire point of social media anyway, isn't it? People there will _love_ this. Ugh ...
Comment by eclipticplane 2 hours ago
Comment by jodacola 2 hours ago
Can you imagine finding a stack of photos in the basement with timestamps of the Before AI times and wondering whether they are real or just got swapped with generated and printed fakes? Scary!
It’s cool, folks… nothing bad is going to happen. Right? Right?!
Comment by AyanamiKaine 2 hours ago
Comment by bakhlawa 2 hours ago
Comment by rs_rs_rs_rs_rs 3 hours ago
How is that sad? Why is your mothers joy sad?
Comment by AyanamiKaine 2 hours ago
Imagine a time where a small moment is actually a smaller amount experienced then the remixed one. At one point you will have more memories of fake events, and more emotional beats for said fake events than real ones.
AI videos showing you a second life if you just had chosen a different path.
People will be depressed from it.
Comment by m_fayer 2 hours ago
Comment by m_fayer 2 hours ago
Comment by handbanana_ 2 hours ago
Comment by breezybottom 2 hours ago
Comment by echelon 3 hours ago
For thousands of years people have told embellished stories, and humanity has thrived upon it. Tall tales, fish stories. Heck, that's still the average person's experience unless they've really worked their critical thinking muscles.
Today's working adults are used to the short thirty-year "safe space" of smartphones and internet. We grew up in a temporary meta stability where "truth" was "recorded" and could be "relied upon", and now that the fundamentals are shifting, we're complaining that the physical world is amenable to storytelling and imagination once again.
Cry me a river. This is awesome and is a direct consequence of everything I ever wanted the future to be: magical creative superpowers. I wanted to graduate into the world that is emerging now rather than spend the first third of my career in incrementalism and slow progress.
2008 - 2020 sucked. What a total lull. AI is healing these things and putting us back on track for the jet pack future we grew up dreaming about. It's taking us back to a creative world without shitty platforms controlling what we say and do, and without a ceiling on what we can accomplish.
It feels like every day we're unwrapping a new present or several. Not small things, but reality-shattering things that fill me with inspiration to build and explore. It's so much fun.
Comment by AyanamiKaine 3 hours ago
There are many many people that will take an image as an undeniable fact of the real world. While before you could fake(photoshop) things it took more time then 30 seconds.
I wouldnt regulate these aspects, because I believe pandoras box is already wide open. Regulating anything wouldn't change much.
Comment by pelzatessa 3 hours ago
Comment by polytely 3 hours ago
Comment by headz 3 hours ago
They probably didn't care to check those images in detail.
Comment by kfarr 2 hours ago
Comment by vunderba 1 hour ago
https://developers.openai.com/api/reference/python/resources...
Comment by vunderba 2 hours ago
Comment by nzach 3 hours ago
Comment by pllbnk 3 hours ago
Edit: I feel stupid I didn't see the original OP already mentioning three finger issue. I'll just leave it here.
Comment by fsloth 2 hours ago
Comment by timbaboon 2 hours ago
Comment by nilsherzig 2 hours ago
Comment by minimaxir 2 hours ago
gpt-image-2.5-sunburst: 1421
gpt-image-2.5-flare: 1399
gpt-image-2 (medium): 1381
mai-image-2.6: 1331
Even with LM Arena being flawed, this is significant. I was planning to do a writeup on the original gpt-image-2 as it crushed every complex image comprehension benchmark I had...I'm glad I procrastinated since ChatGPT Images 2.5 seems like an even better starting point to test out what these models can actually do nowadays.
Many people still think AI images output the wrong number of fingers on a regular basis. (EDIT: this was an ironic comment to make in hindsight and I own it)
Comment by yawn 2 hours ago
There's literally an image of a dude with 3 fingers in the Composite Party Photo.
Comment by minimaxir 2 hours ago
Comment by addandsubtract 35 minutes ago
Comment by yoz-y 2 hours ago
Last time I used a frontier image model it made me a seal with three hands so…
Comment by handbanana_ 2 hours ago
Comment by tyjen 2 hours ago
Comment by vgalin 2 hours ago
Comment by VladVladikoff 2 hours ago
Comment by dude250711 2 hours ago
Comment by bufbupa 3 hours ago
Comment by emadabdulrahim 56 minutes ago
Added support for GPT 2.5
Comment by vunderba 2 hours ago
Comment by gpt5 2 hours ago
Comment by vunderba 1 hour ago
The agentic tooling in some of the latest models is getting pretty crazy too - somebody on Reddit recently shared an example of using GPT-6 Astra to generate sprites with Aseprite, an open-source, popular pixel art editor and it worked shockingly well.
Comment by pigpop 1 hour ago
Comment by ramesh31 2 hours ago
The key is making them one frame at a time, rather than asking for the whole sheet. Adherence frame-to-frame with a reference image is really good, so just prompt with the previous + direction. I finally got Nano Banana to make fairly decent fluid animations that way.
Comment by itomato 2 hours ago
Comment by RobinL 2 hours ago
That said for single images the old model was already okayish for prototypes
Comment by TomGarden 2 hours ago
Comment by Sivart13 3 hours ago
Comment by lefty2 1 hour ago
Comment by gottagocode 3 hours ago
Comment by Jonovono 2 hours ago
Comment by geooff_ 1 hour ago
At this point, I'm much more interested in seeing releases for models chasing down prices on "good enough" to unlock new use-cases than I am SOTA.
I've been using Grok Imagine V1 in prod for months now as its quality for my use-case is already more than enough.
Comment by __MatrixMan__ 2 hours ago
What kind of person cares enough to buy flowers for an arrangement, but is uninterested in arranging them? What's next, will we buy music for our AI's to listen to? Food for them to eat?
There's a lot of drudgery that can be automated away, but this seems to be focused on automating the fun parts away.
Comment by arjie 2 hours ago
Comment by Kloopvram 47 minutes ago
Comment by WarmWash 1 hour ago
I'm 99% sure that that is the main tell AI sniffers rely on.
Comment by throwaway314155 23 minutes ago
Let's see Paul Allen's card.
Comment by redox99 34 minutes ago
Comment by bearjaws 3 hours ago
Everywhere I go now I see places that blatantly used ChatGPT for their menus or posters and they all look the same.
It feels almost existential to the service, I know its a bit of survivor bias but so many images you can tell immediately are ChatGPT vs other image providers, and I feel many people are sick of them.
Comment by mickeyp 3 hours ago
They'd pick a kebab from a menu of professionally made kebab pictures the printer has in stock. Or worse, they'd take pics of their own plated food with a dead-centre point flash in a dark cupboard or something and you end up with awful-looking food, no matter how good.
Comment by shagie 2 hours ago
The original photo I took was at https://www.reddit.com/r/Tovala/comments/1pfuwrb/meatloaf_pa... - it's a cheeseburger meatloaf with potato wedges taken with a phone camera. And I was going for a consistent documentation approach for the photograph, not trying for menu proper.
https://chatgpt.com/share/6aa05d33-3e6c-83ea-9e32-f1371c1ad6... ( https://imgur.com/a/fPB5VKP for just the image)
I suspect that someone doing a menu could take a properly plated meal from the kitchen (rather than me photographing on top of my oven) and have it get redone for a good image for a menu without fundamentally changing what is being served.
Comment by core_dumped 3 hours ago
Comment by infecto 3 hours ago
Comment by breezybottom 2 hours ago
Comment by infecto 2 hours ago
Comment by sebzim4500 3 hours ago
Comment by masswerk 2 hours ago
Comment by ieie3366 3 hours ago
Comment by alexgyurov 2 hours ago
Comment by raincole 2 hours ago
Comment by hspeiser 3 hours ago
Comment by furyofantares 2 hours ago
> Pricing and availability
> Images 2.5 is rolling out today to ChatGPT, ChatGPT Work, and Codex users across all tiers on desktop, mobile, and web.
> GPT‑Image‑2.5 Sunburst and GPT‑Image‑2.5 Flare are available in the API. See pricing details here.
and the link to the pricing details is a page where i am either too dumb to find the pricing details, or they don't exist
Comment by raincole 2 hours ago
Comment by furyofantares 1 hour ago
Same price as image-2, and I'm guessing it's not more tokens (given the speedup).
Comment by bilsbie 2 hours ago
I had trouble weeding through all the marketing speak.
Comment by mudkipdev 3 hours ago
Comment by wrcwill 3 hours ago
Comment by kiririn7 1 hour ago
Comment by teaearlgraycold 2 hours ago
Comment by itomato 2 hours ago
Comment by pllbnk 3 hours ago
Comment by scurnus 2 hours ago
Comment by rarisma 2 hours ago
Comment by _doctor_love 1 hour ago
I super encourage others to learn just a little bit of art technique and your images are going to go to the next level. You need a little construction, perspective, gesture, anatomy, and you're pretty much off to the races.
Knowing names of illustration styles is great too. Look up "style transfer" if you are not familiar.
Comment by clement_b 3 hours ago
Comment by srott 3 hours ago
Comment by starone99 2 hours ago
Comment by benatkin 2 hours ago
Edit: finally, I have my horse on the moon. https://chatgpt.com/share/6aa0627e-b168-83e9-98e5-d5ad5a9cad... Though the horse doesn't have a spacesuit, neither does it have one in the Stable Diffusion wikipedia page.
Edit 2: finally got the kind of result I wanted https://chatgpt.com/share/6aa06813-4818-83e9-9c19-8c8ac9a348... https://chatgpt.com/s/p_359c0c38939c81918addaaa9de196f4e
Comment by vunderba 1 hour ago
For the heck of it, I ratcheted up the difficulty of the prompt (two-headed horse, astronaut with the visor up, etc.), and gpt-image-2 got it right every time.
Comment by nilsherzig 2 hours ago
Comment by iLoveOncall 3 hours ago
Comment by leokennis 2 hours ago
Comment by charcircuit 3 hours ago
Comment by dogomatic 2 hours ago
Comment by UltraSane 2 hours ago
Comment by _pdp_ 3 hours ago
Comment by webdood90 2 hours ago
The majority of the world does not want or need any of this, yet the nerds in SF that can hardly hold a conversation with another human being are pumping it out as quickly as they can.
I truly think we're fucked.
Comment by charcircuit 41 minutes ago
Comment by wetpaws 3 hours ago
Comment by parasxos 2 hours ago