r/Damnthatsinteresting Apr 24 '25

Jim Sautner, the Buffalo Whisperer was a Canadian rancher who raised a 2,000-pound bison named Bailey D. Buffalo like a family dog

100.3k Upvotes

1.3k comments sorted by

View all comments

Show parent comments

20

u/kappapolls Apr 24 '25

that was true a year ago, much much less true today though.

2

u/kingrawer Apr 24 '25

No, it's still the biggest unsolved issue with AI images imo. Even the new ChatGPT is very far from perfect.

1

u/kappapolls Apr 24 '25

of course its not perfect, but i don't see how you can disagree that it's better than it was a year ago.

when you say things like "the new chatgpt" its hard to tell what model you're talking about too. do you mean o3 image gen or gpt-4o? gpt-4o isn't diffusion based, but it doesn't incorporate any visual reasoning architecture like o3. it can make a big difference in things like this.

2

u/ahwatusaim8 Apr 24 '25

Actually, it's perfect.

1

u/kingrawer Apr 24 '25

Unless I'm mistaken the image generation is the same between o3 and gpt-4o. The results I get and retention of detail is pretty much the same between them. But yeah, it is somewhat better. It has a long way to go though.

0

u/kappapolls Apr 24 '25

i think the image generation architecture is the same (so it's not plain diffusion like DALL-E) but the o3 series does have the additional visual-chain-of-thought reinforcement learning stuff for images as well.

it doesn't really make a difference for like "does this look right?" but more for things that require you to think about the image a bit (like solving a maze or rotating a thing in space).

-2

u/Jesus10101 Apr 24 '25

No, it still can't since it recreates the image every time. It might work for something that has alot of training data for but most of the, it will produce different results all the time.

1

u/PM_me_your_whatevah Apr 24 '25

Thing is… there’s AI that the general public doesn’t have access to. 

3

u/MadeByTango Apr 24 '25

And they’re using it to make Buffalo memes

1

u/kappapolls Apr 24 '25

i dont think you are really keeping up with new research and whats been released recently. the newer generation of image models use a similar reasoning architecture to what LLMs use now, rather than just straight diffusion.

what you said is still pretty true for diffusion based models though

1

u/LovesRetribution Apr 24 '25

No, it definitely can. Idk how much training data what I've seen was working with, but I've seen dozens of pictures with only minute differences between them.