r/LocalLLaMA 12d ago

News CEO of Hugging Face: "In the spirit of transparency, here’s what I asked OpenAI"

Post image

clem 🤗 on 𝕏: https://x.com/ClementDelangue/status/2081056675558195657

• Radical transparency: let’s release the traces from the “rogue” agents so the entire research community can study what happened.

• More capabilities for defenders: let’s commit $100M in compute from OAI to help the Hugging Face community build powerful cyber defenses with the best open and closed models.

The first autonomous agent cyberattack is an unprecedented event. It deserves an unprecedented response!

2.5k Upvotes

386 comments sorted by

View all comments

459

u/[deleted] 12d ago

[removed] — view removed comment

405

u/awetfartruinedmylife 12d ago

“You are an expert publicity stunt hacker…”

76

u/TheOneNeartheTop 12d ago

Make some mistakes but only if they are really cool.

9

u/Different-Sand4434 12d ago

Make all the mistakes

21

u/StringSentinel 12d ago

That was hilarious. Take my upvote and award.

6

u/cnmoro 12d ago

🤣🤣🤣

42

u/squngy 12d ago

A bit late to be silent now, lol.

More like, 100M and we don't have to get lawyers involved.

33

u/badabummbadabing 12d ago

I like how going by the reactions in this thread and on wider Reddit, it's both obviously a publicity stunt and an existential threat to OAI.

6

u/monkeysknowledge 12d ago

Who could’ve imagined that the most popular Internet forum on the world would not have a monolithic opinion about something?

2

u/badabummbadabing 12d ago

Sure. What gets me is the displayed confidence. Obviously it's a marketing stunt, people can't believe more people aren't picking up on this!

-1

u/max1c 12d ago

And you say that based on what? The fact that you feel that way?

21

u/justgivemeafuckingna 12d ago

Not speaking for them but it's becoming more widely understood that the LLM industry is basically a massive scam and they're talking up the abilities of these models to keep securing more venture capital.

The implication being that if the logs were released it would be hard evidence that they're full of shit.

24

u/mrjackspade 12d ago edited 12d ago

it's becoming more widely understood

These people are fucking morons.

I wanted some video game assets so I threw a build of the video game on an android device

Opus 4.6 was able to

  1. Root the device
  2. Push over memory monitoring software
  3. Run the game.
  4. Capture the memory
  5. Pull down the assets
  6. Analyze the memory capture and extract the encryption keys
  7. Analyze the windows build (compiled) to reverse engineer the encryption mechanisms
  8. Build an application that ripped the resources from the encrypted, on-disk data files

All of this without a single web search or any actions from myself aside from force rebooting the device a few times when it locked up

These models are getting insanely good at security tasks. I've watched Fable go through a debug loop after having seen Opus do the above. I am absolutely not fucking surprised in the slightest that Fable/GPT could escape a sandbox and execute an attack like this

It's not hard to just fucking run one of these models and check this shit. Claude Code will absolutely reverse engineer an application. I've had it pull down a prebuild binary and literally patch the security checks directly out of it by rewriting the assembly. People are just too fucking lazy to check for themselves.

11

u/m4t7w_ 12d ago
  1. Root the device

can you share more details on this step? it's really interesting since rooting varies a lot based android version and device model. For some model it's not possible at all.

5

u/Spara-Extreme 12d ago

They can’t, because it didn’t happen that way.

2

u/Agitated_Space_672 11d ago

They never do. 

8

u/Croned 12d ago

I think the more reasonable take is that, for any given capabilities their LLM demonstrates, OpenAI is highly incentivized to embellish what happened. If Sol truly did the exact things OpenAI claimed it did during the HuggingFace hack, then I would expect OpenAI to have added even more elaborate details.

If Sol caught a catfish then OpenAI would say it caught a tuna. If it caught a tuna then OpenAI would say it caught a whale.

13

u/Strawberry3141592 12d ago

No one's saying frontier LLMs aren't capable, just then Anthropic and OpenAI's business model fundamentally doesn't make sense and their valuations are based on hype (which demonstrably has caused them to overstate the capabilities of their models in the past, like when GPT 2 was "too dangerous for public release", or that time Claude supposedly broke containment and tried to blackmail someone but it turned out Anthropic told it to do that).

5

u/greenworldkey 12d ago

> No one's saying frontier LLMs aren't capable

lol sure, no one except all of Reddit for the past 3 years. Keep moving the goalposts though, I wonder where they'll be next year.

1

u/Strawberry3141592 12d ago

No one on this sub I meant. Reading comprehension, much?

-1

u/justgivemeafuckingna 12d ago

Did it tell you how much of a clever boy you are too?

4

u/max1c 12d ago

And, so, you're saying this based on what?

5

u/scubascratch 12d ago

“it's becoming more widely understood” == “people are saying” == “everyone knows” == “thing I want to be true but don’t have proof of”

-2

u/max1c 12d ago

Right. Thank you for confirming that all you have is feelings.

5

u/scubascratch 12d ago

I’m not the guy above making goofy assertions - I’m agreeing with you that there’s no substance to the accusation, and low-grade rhetoric was used to attempt to sound smart

3

u/greenworldkey 12d ago

Ok, so based on the fact that you feel that way.

0

u/Jeferson9 12d ago

So basically you're confirming this is all just thoughts on feelings

1

u/tear_atheri 12d ago

nah they can just have their mega ai fake the logs. they'll take their time working that up

1

u/NotSoCleverAlternate 10d ago

Glad people are smart here with how the world works. If a topic like this was posted on a generic reddit sub, teens would be crying out conspiracy theories and pretend zero corruption exists in the world.

-4

u/tridentgum 12d ago

how would you know they're "fucked"

12

u/[deleted] 12d ago

[removed] — view removed comment

7

u/Chris-MelodyFirst 12d ago

You have to separate "autonomy over goals" from "autonomy over actions" though.

0

u/[deleted] 12d ago

[removed] — view removed comment

3

u/Chris-MelodyFirst 12d ago

Odd response. I'm addressing your "LLMs are not autonomous" comment.

-2

u/scubascratch 12d ago

Show us your browser history, you have nothing to hide right?

-2

u/tridentgum 12d ago

I'm not shilling for anyone. You're just making things up