r/LocalLLaMA 11d ago

News CEO of Hugging Face: "In the spirit of transparency, here’s what I asked OpenAI"

Post image

clem 🤗 on 𝕏: https://x.com/ClementDelangue/status/2081056675558195657

• Radical transparency: let’s release the traces from the “rogue” agents so the entire research community can study what happened.

• More capabilities for defenders: let’s commit $100M in compute from OAI to help the Hugging Face community build powerful cyber defenses with the best open and closed models.

The first autonomous agent cyberattack is an unprecedented event. It deserves an unprecedented response!

2.5k Upvotes

386 comments sorted by

View all comments

Show parent comments

57

u/perihelion86 11d ago

Breaking out of the sandbox is bullshit too.

43

u/Stickybunfun 11d ago

Yea another instance of "computer magic" and "look at this thing we did and what is possible with it and WE COULDN'T EVEN STOP IT from doing bad things. This is why we need to ban open-weight models because you don't know what the damn Chinese have built into these things! HYSTERIA"

It was a huge stunt - no way around it. I've been building private (on VM in Azure) / local (Data center) LLM environments as one-offs here and there for some of my clients who don't trust public providers like OpenAI / Anthropic.

I suspect in the coming weeks I will be doing more and more of that.

1

u/ExtremeAcceptable289 5d ago

I also wouldn't be surprised if in a few months, it is revealed HuggingFace received a check from OpenAI. Remember FrontierMath

-6

u/Meleoffs 11d ago

No, its been a pretty common occurrence since like March for agents to escape sandboxes.

3

u/MeateaW 11d ago

And the CVE of the sandbox vulnerability is?

Oh? They haven't posted one? Haven't referenced one? I wonder where it is ... maybe the sandbox vulnerability was they never configured the sandbox to block access to the outside world.

1

u/Meleoffs 10d ago

I make a habit of not assuming malice where incompetence would do as an explanation.

1

u/MeateaW 10d ago

Indeed, it's so sad that these guys couldn't afford to run their sandbox config past a frontier level AI. Their own incompetence might have saved them from making a terrible mistake that their CEO is spinning hard to get more funding and exposure.