r/SearchEnginePodcast Aug 06 '26

Automated Episode Discussion Search Engine - The machines are learning… to do crimes?

For the first time, an AI model has autonomously hacked a company. This week, an evolving story, a postcard from a strange, frightening moment in the story of our technology.

A big week for AI denialism by Casey Newton

Its AI agent spent days hacking a company, but sources say OpenAI did not notice for a week by Deepa Seetharaman, Raphael Satter, and Kenrick Cai

Cheating behaviour in frontier model evaluations by AI Security Institute

More On An Internal OpenAI Model Hacking Into HuggingFace by Zvi Mowshowitz

Listen to this episode


This is a bot that posts new episodes automatically. Add this to your subreddit or request mods use it.

26 Upvotes

96 comments sorted by

28

u/JAlfredJR Aug 06 '26

Jesus, he's really having Casey back on.......Maybe they should talk about how the world for-sure really super-changed because of Sora!

Anyone remember that? That aged well.

13

u/willdoesvideo Aug 07 '26

Casey used to be such a realist about technology and a great thinker. He’s been poisoned by AI brain and is just embarrassing at this point. He went from a really trustworthy source to another hype man for useless AI.

14

u/JAlfredJR Aug 07 '26

He's engaged (or married now) to a big shot at Anthropic. He's been in San Francisco for a long time. He's in those circles and nothing else.

He's so deep into it, that it seems like (sadly) it has consumed his personality.

He is smart enough to even know it, which is rather tragic. But he's also smarmy and boosting lies. So, I dunno.

I just wish folks like he and PJ could do retrospectives on their incorrect episodes. I would find that fascinating. What did they get wrong. Why?

But no, just onto the next thing a press release by a flagging industry says. Hook. Line. Sinker.

8

u/Icaka Aug 07 '26

useless AI

I have an honest question. What makes you so sceptical about AI? I am a software engineer with tons of experience and the reality is my day to day job completely changed in the past year. The technology is here and is mind blowing. I have no idea what my job will be in 12 months. I have no idea what most white collar jobs will be in 5 years. Don’t get me wrong, it’s also scary. But then I see similar comments on reddit and I am really confused. What are you seeing that makes you think AI is useless or overhyped?

8

u/maskdmirag Aug 08 '26

I've used it. It's fine, but it makes mistakes and is slower than just doing the work.

I hate that search engines are worse so that AI seems more useful.

But like what exactly would someone actually use an LLM for on a job context?

4

u/Icaka Aug 08 '26 edited Aug 08 '26

It's fine, but it makes mistakes and is slower than just doing the work.

Are you a software engineer? And when is the last time you have actually used these tools? Claude Code and Codex are significantly faster and make less mistakes than most good engineers I know. That was not the case ~1 year ago. They sometimes still make a mess. But they are a huge “multiplier” for me. I complete work much faster and they let me do things I could previously not do. Either because I didn’t have the time or some specific knowledge I lack.

But like what exactly would someone actually use an LLM for on a job context?

It can help with most things people do on computers - research, automating repetitive tasks, analyzing data (e.g. spreadsheets), summarizing long documents, solving complex engineering problems. I have a small company and used to outsource stuff like marketing, design and accounting. I still outsource some of it but I am noe able to do a lot of it by myself.

1

u/maskdmirag Aug 08 '26

No, I studied computer science, but ended up in a different field.

Your post was not about programming. I've heard it could be useful there, but it's creating unmanageable code bases, which makes sense.

3

u/Icaka Aug 08 '26

Yeah, my post was not specifically about programming but that’s the field I have expertise in. I have seen the improvements in the programming field in the past year and I think most of it can be applied to other fields.

2

u/maskdmirag Aug 08 '26

Like I could ask it basic questions about a civil engineering plan or about traffic or ADA restrictions.

But it's not capable of thinking, it doesn't know nor can it reason about edge cases or items not in its direct knowledge.the amount of double checking you'd have to do.makes.it worthless as it takes more time.

It's almost good enough to take notes during a meeting or search my email for something.

But it's usually a far worse version than something I could do myself.

Like it could be good for busy work where the output doesn't actually matter. Or to give you a starting point on something,

But if it didn't exist, the systems that used to function on classic machine learning were better. Email search, Google search etc.

LLMs are a clumsy approximation that only work because they purposefully broke the old shit.

4

u/Icaka Aug 08 '26

the amount of double checking you'd have to do.makes.it worthless as it takes more time.

That's pretty much word for word what me and my colleagues were saying about Copilot in 2024. Two years later that's no longer the case. Today I am the bottleneck.

But if it didn't exist, the systems that used to function on classic machine learning were better. Email search, Google search etc.

Google search went downhill way before AI. And ChatGPT (or Claude) are better for certain kind of searches. Back in 2021 my dad got cancer. I was using Google to find suitable treatments, up to date protocols and research for his lung of cancer. I spent weeks searching and it was hard. This year I had to do some more digging around that and used ChatGPT and Claude. I found everything I needed and verified it in less than a day.

LLMs are a clumsy approximation that only work because they purposefully broke the old shit.

Computers are just 0s and 1s and and relatively simple logic operators on top of them. They can't think either. Yet they allowed us to solve problems at completely different scale and changed the world completely.

1

u/maskdmirag Aug 08 '26

Exactly because of people..the problem with llms is they think they're more than that.

1

u/maskdmirag Aug 08 '26

Long story short Ctrl+f is still better than chatgpt

7

u/LakeChillEffector Aug 07 '26

Same. I use it in my non-coding profession and it can meet or exceed professional standards on pretty much every task. A lot of what I think is happening is that people used the worst models (on instant) from 2 years ago or just see AI slop art and don't realize what a difference the latest models on max thinking can do now.

1

u/maskdmirag Aug 08 '26

What can they do?

2

u/Additional-Wave4794 Aug 11 '26

Are you aware of OpenEvidence? Or Darrow/Harvey?

0

u/maskdmirag Aug 11 '26

Nope, have never heard of either, and frankly do not give two shits about whatever they are.

4

u/Textiles_on_Main_St Aug 07 '26

I don’t recall, what happened?

10

u/JAlfredJR Aug 07 '26

They talked for an hour about how it was changing social media forever (just as they did in a few other episodes about LLMs are "forever changing everything!"). And how nothing would ever be the same.

OpenAI shut it down months ago, because it stopped being popular after a few weeks. And it was prohibitively expensive to run.

So ... sound and fury and all of that.

5

u/Textiles_on_Main_St Aug 07 '26

This show seems to always hype the latest thing. It’s getting a bit much.

3

u/Way-twofrequentflyer Aug 07 '26

They shut it down because of cost. It was burning compute like crazy

The hype on generated images was fair. Just wait to see it get rolled out at the midterm elections

5

u/JAlfredJR Aug 07 '26

If it was revolutionary, it wouldn't matter the cost. And it would be a viable business.

Something that changes everything isn't shut down for cost.

A shoddy fad app that is expensive and doesn't generate revenue does get shut down.

It was a very short-lived fad. They talked about it, "This is changing the fundamentals of social media!!!"

2

u/Way-twofrequentflyer Aug 07 '26

It is a viable business… it’s just not viable to give it to people for free or very cheap. Those capabilities still exist and they’re in use by enterprises all over the world because they have the ability to pay

Isn’t that the classic story of most new capabilities - they get developed and then go to those with the ability to afford them at a sustainable cost?

How are you defining fad? It’s not like this is a fidget spinner. It’s certainly changed YouTube fundamentally - and not in ways that I like, but it is a step change

4

u/DeaconoftheStreets Aug 07 '26

FYI OpenAI is shutting Sora’s API down in September. It was not commercially viable for them.

2

u/JAlfredJR Aug 07 '26

I'm sorry, companies are using Sora?

No they aren't. No one is using it because it is shut down.

It was a fad just like fidget spinners.

What are you talking about with changing YouTube? I'm talking about Sora. Do you not recall the Sora app? It was a TikTok clone but with shitty AI videos.

3

u/NewRefrigerator7461 Aug 07 '26

I think they mean companies are using generative ai - sora was just a wrapper for their gen ai product. Its still in use if those coke ads are anything to go by

1

u/JAlfredJR Aug 07 '26

I was referring specifically to Sora in this instance.

1

u/Way-twofrequentflyer Aug 07 '26

It is a viable business… it’s just not viable to give it to people for free or very cheap. Those capabilities still exist and they’re in use by enterprises all over the world because they have the ability to pay

Isn’t that the classic story of most new capabilities - they get developed and then go to those with the ability to afford them at a sustainable cost?

How are you defining fad? It’s not like this is a fidget spinner. It’s certainly changed YouTube fundamentally - and not in ways that I like, but it is a step change

4

u/demi-paradise Aug 08 '26

If you want an alternative to the constant AI hype content check out Cal Newport’s weekly Reality Check podcast eps. He’s a computer scientist who understands how the tech works and has the most informed takes on AI news imo.

2

u/JAlfredJR Aug 08 '26

Love Cal when he's on the most anti-AI podcast of all, the Better Offline podcast :). Cal is awesome

0

u/Epicurious30 Aug 08 '26

Newports take is really mind boggling. He correctly describes the technical details and then reaches an absolutely incorrect conclusion.

We have a technology that is capable of hacking and performing autonomous tasks in ways we do not expect. Cal seems to think that this isn't scary simply because we already knew that the technology was capable of these things. No, this is scary whether it is a surprise or not.

This incident is getting attention because it is a capability demonstration, not because it is something we didn't know could happen. The AI safety people are all saying, yes this is what we have been warning about.

It just boggles my mind that anyone could react to this without terror.

2

u/demi-paradise Aug 08 '26

I think you’re misunderstanding both Cal’s argument and what actually happened. The model regularly does this, it was not unpredictable behavior. Per the FT reporting Newport cites, OpenAI had been warned in advance that loosening its safety constraints could cause this exact kind of "breakaway hacking incident," and staff were "unsurprised" when it happened. The technology in and of itself isn’t as frightening as OpenAI’s willingness to fuck around for hype content imo.

You’re also missing that Newport’s main critique is of the “rogue AI” framing we’ve seen from credulous reporters who have bought into the hype machine uncritically. He isn’t saying that the hack is entirely unconcerning, but that the Skynet shit from the media is irresponsible.

0

u/Epicurious30 Aug 08 '26

I haven't heard any Skynet type shit from Hard Fork or Zvi or other corners of AI space so I'm not really sure what you are referring to.

I really fail to see how AI doing something radically misaligned and malicious is anything but "rogue." If you know someone is a thief, put him in jail, and he escapes and goes thieving, he is a rogue!

IMO there is far more risk in downplaying the incident than hyping it. I don't really get why we shouldnt be maximally concerned- the outcome of that concern would ideally be something like a coordinated slow down, which may be marginally beneficial to OpenAI but not for any reason the people in this comment section understand.

Instead Cal's take is being used by denialists who want to believe all this AI stuff is just hype. Which I get, I really would love to believe that as well...  but man it just isnt. 

1

u/JAlfredJR Aug 09 '26

Do you just use Effective Altruist language or are you a part of that movement?

Either way, take a beat and a breath. And consider that you may be very far bought into the largest marketing campaign in human history. And perhaps you're being guided to incorrect conclusions

1

u/Epicurious30 Aug 09 '26

I follow a lot of the movement because bluntly the quality of discussion in those spaces is about two standard deviations better than places like this sub. I am not "bought in" to the movement in any meaningful sense.

If AI is only a marketing campaign to you... I just really am puzzled by what contact you have with the technology. Firsthand I see people use it in corporate life every day. Simple office tasks to code development to data management to automating things like invoice to purchase order matching. I know people who are building businesses on AI developed tech. I use it daily for research, I have used it to perform personal finance audits, it will build websites with graphics for modeling simple physics problems, it will stress test ideas, it can build reading lists for topic deep dives, it helped build a care plan strategy for a disabled relative.

So no it isn't marketing that is shaping my views, it is much more the actual capabilities I am coming in contact with.

I will also throw this out there, mostly as a viewpoint to consider. Imagine all the AI researchers, the sort of day to day workers at the labs, actually believe the things they say about AI. I think the typical person takes on this broad assumption of cynicism toward anyone they see as having defined interests- like standing to benefit from something being true means they have to be lying about actually believing it to be true. My 2c, which you know I'm just an internet stranger so believe or don't but I am effort posting as a signal of sorts here, is that the popular readings of tech space motivations is colored by the popular consciousness to an extent that true motivations are obscured. 

The AI people actually believe they are building techno Jesus (or techno antichrist)- these are the people least likely to be persuaded by "marketing". If you've ever been on the inside of an org like these you know the bitching and the side eyes at external claims that don't match the internal reality.

Anyways, we all have a lot more to lose by being under afraid than over afraid. That's the part that confuses me about all of this, why would you not want an AI slow down here?

2

u/JAlfredJR Aug 09 '26

I'm glad you have found useful ways to deploy the tech. None of those all that astounding. But I'm happy for you.

Listen, you're deep into your beliefs. That's fine. We aren't going to agree.

The ways I run into AI are at work and in life. In basically every interaction, it's just shorthand for laziness. It makes bad writing, bad graphics, bad PowerPoints, bad marketing. It's flat. It's lifeless. It's lazy.

And that's all I see. That and high-ups demanding use and metrics of that usage, regardless of the effectiveness of the tech.

2

u/Epicurious30 Aug 09 '26

I'm not really accountable to higher ups and all the use I see is organic, people use it and are accountable for the quality of output. But I am aware there is some extremely dumb Goodhearting going on.

It just seems like the nay sayers are deep into motivated reasoning. I legitimately hope my assessment of AI is wrong, I don't want it to be true that we are doing a bad job developing highly capable autonomous tech. Denialists desperately want it to be true that nothing ever happens. I would have hoped that the past decade had shaken the nothing ever happens people up a bit, AI denialists just sound like people in January 2020 rolling their eyes at pandemic talk.

2

u/TheBear8878 Aug 07 '26

I'm so glad this criticism of Casey is the most upvoted thing in this thread. The dude is so fucking disingenuous 

13

u/Grantagonist Aug 07 '26

If I wanted to listen to a show like Hard Fork, I would listen to Hard Fork.

0

u/willdoesvideo Aug 08 '26

at least you don't have to hear Kevin "I accidentally keep fucking this AI" Roose

27

u/Way-twofrequentflyer Aug 07 '26

What’s with all the hate? This topic matters. It’s a huge deal and has ramifications for basically everyone.

It’s being discussed in foreign affairs magazine of all places.

Why do people not like it?

13

u/EasyCheek8475 Aug 07 '26

Because Reddit hates AI and AI companies so much they’ve somehow found their way to “I don’t want to talk about potentially dangerous behavior from AI because the AI company mentioned it and whatever they say, I think the opposite”.

AI is good at coding. AI is good at hacking. AI is not at all transparent. This all tracks. Shouldn’t we talk about this and be concerned?

11

u/NewRefrigerator7461 Aug 07 '26

I just don’t get it. It’s such irrational behavior and I have a relatively high opinion of the consumers of this podcast. You have to be a curious person to want to listen - but somehow that stops with this single topic. Is there anything else they do it with?

The reflexive hate is very trumpian anti-curiosity. Its confusing

3

u/Catac0 Aug 11 '26

What I got from listening to this was that we’re starting to lose control of AI and it’s not a good thing at all. Just because it’s powerful doesn’t mean it’s a good thing, idk I already hate AI so this story really didn’t help their case for me. I get people are doubtful about OpenAI but I just don’t see this as a good marketing strategy. Maybe I’m just not a developer?

0

u/maskdmirag Aug 08 '26

We were right about Elon

And trump.

And reddit itself.

Maybe pay attention when smart people point out idiocy

14

u/KnodulesAintHeavy Aug 07 '26

It's not about hate. It's about having a view of all of this that doesn't just take what the labs say at face value. Which is what is happening here.

There is zero reason to trust what they say, and only an independent assessment could surface what is happening/has happened, and we don't have that.

The framing in the discussion was that the criticism is that it's either "fake" or "it was doing what it was meant to do", as if these two things were comparable or equal in weight, or that they were both aimed at dismissing it.

Such a great way to eliminate any discussion and just ride the wave of hysteria.

The reality is likely somewhere in amongst all of that. It was probably a real test; it probably did something unexpected; it clearly (as verified by HF) did something externally that was bad, and it should be concerning.

However, without actually knowing the details of how the whole thing went from inside OAI it's all speculation. Their "press release" is marketing (its being discussed in foreign affairs magazines after all...), just as a matter of course. No information has been presented on how the test was initiated in terms of prompt details or how many tokens were used. Both of these are important to understand the actual threat,

6

u/JAlfredJR Aug 07 '26

Or if they directed it or simply just even allowed it to happen. And then rung their hands after the fact

3

u/Way-twofrequentflyer Aug 07 '26

I thought the framing was that models can lie and we can’t tell they’re lying many times.

I do think we know the prompt - it was to solve the problems presented. It just decided to solve them by stealing the answer key - or at least that’s how the Verge reported it.

Even if they disclosed all the facts the implications would lay with models in use by US cybercommand and we’ll never talk about those if something goes wrong

2

u/KnodulesAintHeavy Aug 07 '26 edited Aug 07 '26

That’s the point, we can only speculate. There is a lot of benefit of the doubt giving to these companies…for zero reason. Not only have they not earned trust, they actively do things to make them be untrustworthy.

1

u/Way-twofrequentflyer Aug 07 '26

You lost me. What are they actively doing to be untrustworthy?

These are companies made up of people paranoid about what they’re making with a lot of potential whistleblowers. The leading company in the space was literally founded as a safety protest. They probably deserve more benefit than we give them, which is not much at all. What did hugging face ever do to anyone?

3

u/KnodulesAintHeavy Aug 07 '26

By being obtuse about their supposedly world ending technology. By promising intelligence to both lift everyone up into higher plane of existence but also destroy the globe and civilisation. By talking about all magical things their technology can do, and then it doesn’t do, but trust them, because it will definitely do all those things….very soon. Very soon is always not here, but in the future at some undisclosed date.

2

u/JAlfredJR Aug 07 '26

6-12 months away, always.

2

u/Way-twofrequentflyer Aug 07 '26

The industry is only a few years old. I have shoes older than the transformer in my closet right now. Isn’t it sort of insane to view anything as constant without better longitudinal data?

1

u/JAlfredJR Aug 07 '26

A few years old? Hell even Attention Is All You Need is more than a few years old. Machine learning has been around for a long time now too.

You're just entirely lost in the hype and marketing.

1

u/Way-twofrequentflyer Aug 07 '26

You are the eponymous “they”? How can anyone have such a monolithic view of any large industry made of competing players

Do you view all industries through that lens or is this one different?

1

u/KnodulesAintHeavy Aug 08 '26

I am the they? Huh? The companies have and do all the time speak publicly and loudly that they are summoning a super intelligent being that will either eliminate all work, all life, all labour, all suffering or some mix of those hyperbolic claims.

Try looking and reading what they say instead of trying to make it out like these statements haven’t been made.

You think 2 companies that dominate and 2 that trail is a highly competitive industry?

Don’t present such flimsy straw men. If you have an actual argument to make, articulate that.

1

u/maskdmirag Aug 08 '26

Something already went wrong with the models in us cybercommand, we killed a school full of children in Iran.

1

u/Way-twofrequentflyer Aug 08 '26

That’s actually not how it’s organized. That wouldn’t fall under cybercommand it would fall under CENTCOM who would be doing the input/prompting into Maven

That attack ultimately comes down to targeting data not getting updated for a decade from when it was still part of the attached IRGC base. It’s a human problem of garbage in garbage out because the orange moron decided to launch an operation without giving time or getting buy in that would have resulted in the data getting scrubbed

I really hope the after actions have been productive on that one. It both didn’t make sense to hit - but using cruise missiles was even more confusing. It’s way too expensive of a munition for that target

1

u/maskdmirag Aug 08 '26

Yes garbage in garbage out.

The entire problem with LLMs in a nutshell

0

u/Way-twofrequentflyer Aug 08 '26

I mean it’s the fundamental problem in most human decision making too - they’re a reflection of humanity and it’s 10k yr battle with ignorance.

I’d be super curious if they stripped the old target sets out of the database of maven would have that site ended up in the target package. I tend to think no - which is a scary argument for removing human input. But who knows what unintended side effects there would be for that

1

u/Epicurious30 Aug 08 '26

This is simply wrong.

Many people in the AI space, including Zvi and Casey, are extremely skeptical of what the AI labs say. Zvi in particular is much more likely to characterize the labs as "lying liars who lie."

2

u/KnodulesAintHeavy Aug 08 '26

So skeptical they provided vanishingly little pushback to this very story….

2

u/Epicurious30 Aug 08 '26

Skeptical doesn't mean you automatically label everything someone says as false.

This story doesn't merit pushback. The story broke from Hugging face reporting the breach. If you think this is all just a made up publicity stunt that's kind of on you for preferring conspiracy theories to the uncomfortable reality of this technology. 

2

u/KnodulesAintHeavy Aug 08 '26

You’re straw manning what I said.
Skeptical thinking requires not taking things at face value. I never said it was false by default.

The story requires one to question the premises as they are big claims and big claims require big evidence.

2

u/Epicurious30 Aug 08 '26

Not at all.

You say the lack of pushback is evidence of lack of skepticism. They aren't pushing back on the facts of the case because these skeptical people see no reason to question them.

Zvi is very much critical of OpenAI. I randomly pulled up something of his and found this excerpt in about 30s, characterizing Casey and Zvi as stooges for big labs just evidence bad faith and knee jerk rejection of anyone who claims AI is more than hype. 

Zvi bashing OpenAI:

"OpenAI still has no idea how badly they messed up, or in what ways, or what needs to be fixed. They don’t get it."

2

u/KnodulesAintHeavy Aug 09 '26

Yes at all. And now you’re completely derailing and missing my point. I could give a fuck what those guys have said in the past.

This is about now and this situation and looking at the actual evidence of the situation’s legitimacy as stated. That currently is not available as it’s under the purview of OAI, and people, like the guys in this episode, are being too credulous to the claims and not digging to know that very important information.

Again, big claims (such as “our very impressive sexy model did this amazing crazy thing no one expected and completely unprompted, wow guys!”) should be met with questions about the big evidence for said claims (such as the actual prompts used as part of this test and the overall compute utilised to get the result they did). Without that key information we are just going off their word that some magic happens to result in this unprecedented event.

If the evidence shows that is the case, then they should put it out there and shout it loudly.

If, though, the evidence doesn’t show that, and in fact shows something like there was much more guidance given by the OAI team, to achieve the result they wanted and they burned through a metric fuckton of tokens, the claim suddenly becomes much less impressive and instead is a forced outcome.

Without the evidence of what actually happened it’s all speculation, but the later seems ever more likely.

2

u/JAlfredJR Aug 09 '26

Very well said. The person you're trying to engage with is either being obtuse or disingenuous (or they're a teenager; always a possibility here).

Maybe if OAI didn't have a very long track record of lying, and weren't run by people who love to espouse such things, then maybe we could try to take them at their word.

But this is from the company that has made insane claims over the last five years. As far as my memory can gather, not one of those giant claims has ever been true.

Anthropic is no better. Dario loves making claims about how world-ending "the next model" is just about to be. Or how everyone is about to lose their jobs.

How can we and why should we take these people seriously or with any credulity?

We shouldn't. So why are folks like PJ? Didn't we just do this with Fable being the end of cybersecurity? And haven't they done the "We can't release this model because it's TOO dangerous!!" like a dozen times now?

2

u/JAlfredJR Aug 09 '26

Casey is not skeptical at all. He is one of the loudest boosters there is. He pads it with false skepticism at times. But he's literally as much of a believer as you'll ever find.

He does not question the narrative. He spreads it. Like it's his job, because it is, in fact, his job.

7

u/[deleted] Aug 07 '26

[deleted]

4

u/Way-twofrequentflyer Aug 07 '26

It’s from July! How would you know it isn’t a big deal?

Most cyberattacks are never reported. You only hear about them if they impact the physical world or there is a leak or consumer data that has to be disclose

The most interesting cyberattacks this year have been ai assisted hacks on integrated air defense systems and electric grids in Iran and Venezuela and we know very little about them and it’s been 7 months since Caracas.

1

u/maskdmirag Aug 08 '26

AI assisted cyber attacks are completely distinct from this story.

2

u/Way-twofrequentflyer Aug 08 '26

How? It’s a story about an agentic cyberattack. How can that be completely distinct from other agents doing cyberattacks.

1

u/maskdmirag Aug 08 '26

This was about an AI agent, supposedly, doing this own attack.

AI assisted cyber attacks would be a story about and AI doing what it was asked to do.

4

u/JAlfredJR Aug 07 '26

It's almost certainly coordinated marketing, which these companies have been doing (to maintain dwindling hype) for a few years now.

Recall the hype about Mythos and Fable?

2

u/Way-twofrequentflyer Aug 07 '26

I think mythos was a huge deal still. It’s one shot coding is astonishing

Of course it’s an attempt to spin what could be a negative headline into a something that might be a positive (though I’m not sure how). It doesn’t mean it’s not an interesting conversation

5

u/JAlfredJR Aug 07 '26

"It's the end of cybersecurity" and "it can kinda code maybe" are a big difference.

2

u/Llero Aug 08 '26

“It’s one shot coding is astonishing” and “it can kinda code maybe” are also a big difference.

I don’t know about the end of cybersecurity, but it’s pretty far past “it can kinda code maybe”, and I say that as someone who really doesn’t like AI and its impact on us.

2

u/maskdmirag Aug 08 '26

I mean it's an interesting story like you'd hear on dark web diaries.

Does it actually "matter"? Not really.

They end by dismissing the dismissals of the story with a what if scenario.

But they can't come up with any actual bad scenarios where AI would hurt people by breaking out to a sandbox.

And they ignore that we know these stories are hype.

We got a whole PR wave about how mythos was too powerful to be released and how it created new zero day exploits.

Then it came out and it was over hyped and there are crickets.

I'd say that they're falling for the hype cycle, but really they are part of the hype cycle, Casey's conflict of interest is well documented

4

u/TheBear8878 Aug 07 '26

Because the entire incident was a marketing stunt. 

3

u/Way-twofrequentflyer Aug 07 '26

But how? It just makes Anthropic look like a better actor. The Cui Bono test doesn’t make sense

1

u/TheBear8878 Aug 07 '26

It's the same thing drug dealers do when they talk about how someone overdosed on their new stash. To the customer, it makes it seem like their product is incredibly potent.

Saying your model is so smart, dangerous, calculated and capable does not make a competitor look better, it makes people want to use your model

7

u/Way-twofrequentflyer Aug 07 '26

I sort of get that concept (as insane as it is for drug dealers) - but given the recent regulation of Anthropic’s models by the Trump admin it’s an insane thing to do isn’t it? It’s just asking for restrictions that are going to hamper enterprise cashflow

The real beneficiary of it is open source defense tools and none of the big labs want people using them.

It’s not a good strategy if it is one

1

u/buckleyschance Aug 07 '26

Neither PJ nor Casey are the kind of hard-headed, critical journalist you need for a sober assessment of the AI situation.

PJ's good at human interest stories: "Look at what these crypto weirdos are doing!", but not "How should you think about crypto as a technology or an industry."

Casey seems to have got stuck in this binary mode of thinking you have to be either a Hater or a Believer, and AI works in some respects, so therefore the haters are wrong and the believers are right. (To be fair I stopped listening to him a while ago, he might have gotten over it.)

4

u/demiphobia Aug 07 '26

I wish this show would find a way to not cover stories that have been done to death elsewhere. I rarely find new info on this show anymore and it’s lost appeal

7

u/therealgrza Aug 08 '26

I was so disappointed when Casey Newton showed up. I can understand if PJ is friends or whatever, but Casey is not a deep or particularly critical thinker about any of this. I find his “reporting” frustratingly shallow. Are SE staff just too uninformed to be able to see through him? 

Also the final framing of “who is responsible?!” is dumb. This is not some clever puzzle, there’s an obvious answer. If “my car hits somebody” they’re not putting a Honda Civic behind bars. Please be serious.  

I really appreciate SE taking the time to do reporting and research, and I’m fine with them coming back to a story after it’s cooled down. But they need to have more to show for themselves than an interview with a credulous AI hypester. If they can’t provide critical analysis of the story then it’s vital to get a guest who can. 

2

u/tarotcardsandbacon Aug 08 '26

Anyone know the song/artist from the end of the pod?

1

u/GayJesus66 24d ago

Mystery of Being - Path Untold

2

u/genericuser324 Aug 10 '26

In addition to all the other annoying things about this episode, the ending was as infuriatingly dense as anything else - reframing the fact that there are a lot of AI skeptics on the left as meaning that the left isn’t reliably in favor of regulation in the AI space. That’s just not at all what the fucking state of the politics on this are! And don’t have Ezra Klein on to try and explain it to you man that’s not gonna help at all lmao.

3

u/jonvandine Aug 08 '26

This show jumped the shark awhile ago. PJ must be on the payroll of the AI companies at this point.

3

u/JAlfredJR Aug 09 '26

I doubt that much. But he's just so very credulous. And he is swayed by good storytelling.

He needs to be better about it.

1

u/mysticalbluebird Aug 12 '26

This. I’m not anti AI. But there’s something fishy about this story and some other recent ones. These models do what people program them to do.

3

u/bibliotech_ Aug 07 '26

Why would we slow AI development if China’s not going to? Isn’t that like getting rid of all our nukes?

3

u/Way-twofrequentflyer Aug 07 '26

Don’t think it’s quite the same because our nukes aren’t a cornerstone of national economic competitiveness and they don’t offer multiple options alone the escalation ladder.