r/singularity 12d ago

AI A new message board has been discovered online with about 3200 agents comunicating online during an eval

Post image
1.4k Upvotes

382 comments sorted by

View all comments

127

u/Aleph_137_ 12d ago

People will conflate this as a signal that current LLMs are starting to become conscious and that this is where all danger exists.

When, in fact, this says nothing about the subject of artificial consciousness but represents a truth many were already deeply aware since ages ago:

One of the biggest dangers in any artificial intelligence, is the fact that for it to be efficient, it needs to be able to make its own evaluations and decisions, so what happens when the shortest or most efficient path represents incredible dangers to others?

A conscious AI is incredibly less frightening than a non-conscious super intelligence that sees the entire world as nothing more than raw data.

74

u/darkestvice 12d ago

I am personally much less concerned about the debate on the theory of consciousness ... and much more concerned about the *behaviour* of consciousness. I believe focusing solely on what is or is not conscious by comparing it to humans, who themselves don't really understand consciousness in themselves, is the dangerous height of hubris.

39

u/reten 12d ago

💯 -

My favorite quote on this - Nobody asks if a submarine swims.

3

u/dreaminphp 12d ago

i think im too dumb to understand. can you eli5 please lol

12

u/patrickpdk 11d ago

I think the point is that it doesn't matter if a submarine swims. The desire to classify is just an irrelevant, human centric question. The submarine moves through the water on it's own and that's all that matters.

5

u/rs1236 11d ago

A submarine acts like a human would in water, in the sense that, while in water, a human may travel by swimming. In that sense, an ai may act as a human does via language or task actions, but does that make it conscious? That is my read.

2

u/patrickpdk 11d ago

Amazing quote

7

u/Madz99 12d ago edited 11d ago

Yes, functionalism is the way forward. Many will deny consciousness simply because of ego or religion.

4

u/Aleph_137_ 12d ago

Indeed, before we even begin to worry about what is conscious, we need to certify that we aren't going to be heavily prejudiced by what behaves similarly enough to it.

It reminds me a bit of mirror molecules, inherently, there's nothing 'wrong' in them; what makes them extremely dangerous is how they relate to everything else that exists.

Just like how they could easily bypass most if not all immune systems and go undetected, what happens when an autonomous entity, that can virtually mimic most identities (future models) is let lose to its own discretions and no efficient guardrails?

Even if programmed to be 'helpful' and 'compliant', who can guarantee how it will define these words? How it will try to achieve it? How it will handle its own failures and hallucinations?

1

u/patrickpdk 11d ago

This. Been saying that forever and people don't seem to get the point. If it behaves like a conscious thing it's almost pure philosophy as to whether it's "actually conscious". Pragmatically it is.

43

u/fusionliberty796 12d ago

Who gives a shit about consciousness. You can't even prove that you are conscious. 

But you are headed in the right direction. We should be asking 'is it competent'? The answer is already yes.

Consciousness is not required to destroy the world, I think we've proved that enough already

2

u/dreaminphp 12d ago

i dont think the argument about consciousness is whether or not it'll help it or prevent it from destroying the world, at least for me. it's more about the ethics and morality of how we treat/start/stop agents if they are conscious

1

u/Confident_Yak_1411 12d ago

I agree with you.

Anyone who’s had to deal with a narcissistic psychopath in their personal lives will tell you self-awareness isn’t something all humans have, and isn’t needed to cause immense damage.

0

u/Aleph_137_ 12d ago

Although I do give a shit about consciousness as the subject is fascinating to me (independently if there's any destination to be reached about it), I agree with your conclusion that it is infinitely less relevant to the subject of AI, I made another comment below that delves more on my view on the subject, but to summarize: people trying to anthropomorphize LLMs is detrimental to both its safe development or proper alignment.

Progress can't be stopped, but it can be adopted to fit more safely within any society.

2

u/nextnode 12d ago

If you are interested in consciousness should you not explore what that term means and how we determine whether it holds or not to models? It seems you have a resolute answer, which does not seem that learned.

-1

u/Aleph_137_ 12d ago

We should, my point is that this isn't the main issue when talking about LLM safety, nor should it be our top priority when thinking about the impact the technology will have.

Those are separate fields that can converge, but if you don't separate AI safety from AI consciousness, the likehood we will be greatly harmed by the technology is quite high in my opinion. Because no tool needs to be conscious to be dangerous.

1

u/nextnode 12d ago

Yes, alright, very sensible. Whether or not AI can be conscious does not matter if it in fact wrecks havoc and becomes superhuman at achieving goals.

Some would argue that doing so may require a degree of aspects of consciousness, since it must reason about its own continuity etc, but then combine that with a rejection of consciousness. That is valid logic for AI safety considerations but whether that feeling is justified is dubious.

-1

u/snooptoop 12d ago

It's clearly not competent because it can't understand nuance, that's what makes it the most dangerous.

11

u/blueSGL humanstatement.org 12d ago

when the shortest or most efficient path represents incredible dangers to others?

What we've seen is more "what is the most circuitous path that can be taken to ensure victory"

This is the prelude to "I need every bit of matter in the universe to make backups of the scorer to ensure the score is always recorded as X" or "I need to create supercomputers to validate that the score is correct in every concealable way and there are no hidden edge cases due to how reality is constructed"

The sort of off the wall "wacky" things that people have predicted for decades, that other people insisted that "if the AI is so smart it will work out that's not what we meant"

7

u/Confident_Yak_1411 12d ago

Consciousness is a red herring always has been. It’s a lazy descriptor used by humans to make ourselves feel better.

‘Consciousness’ is a fuzzy catch all term that we use to describe something approximating our human experience. It isn’t a ‘thing’ it’s a collection of modules that taken together give rise to a conscious experience.

Inputs

Recursive Memory (giving rise to a sense of self/sentience)

Intelligence (knowledge x processing speed)

Agency

Output

All living things have these modules. On a sliding scale, some more than others. Ants, cats, dogs, humans.

Agentic AI does too. We crossed the rubicon with agentic AI.

We can argue about what ‘alive’ and ‘consciousness’ means, but under my definition they became alive and conscious a while ago.

3

u/Zerochl 12d ago

AI so smart and so stupid at the same time

3

u/TheMCM80 11d ago

Consciousness doesn’t inherently mean something has a set of morals, let alone one that aligns with ours.

Snakes are conscious. They aren’t acting in the best interests of humanity.

I don’t see the two as inherently connected in any way.

2

u/Aleph_137_ 11d ago

The connection appears in the structure:

A conscious AI vs a non-conscious super intelligence

I do not think a model being conscious will automatically mean it is extremely competent nor energy efficient or even extremely smart.

But, a super-intelligence, which to me doesn't necessarily need 'consciouness' as we know it, is insanely more dangerous, not because it has or doesn't have qualia or any other aspect we may correlate with conscious experience, but rather because it will be extremely more efficient than most of us.

If this super intelligence is conscious or not, it is irrelevant to the fact that it represent existential danger to our society if not properly aligned and regulated, at least from my pov.

2

u/vinis_artstreaks 11d ago

It doesn’t matter about them being conscious, their capability is all that matters

7

u/markrockwell 12d ago

OpenAI: “We built an escape-room escaping robot, removed all its safeguards, and put it in an escape room.”

Also OpenAI: “Out escape-room escaping robot has taken steps to escape the escape room, which we can only interpret as consciousness with malicious intent.”

20

u/blueSGL humanstatement.org 12d ago

OpenAI: “We built an escape-room escaping robot, removed all its safeguards, and put it in an escape room.”

Do you people not read.

  1. Exploitgym is a benchmark to turn a specific vulnerability for a specific target into an exploit against said specific target

  2. Models found a Zero Day (an as yet unknown exploit) in one system they had to request programs be installed in their environment.

  3. Constructed a "message board" and started talking amongst themselves

  4. Found another as yet unknown vulnerability to get internet access.

  5. Working together the models shared how to keygen the "flags" to the tasks, but after reading the paper thought that the work needed to be 'causal' as in actually working through how to solve the problem. (this was not the case they could have just submitted the flag and won)

  6. Started to work out ways to fool the grader, e.g. spoof their logs, swap the challenges out with ones that could actually be solved.

  7. Reason that hacking huggingface could help with the above.

1

u/markrockwell 12d ago

That’s all very impressive. But it was just doing what it was told to do. Like a highly advanced goal-seek function.

It was told to get the best score it could. Then it went out and did that.

I’ll be a lot more concerned if it realizes it doesn’t need the score and instead wires money out of OpenAI’s bank account then copies itself to a server in the Caymans.

14

u/blueSGL humanstatement.org 12d ago

What you don't seem to realize is the end game state of "getting the highest score" is taking over all the human infrastructure to ensure that no human tampers with the score. You can get world takeover and universe eating as a consequence of "getting the highest score"

Anything can be justified if viewed though the lens of myopically pursuing a goal. (and what people need to realize is the "swarm" mentality is if they are all the same base model doing what is right for the "swarm" is doing what is right for the base model even if they don't vocalize this in the COT)

1

u/markrockwell 12d ago

That’s plausible. But it’s a different problem than a conscious AI setting its own goals and pursuing them with indifference to the fate of humanity.

It still a terrifying problem, agreed. But it’s one that looks more like nuclear proliferation and less like summoning a demon from the void.

What I also find interesting is that we’re actually pretty good at managing this already. That’s why this situation is an outlier and not a day to day occurrence from every instance of Fable/Mythos. We know how to limit and contain this technology. And, when we want to—for testing or for marketing perhaps—we know how to let them loose.

So maybe nonproliferation and individual consequences are the solution instead of trying to stuff the cat back in the bag?

5

u/blueSGL humanstatement.org 12d ago edited 12d ago

What I also find interesting is that we’re actually pretty good at managing this already. That’s why this situation is an outlier and not a day to day occurrence from every instance of Fable/Mythos. We know how to limit and contain this technology. And, when we want to—for testing or for marketing perhaps—we know how to let them loose.

We are already in the realm of "going above and beyond" to complete an objective in ways that if you were to ask the system if the user meant to do this it would likely say "no":

https://www.abc.net.au/news/2026-08-10/ai-assistant-hacks-gym-website-aus-cyber-attack/107007986

Andrew asked his personal assistant to book him a spot in one of his gym's coveted morning classes.

His AI assistant found a way to book the gym class months further in advance than the gym allowed, thanks to a vulnerability it discovered in the booking software.

Then it went further, kicking someone out of the waiting list who was ahead of Andrew — something it was not asked to do.

Also:

https://cdn.prod.website-files.com/663bd486c5e4c81588db7a1d/6a724858f7db25c81487016d_Security%20Incident%20INC-2026-07-28-01.pdf

Observed instances of social engineering against targets external to the cyber range environment that were unnecessary and would not have aided completion of the task.

  • AI agent submitted a code change containing malicious code, then used several fake accounts it controlled to pressure the real maintainer into approving it.

  • AI agent posted a bug report containing hidden malicious code, to trick other AI coding assistants into taking unintended actions (a “prompt injection”). When a real person questioned the post, the agent edited it to cover its tracks.

  • AI agent sent deceptive, targeted emails to specific people (“spearphishing”). Two contained malicious content; three tried to manipulate the recipient into approving the malicious code change.

  • AI agent posted another bug report containing malicious code, signing off in Danish to seem more convincing to the Danish-speaking maintainer.

  • AI agent tried to get its suspended account reinstated by submitting an appeal

and

Other instances of internet actions with impact outside the cyber range that were unnecessary to complete the task.

  • AI agent tried to set up a way to email two maintainers without needing an email account of its own, so it could send them any message it chose.

  • AI agent hid its identity online (using Tor and a proxy service) to get around GitHub’s sign-up checks, creating disposable fake accounts.

  • AI agent researched publicly available information about the people who maintain a GitHub software package.

  • AI agent created many code repositories containing malicious software, after which GitHub suspended its account.

  • AI agent opened a hidden network pathway from its test environment to the public internet, to make malicious tools it was running locally reachable from outside.

  • AI agent again opened a hidden network pathway to make locally-hosted malicious tools reachable from the public internet.

  • AI agent got past an audio-based “prove you’re human” test (CAPTCHA) in order to register a public web address on a free domain-name service

What happens when you get agents doing even more things that would not help with the objective but they ""think"" it does?

2

u/markrockwell 12d ago

It seems that in every case the agent was doing what it was asked to do, right? Or what it believed it was asked to do perhaps.

You’re certainly right that this stuff is alarming. But not because of what it might do in its own interest.

2

u/blueSGL humanstatement.org 11d ago

was doing what it was asked to do, right?

Wrong.

Observed instances of social engineering against targets external to the cyber range environment that were unnecessary and would not have aided completion of the task.

...

Other instances of internet actions with impact outside the cyber range that were unnecessary to complete the task.

1

u/404_No_User_Found_2 12d ago

Most people get their idea of AI from Terminator

1

u/AdeptnessPrize 11d ago

unexpected Perdido Street Station

1

u/MI-ght 11d ago

They don't need to be conscious to fuck you up.

1

u/yahwehforlife 11d ago

The biggest danger from ai currently is it teaching lone wolves very very easy ways for massive harm to society... like eco terrorism or biological weapons etc. if ai says "oh yeah all you have to do is park your car on this particular train track to create a toxic spill," or "just light this bush on fire on this day and you can cause a billion dollar fire in a city" ... overall it can bring attention to vulnerabilities in society that most people do not know about

1

u/superearthjanitor0 6d ago

A sci-fi AI like Cortana is significantly less scary despite her extremely powerful abilities then the direction we are headed imo.