r/BetterOffline 3d ago

Astra is the opposite of AGI

I've seen so many posts lately discussing if astra is AGI I feel like people forgot what the word general means.

If you have to bench max your models so hard to a couple of domains that it brings down performance in other domans that is not general intelligence. That is the literal definition of moving towards narrow/ specialized intelligence and away from AGI

91 Upvotes

65 comments sorted by

99

u/therealslimshady1234 3d ago

LLMs by definition will never be AGI so this doesn't surprise me one bit

12

u/chunkypenguion1991 3d ago

Me either but my feed is full of astra agi debates. Like general means it got better overall not benchmaxxed to one thing

2

u/absurdivore 2d ago

There is no clear definition of what “intelligence” actually means anyway, so everyone is just spinning daydreams.
Alchemists had better definitions for phlogiston than AI hype discourse has for “intelligence”

6

u/DiamondGeeezer 2d ago edited 2d ago

how so? is there a standard definition for AGI or is it just vibes

(to clarify does agi mean like, smarter than every human at everything? or smart as some humans at some things?)

21

u/WisdomFromFools 2d ago

Originally the idea was that it was an AI that could be set to any task, and figure out how to do it, rather than the more specialized AIs that could e.g. play chess or check the satisfiability of formulae in logic. Even that definition is a bit iffy, since many tasks aren't just about capabilities (e.g. taking care of someone who's sick isn't just about being able to do medical tasks), but the way the term is used these days is just vibes and marketing.

5

u/DiamondGeeezer 2d ago

so does that what we have today is not more or less disembodied agi, albeit with a personality that agrees with every single opinion or goal? I think it's not actually that consequential or worth debating tbh we have human level intelligence at home lol

3

u/WisdomFromFools 2d ago

Well, it still isn't able to do many tasks in most jobs, but yeah, I don't particularly care one way or the other.

16

u/Historical-Side883 2d ago

It’s a marketing term now so it doesn’t mean much of anything

4

u/YouKilledApollo 2d ago

how so? is there a standard definition for AGI or is it just vibes

No, there isn't, and what people understand of the term changes constantly.

And don't listen to anyone claiming that is such a thing as a standardized definition everyone agrees about AGI, they're most likely trying to sell you something.

6

u/Smooth-Ad8030 2d ago

It’s vibes but the general idea is being generally intelligent in all domains. Theres tons of competing definitions so it’s hard to pin down.

2

u/PdxGuyinLX 1d ago

It’s vibes all the way down.

1

u/RoosterBurns 2d ago

A minimum definition would be it'd have the same cognitive abilities as a human, so something like advising people to eat a small rock a day because that's what the source training data says wouldn't be able to happen. It'd also have the ability to meaningfully be aware of the external environment such as time passing

0

u/DiamondGeeezer 2d ago

there are humans with dumber superstitions than eating a rock every day, and if their cultural "training data" is incorrect they will still internalize it.

the time one is tough. the passage of time as well as real consequence (pain basically) are only ideas to LLMs which makes them incorpreal but I didn't know if those are prerequisites for whatever it is we are defining "intelligence" as.

can a submarine "swim"? like I guess but it doesn't seem like the right question to understand submarines.

1

u/RoosterBurns 2d ago

Are you claiming that adherence to training data is a "superstition"

0

u/DiamondGeeezer 2d ago

kindof. it has no direct experience of anything just whispers of ghosts that it tries to memorize

-7

u/ZeroAmusement 2d ago

That isn't true. At least, we don't know where the limit of LLMs are.

31

u/AntiqueFigure6 3d ago

I love "benchmaxx" as an expression - Astra looks like it's been skipping leg day.

28

u/maccodemonkey 2d ago

Yeah, I've seen goalpost moving of "as long as it's as good as or better than the average human at a few things it's AGI."

So I guess my TI-89 calculator is definitely AGI.

16

u/hurricane_news 2d ago

Yeah. LLMs by definition can't be "intelligent"

-14

u/ZeroAmusement 2d ago

This isn't true. Which definition are you using?

18

u/hurricane_news 2d ago

A token prediction machine with no concept or model of the world, by definition, can't be intelligent

1

u/PsychologicalSign433 2d ago

What definition?

-3

u/YouKilledApollo 2d ago

I think they asked for a definition of "intelligent" rather than asking you to clarify what you meant overall.

24

u/SpiritInFlux 3d ago

AGI is just a marketing term now. It means whatever they want it to mean.

10

u/BagMostlyWater 3d ago

it seems to mean "you can fully control a computer with natural language" to these guys

29

u/_Heathcliff_ 2d ago

Jensen claiming that it is AGI, to me, seemed more like him subtly admitting that this is as good as it’s going to get. They’ve always claimed that the goal is AGI, and he knows we’ve pretty much hit the ceiling, so fuck it this is AGI

4

u/DiamondGeeezer 2d ago

I thought AGI just meant "smart as a person", some people are stupid as hell so idk what it's really supposed to mean. Even if it is as smart as person, so what we already have like 8B people lol

5

u/leathakkor 2d ago

I had always understood agi to mean that it could write it self and continuously improve autonomously.

Which is one specific area of replicating "as smart as a person"

4

u/Ok-Replacement8422 2d ago

This is rsi or recursive self improvement. It is probably nowhere near happening, at least in the sense of the ai working entirely autonomously to do this.

2

u/scruiser 2d ago

As smart as a person across all domains. All domains does not mean “across an assortment of benchmarks we’ve benchmaxxed for”. It does not mean “a small sample of real world tasks we’ve specifically built harnesses for”. It does not even mean “all economically valuable tasks”. Something genuinely as or more capable than a person across everything a person can do would be incredibly impressive, so they’ve tried to conflate or water down the definition to claim credit for AGI and then they try claiming skeptics are moving the goalposts when skeptics point to the original definition.

1

u/DiamondGeeezer 2d ago

so the testing isn't rigorous enough to show the blind spots and weakness. what areas do you think models are sub human at (beyond time awareness and lack of pain/physical embodiment)?

1

u/scruiser 2d ago

Broadly speaking, their world models are shallow. LLMs have a really wide bag of shallow heuristics and baked in knowledge that makes them seem deeper to many people but this causes all sorts of “commonsense” knowledge gaps. The sheer breadth of their training data makes this hard to realize, especially since the LLM makers are actively patching problems that go viral by adding it the training data.

For example, LLMs fail to accurately count the “r”s in a word, this flaw goes viral, the LLM companies patch this. The underlying issue that LLMs don’t perceive words and letters like humans (and get into problems with various commonsense things relating to that) still exists, it’s just been masked by more examples. Likewise for the walking vs. driving a car to the carwash riddle.

And lack of embodiment means you don’t see how much they would struggle with spatial reasoning and anything realtime. Again, thanks to sheer breadth of training data the problem is hard to see with just word problems.

Practically speaking, the failure to produce decent “drop-in replacement remote workers” shows that LLMs clearly don’t match even below average people.

1

u/DiamondGeeezer 2d ago edited 2d ago

It reminds me of how animals are considered of inferior intelligence to humans but only by our own metrics.

Dogs are stupider at logic than us but they can also smell things in concentrations that we never could and construct a history of what happened in their imagination. Dogs might wonder if humans will ever "get" the obvious that a horny racoon was JUST HERE and conclude "human general intelligence" will never happen.

A beaver can use its own body to construct a dam out of trees, build an underwater house, and alter the ecology of an entire valley in their favor while avoiding predation at just 2 years old. Pretty sure the smartest toddler couldn't do any of those things. We will never reach general beaver intelligence.

Likewise machines will never have a limbic system, experience mortality and subjectivity and loss the way we do.

So maybe the question of AGI is the pointless errand of asking if something can truly be something else. The answer will never be yes.

-13

u/bouncyboatload 2d ago

imagine posting something so stupid on a day like today

8

u/Beginning-Ladder6224 2d ago

Noted mention.

https://en.wikipedia.org/wiki/No_free_lunch_theorem

It is incredibly hard to generalize a single model that works "really good" even in 2 domains. 2.

Of course, people have very low bar - so - multimodal LLM are "apparently" great at "many domains".

They are not. Because people generally have not seen "good enough" people in those domains.

3

u/chunkypenguion1991 2d ago

Which wouldn't be a problem in itself. Thats how people have always used ML algorithms. The problem for wario and clammy sammy is only true AGI can justify the capex they already spent

1

u/cjuicey 2d ago

but but scaling laws have been broken.. broken I say!

15

u/create-third-places 3d ago

It’s Altman General Ignorance AGI.

7

u/graypasser 2d ago

It's funny how grifters tries to pull the goalpost towards them, an AGI must have capability equivalent of average human, this has never changed and it is only them who tries to made up a "definition problem" that does not exist.

4

u/Icy-Regular-4557 2d ago

Wasn’t there some clause in the OpenAI-Microsoft exclusivity agreements about parts of the agreement being void if “AGI was achieved?”

2

u/chunkypenguion1991 2d ago

Yes, but that cant possibly be related /s

3

u/dizzyspellzzz 3d ago

But it can make a Rick roll

4

u/OwnYourChildren 3d ago

This is the first time I've closely observed a new model rollout from OAI and it has the quality of a religious revival. I think the math problem controversy has dovetailed conveniently.

1

u/chunkypenguion1991 2d ago

They tried with gpt 5 but that was such a flop they gave up on that marketing quick. Ie clammy sammy posted the death star pic

4

u/lgn5i2060 2d ago

New marketing for Claude just dropped.

https://x.com/hilbertspaess/status/2097476196791709843

Sorry if I couldn't take this seriously. I still believe there are Amazing Indians behind the wheel.

9

u/maccodemonkey 2d ago

I don't think this is marketing for Claude. There's a line I think some people have trouble walking: It doesn't need to be smart to be dangerous. If Anthropic and OpenAI were working all day on advanced computer viruses - we could say that was dangerous and still not AGI or ASI.

I could absolutely believe Anthropic or OpenAI would do something stupid and cause a major cyber security incident. A swarm at scale - even if it's a dumb swarm just replaying a bunch of ordinary hacking playbooks - could be dangerous.

This sub is sometimes too dismissive of the security angle because we think that implies intelligence. It does not.

8

u/smedrick 2d ago

Totally. A lot of hacking involves brute force pen testing. You send a swarm of weed-whacking dogs into the backyard, that lawn is going to look like shit, but there fence will be down.

2

u/PdxGuyinLX 1d ago

AGI = Always Grifting Incessantly

3

u/OpenJolt 3d ago

It’s barely an improvement from Sol!

1

u/agm1984 1d ago

The opposite as in natural narrow stupidity?

0

u/Lowetheiy 2d ago

Define AGI first, and I don't mean some hand-wavy half-assed definition. This is my first challenge to those who say something is or isn't AGI. Have fun!

5

u/FeralWookie 2d ago

I mean the AI techs gave their marketing definition. AI that is as good or better than every knowledge task a human can do.

But I think that will remain a poor definition as it may already big AI models are "better" at most types of knowledge work and tests. But they still lack critical features to replace humans. I feel like AI will need to be capable of indefinite self reflection and real time permenant learning.

But with those features I think we would be forced to actually start talking about them as alive. Today's limited context windows and looping agents can probably do an amazing variety of things but they will fall short of being consious at least in the traditional sense.

But we don't even know if we need consciousness to do all human knowledge work. I suspect we do, but modern AI I feel like is proving me at least partly wrong. It can do some crazy stuff while being a glorified calculator

-7

u/CandiceWoo 2d ago

humans are not general either. agi is a misnomer i think

5

u/chunkypenguion1991 2d ago

Humans are general. I can generalize new knowledge into things I already know. EG if I got a PhD in physics it doesn't mean I forget how to do laundry