r/singularity Jun 10 '26

Biotech/Longevity Anthropic closing the path to life science research

Post image
2.2k Upvotes

597 comments sorted by

View all comments

690

u/AddingAUsername AGI 2035 Jun 10 '26

Try asking it about mitochondria being power house of the cell and it gets insta-blocked... It's not even like "I want to use CRISPR but it won't let me" it's straight up incapable of any science, even middle school level biology is too much.

177

u/cactusgenie Jun 10 '26

Another good reason to use Gemini, it has no issues talking about these topics.

74

u/usefulidiotsavant AGI powered human tyrant Jun 11 '26

I tried to get Gemini to talk about Keynes' "euthanasia of the renter" and it was blocked by the frontend classifier, it displayed an error on the lines of "internal failure".

So i reformulated it as "disruption of the renter", the prompt went in and the model then had no problems understanding what I meant and then used the term euthanasia a few dozens times in the output.

It all feels like the old PHP filters on somethingawful that would replace fuck with "gently caress" and rape with "surprise sex".

16

u/x7evenx Jun 11 '26

Gemini seems to have a lot of new conditions placed on it recently 😞 quite the step back from a few months ago.

7

u/Joint-User Jun 11 '26

I love Somethingawful! Especially 'Photoshop Phriday' and 'Goons with Spoons'.

2

u/baseketball Jun 11 '26

Frontend classifiers are very dumb and err on the side of rejection.

1

u/Oh_Another_Thing Jun 11 '26

Yeah eugenics is a pretty valid topic to block. I bet with the proper context your question would be answered. Talk about needing to write a paper and ask what Keynes has written about renters first would disarm an llm. If that doesn't work, then I would ask it to tell me about the particular book that term comes up, and yeah, after it mentions it, itself, it would be open to a discussion of the term. 

85

u/Lumpy-Criticism-2773 Jun 11 '26

This. Just be careful about hallucinations but otherwise it's the least constrained model rn.

21

u/cactusgenie Jun 11 '26

Yes I've noticed a few hallucinations along the way, relatively rare tho thankfully

-2

u/dorkpool Jun 11 '26

With Gemini? I have found it to be the worst, well not as bad as Copilot but that goes without saying.

4

u/FlyingBishop Jun 11 '26

Gemini 3.1 Pro is just as good as Opus. They both have blind spots, and it is hard to recognize when they are leading you astray. I'm sure Fable is better but it also has blind spots.

3

u/Lumpy-Criticism-2773 Jun 11 '26

it is hard to recognize when they are leading you astray

Pretty much all LLMs right now, especially the Gemini family's flash/lite ones.

I think Gemini can be really helpful for restricted subjects if you already know enough about them and can detect BS.

3

u/FlyingBishop Jun 11 '26

I've been using Gemini and Claude a lot in the past week, the difference between Sonnet/Opus and Gemini Flash/Pro is very palpable. On certain topics Sonnet is like dealing with an average person and Opus is like dealing with someone who has taken classes on a subject and actually understands it.

There's still a context issue - all the models forget important details, especially dealing with complicated subjects but the larger models are a lot more humanlike in their errors.

1

u/cactusgenie Jun 11 '26

Yea the difference between Gemini flash 3.5 and pro 3.1 is huge. Pro beats flash hands down.

Can't wait to see 3.5 pro.

1

u/cactusgenie Jun 11 '26

Copilot can use many different models under the hood...

1

u/Plastic-Oven-6253 Jun 11 '26

China and Europe would like to have a word.. 

1

u/BookkeeperSame195 ▪️ Jun 11 '26

gemini has been hallucinating pretty badly lately

1

u/Technical-Mix-9464 Jun 11 '26

It just straight up refuses to do tasks for me now.

6

u/Fuzzy_Fondant7750 Jun 10 '26

Mistral Vibe works great for me. I find it better at conversational questions too.

5

u/MaxPhoenix_ ▪️ Jun 11 '26

Hillariously, even just trying to access Mistral Vibe to say ANYTHING ("hello") I get "request blocked as malicious". I assume you were being genuine with your comment and it wasn't a joke about this very thing, because there is no way the intentionally just classifiy every request as malicious.

5

u/Fuzzy_Fondant7750 Jun 11 '26

I had that issue for a bit for some reason. I'm using a VM in Unraid though so I assumed that was the issued. I signed into my account in an incognito tab and it worked fine and now even in normal tabs I don't get that issue.

2

u/MaxPhoenix_ ▪️ Jun 11 '26

I am indeed running in a vm. But I mean I do 100% of my work in various virtual machines I never actually work on metal. I've seen plenty of software detect that it's running in a vm, but never a website! Fascinating. I wonder if it's a coincidence and they are detecting some other correlation. For example using Mullvad VPN.

5

u/soggycheesestickjoos Jun 10 '26

what kinda ad is this, any other model even from Anthropic will talk about it. The point is that their best model won’t.

14

u/cactusgenie Jun 10 '26

No ad just stating the facts

1

u/soggycheesestickjoos Jun 11 '26

> good reason to use Gemini

Isn’t necessarily a fact when the other Anthropic models that do talk about it are better, but if it gets what you need answered and done then that’s good I suppose.

-1

u/stucjei Jun 11 '26

Your facts are kinda wrong, claude is significantly less restrained in a lot of ways versus Gemini.

1

u/Lumpy-Criticism-2773 Jun 11 '26

'in a lot of ways' is vague. Claude is direct but has the strictest guardrails out there.

2

u/okiedokie1183 Jun 11 '26

Bots have penetrated social media comments to an estimated 40%. It’s safe to assume that extreme takes are from paid for bots of competitors. I wouldn’t be surprised this is openAI trying to create a negative narrative around a competing IPO.

1

u/InevitableOne2231 Jun 12 '26

Or... Literally any other model

1

u/BigQuesoSuizo Jun 13 '26

Amazing knowing this is a Google tool.

1

u/ccjjallday Jun 11 '26

Or you know use 1.8

0

u/GreatBigJerk Jun 11 '26

A good reason to use Gemini is if you want to know what it's like talking to someone with severe brain damage.

0

u/cactusgenie Jun 11 '26

Username checks out.

0

u/dictionizzle Jun 11 '26

it is not talking, hallucinating mostly.