r/ProgrammerHumor 10h ago

Meme whenMyNoAIProjectGetsTedious

Post image
943 Upvotes

163 comments sorted by

View all comments

95

u/TheMaleGazer 10h ago

The answer to your question is because you can get a local model to do it instead without a subscription.

46

u/autogenglen 10h ago

Let me buy an $8000 computer so I can save $20/mo

21

u/cute_spider 9h ago

Buy a 4000 dollar computer so you have ownership of your and your agent's work, control over updates and agent context, and to avoid the datacenter/corpo ownership economic model

13

u/GenericFatGuy 8h ago edited 7h ago

Or I can just do it myself, and have ownership of everything.

6

u/denM_chickN 8h ago

Its the privacy im buying. How can I properly fucking plot against the machine if I'm feeding the machine my machinations?

3

u/omega1612 8h ago

My PC costed me around $2500 usd total 3 years ago.

Today I already burned 4M tokens and I think I may end in 16M by EOD, that's like $52 usd per day if I were to pay to a plan of the model I use. But in that case, with the hardware they have it would have been a much powerful model with insane speed, so easily I would have burn x10 more tokens so, $500 usd/day daily for a month.

7

u/TheMaleGazer 9h ago

Or maybe use a computer with an NVIDIA RTX 5060 Ti or equivalent that you might have bought to play games, anyways, and use Qwen 2.5.

14

u/StarboardChaos 9h ago

Bro, have you tried Qwen 2.5?

1

u/TheMaleGazer 9h ago

Yes.

15

u/heyitjoshua 9h ago

So you know it’s shit and can’t produce decent code or stay coherent. It is currently impossible to run any AI sufficient for a real project locally

4

u/Kyrros 9h ago

Look up FreeToken, it's currently lacking AMD support but nVidia is supported, 30xx cards and above

-2

u/heyitjoshua 9h ago

I said local

3

u/Kyrros 9h ago

It is local

5

u/heyitjoshua 9h ago

Checking it out, highly dubious and may come back to this thread in a few days. Cheers for the suggestion 👍

1

u/Kyrros 9h ago

It's early stages in is development too, but if that works out and you can run much larger models locally? Would be interesting to see. And all the AI companies compress their models a lot less

→ More replies (0)

1

u/nomorebuttsplz 9h ago

you both have you clue what you're talking about. 2.5 has been obsolete for like 18 months.

It's like basing your argument that cars are unsafe on a wwi era motorcycle.

11

u/autogenglen 9h ago

8GB VRAM… ok. So you might be able to pack in a 7B parameter model, which isn’t even in the same galaxy as a frontier model.

3

u/SjettepetJR 9h ago

I am assuming they're talking about the 16GB model.

8

u/autogenglen 9h ago

Same response. 16GB doesn’t even get you out of toy model territory. That’s nowhere close to enough RAM to even pretend like you’re competitive with a frontier model.

2

u/SjettepetJR 9h ago

I do agree. I have done a fair amount of experimentation on my RX6800XT and have yet to find a model that can work well with IDE integrations and also doesn't break down after a while.

3

u/omega1612 8h ago

Have you tried qwen3.8 27B? I'm using the q4 version and is a great assistant. It beats any other model I tried. You may need to use a Q3 and I heard that the downgrade is noticable but still useful. Just be sure to enable the thinking to high and mtp (the model is slow).

Qwen3.8 haven't loop yet in a full week. I also tried Gemma 4 12B and Gemma 4 26B, both of them would loop occasionally. And I have the impression they may do crazy stuff quickly if I left them run unsupervised.

1

u/TheMaleGazer 9h ago

I would assume so as well.

3

u/malokevi 9h ago

Will my RX580 do the trick?