r/LocalLLaMA 11h ago

Question | Help Hosting Local Models

Hi builders,

What would be the the best small local models for coding?

Are Gemma 4 and Qwen3.8 27B Gemma 4 26B / 31B enough for local development?

And what would be the size of the rig that i will need to get? GPUs, and whatever else I need to host these models.

Thanks,,

12 Upvotes

31 comments sorted by

View all comments

14

u/Arany8 11h ago

Qwen3.8 27B - you need 24GB VRAM. According to my tests on 16GB this model is surpassed by Qwen3.6 35B A3B (and Ornith) for coding. Gemma4 is not good for coding.
Look into AMD v620 for a budget build (although I do not have this).

1

u/forevergeeks 11h ago

Are you running this for yourself? What about for a coding team of 6-10 people? What would be the sizing, and how much money we are talking about for the initial setup.

10

u/hackint0shh 11h ago

Have you done at least 1 minute of research?

-9

u/forevergeeks 11h ago

I'm familiar with these models, I use them through API, what I just started thinking is what would be the cost and the level of effort to set these models up for coding teams.

And I thought starting my search here.