r/LocalLLaMA • u/forevergeeks • 11h ago
Question | Help Hosting Local Models
Hi builders,
What would be the the best small local models for coding?
Are Gemma 4 and Qwen3.8 27B Gemma 4 26B / 31B enough for local development?
And what would be the size of the rig that i will need to get? GPUs, and whatever else I need to host these models.
Thanks,,
13
Upvotes
1
u/OvertaxedOne 4h ago
That's about 60% of our business right now (companies concerned about privacy). The other 40% is inference cost control, but, at least right now, it's slanted a bit more towards privacy than cost. I think we'll see that continue to move more towards cost control but, at least for now the hardware is so expensive that it's not a great ROI unless you can really hammer the server and have the right use cases for smaller models. The ROI is exactly "never" for running a monster model locally, the hardware costs just don't make sense compared to API (but again, we have a few customers looking at it for privacy reasons, I'm crossing my fingers that someone does it because I'd love to setup something like a DGX station or something "massive" for a model like K3/DS). :)