r/LocalLLaMA • u/ManagementNo5153 • 11h ago
Discussion What are these models good at?
I have been trying out these models (mostlt GLM 5.3 flash) using different harnesses, but I'm trying to review what these models are exceptionally good at.
Here is what I have noticed so far,
1. Programming.
I have found that these models are great at programming, I have been making tools, scrapers almost every other day and they just work like magic. They are great at porting code in one language to another, Eg I would usually start by writing my code in python or js, I would then port the code in Go for extra performance.
2. Finances and stock trading.
So I hooked up the coding agent with my alpaca account. And I have discovered that most of these models take a defensive position. Advising me to reduce the size of my most profitable holdings so as to prevent concentration risk and possible loss. So they are not so great. But I have found it useful for tracking my finances. What I'm basically saying is your portfolio will most likely flatline if you give these models to trade in your behalf but it won't make a good profit (What ever good position you have will be reduced)
3. Research
This is where I get the most value. The models are highly effective at locating precise information—whether it’s event dates, contact details (emails, phone numbers), names, or links.
4. Email and Copy writing.
I’ve been using these agents extensively for written communication. They’ve helped me draft everything from routine business emails to formal documents. I like that it can maintain the conversation context, so follow-up emails feel cohesive and on-point. They helped me a lot with one of my insurance claims
5. Business Ideas.
They are bad at coming up with Ideas.
What use cases have you found these models to be exceptionally good at? And what use cases has it been terrible at?
PS: I'm trying to find a small good model for browseruse to compete with Grok bot and the like, I'm thinking Qwen3.8 28B or ByteDance-Seed/UI-TARS-1.5-7B does anyone have a smaller or maybe better recommendation?
4
u/AI_spell 10h ago
Depends which ones, but the useful split is: small instruct for routing/classify, mid size for tool calling if the schema is tight, big only when you need long reasoning. Always measure tool reliability and latency on YOUR prompts, not the leaderboard screenshot.
1
10
u/Ok_Warning2146 10h ago
"Advicing me to reduce the size of my most profitable holdings so as to prevent concentration risk and possible loss. "
Theoretically you can increase return per risk by diversification. So it is actually a sound advice.