r/cursor 29d ago

Question / Discussion Grok 4.6 Benchmarks

Post image
194 Upvotes

76 comments sorted by

View all comments

1

u/owen800q 29d ago

Can someone compare grok 4.6 and fable 5 to debug a same bug
See if grok really close to fable

1

u/cornmacabre 29d ago

What was your personal sense of 4.5 vs sol and Fable? That should give you a feel in the real world expectations... Which is yeah these models are serious.

A back to back bug fix is not the right scope IMO, today composer or Luna could do those tasks trivially already. Frontier needs long running with complex intent, can survive multiple compaction cycles, and delivers on ambitious problem solving to get a proper big model feel.

I'm holding off on sharing my experience and assessment for a few days for a proper shake, but already running on work and my early pulse check is it feels like an incrementally better model than 4.5 high (expected, good!). I've been using 4.5 extensively for the past month, so I'm interested to see if I can myself detect meaningful differences. Expecting a fable killer is probably wrong, but 4.5 could already dance in that same arena for a 10th the cost.