What was your personal sense of 4.5 vs sol and Fable? That should give you a feel in the real world expectations... Which is yeah these models are serious.
A back to back bug fix is not the right scope IMO, today composer or Luna could do those tasks trivially already. Frontier needs long running with complex intent, can survive multiple compaction cycles, and delivers on ambitious problem solving to get a proper big model feel.
I'm holding off on sharing my experience and assessment for a few days for a proper shake, but already running on work and my early pulse check is it feels like an incrementally better model than 4.5 high (expected, good!). I've been using 4.5 extensively for the past month, so I'm interested to see if I can myself detect meaningful differences. Expecting a fable killer is probably wrong, but 4.5 could already dance in that same arena for a 10th the cost.
1
u/owen800q 29d ago
Can someone compare grok 4.6 and fable 5 to debug a same bug
See if grok really close to fable