Hacker News new | past | comments | ask | show | jobs | submit
Apparently Grok 4.7 has 40% more weights than Grok 4.6, but the price ($6 output token, $2 input) is the same.

Given that the decrease in their margin and the fact they delayed the release of Grok 4.7 almost two weeks past the original date, XAI must not have been happy with the results for 4.7. And XAI also waited the day before Opus 5.5 is rumored to launch. I imagine Opus 5.5 will blow Grok 4.7 out of the water benchmark wise.

However, I have become skeptical of benchmarks. Grok 4.5 solved some issues setting up a buildroot system that Fable 5 couldn't do. I find the post cursor groks are phenomenal at frontend web development, though Claude is much better at backend ruby.

My favorite part of the new Groks has been how they speak in plain english. I simply cannot stand Claudish. Or even GPT, which doesn't have Claude's ticks but definitely likes to handwave explaining technical concepts. Still, nothing beats Claude 3.5 and 4 with explaining since it seems all models have regressed. I wonder if Grok 4.7 will also regress with English because of all the RL.

loading story #49793659
loading story #49790348
loading story #49789415
loading story #49789842
loading story #49790136
loading story #49790628
loading story #49790733
loading story #49792144
loading story #49789516
loading story #49792304
loading story #49790561
loading story #49794210
loading story #49790381
loading story #49790626
loading story #49792695
loading story #49794686
loading story #49794110
loading story #49789857
loading story #49791151
loading story #49790143
loading story #49791989
loading story #49794116
loading story #49789728
loading story #49792318
loading story #49792291
loading story #49792138