比特币橙子Trader
比特币橙子Trader|Sep 21, 2026 16:57
Whoa, Grok 4.7 is here, and it’s only been a little over a month since Grok 4.6. The new model switched to a larger base model and extended training specifically to tackle tasks that require working for several hours straight. Major improvements were made in coding, long-context understanding, self-checking, documentation, and agent tasks. The most insane upgrade this time is in Terminal-Bench, jumping from 20.3% straight to 38.0%, nearly doubling. CursorBench went from 40.4% → 46.3%, and DeepSWE improved from 65.2% → 71.0%. Plus, it’s starting to show some weird “specialized advantages”: Legal tasks at 19.6%, while GPT-5.6 Sol is only 2.5% and Fable 5.1 is 6.7%. Electrical engineering at 64.0%, also outperforming both competitors. However, coding isn’t universally the best—DeepSWE is still slightly below GPT-5.6 Sol, and Terminal-Bench is noticeably lower than Fable 5.1. But honestly, the craziest part is the price. Grok 4.7 switched to a larger model, with capabilities improving across the board, yet the API pricing remains at $2/M input and $6/M output, exactly the same as Grok 4.6. In SpaceXAI’s own comparison, GPT-5.6 Sol is $4/$20, and Fable 5.1 is $10/$50.
+3
Mentioned
Share To

Timeline

HotFlash

APP

X

Telegram

Facebook

Reddit

CopyLink

Hot Reads