Top
Best
New

Posted by meetpateltech 10 hours ago

Grok 4.7(x.ai)
501 points | 416 commentspage 7
finnjohnsen2 6 hours ago|
Is Grok relevant? Who uses it?

Maybe I'm in some kind of bouble but I have never met or talked to anyone who has used Grok.

Brendinooo 5 hours ago||
I bought a month of use for $100. Early impressions of Grok 4.6 was that it's good at talking about ideas (got me unstuck from a piece of writing that I was working on; that bought it a TON of goodwill) and just okay as a coding tool compared to my more extensive use of Fable and Opus. Not as smart as Fable but cheaper; not as capable as Opus 5 but way less annoying along the way. And it generates images, which Anthropic doesn't do.

Not sure if I'll hold the subscription but I could see myself working with it more.

chronogram 5 hours ago|||
I use it in the car to talk to because it's built-in, mostly for navigation via voice. It seems to be an old or quantised model, because it seems a few years old, and it has strict rate limits unlike Gemini, where it really just shuts off if you talk to it a few times in one day, but it's still fun to show passengers who are still new to LLMs.
cvwright 4 hours ago|||
I tried it last month after hearing here that it didn’t speak in Claudisms.

As a chatbot it’s totally fine, virtually indistinguishable from Gemini or ChatGPT or Claude.

For coding it’s… okay. I tried 4.6 and it feels similar to Opus from 12 months ago, or maybe Sonnet from 9 months ago. YMMV.

moomoo11 5 hours ago|||
[flagged]
dofm 5 hours ago||
The only people I know of in the UK who will tell you that they use it are performatively alt-right or right wing. The kind of people who say "nanny-state" or "wokerati". GB News viewers. People who have an opinion on Meghan Markle that they think other people need to hear.

It's just an observation but so far a pretty solid correlation. Musk has so severely poisoned the well in terms of his UK reputation that the only people who are open about using Grok are... well, wankers is as good a word as any.

FWIW among the AI-using people, it mostly goes Claude Code, then Codex, then whatever runs on their Mac. The only Cursor user I knew has jumped ship to OpenCode.

soaaa 5 hours ago||
[dead]
enraged_camel 8 hours ago||
This thing is DOA. They compared 4.7 xhigh to 4.6 high to make it look like it improved. The reality is pretty bad: https://x.com/chetaslua/status/2102087511367618942
outside1234 7 hours ago||
Who uses this trash?
spiderice 6 hours ago|
117 million people per month. Anything else I can google for you?
mrtesthah 3 hours ago||
Is that the number of X users who inadvertently viewed Grok-generated child pornography?
spiderice 2 hours ago||
Not sure. I've never seen Grok-generated child pornography. I take it you're intimately familiar?
BoumTAC 8 hours ago||
Vals AI just affirm that Grok 4.7 is worse than Grok 4.6 (It ranks #24 on the Vals Index at 54.2%, down 5.0 points from Grok 4.6 (#14, 59.2%))

https://x.com/ValsAI/status/2102086608476590432

nostrebored 8 hours ago|
Is this an ad for Vals AI? Looking at their website, the rankings don't mesh with my observed utility for almost any model outside of fable and astra being good-ish.
BoumTAC 8 hours ago||
Absolutely not. I know Elon retweet them a lot when Grok is good. This is how I discover the company.

I like to follow them and look for benchmark for each LLM release.

nostrebored 7 hours ago||
Ah gotcha, not on twitter so just hadn't seen them before!
zug_zug 9 hours ago|
Well I "tried it out" I asked it one question, and it gave no answer and said "Sign up to use more!" I don't think I'll be doing that, no.

I can't think of a single dimension grok is winning on (capability, cost, voice), but want to stay open-minded -- anybody want to vouch for its capabilities in any domain?

sejje 9 hours ago||
If you haven't used it, how do you know if it's winning?

I think it's winning on UI for normies (grok bot) and they made some claims about being pareto SOTA (lowest cost per task completed) a while back with 4.6.

I find it to be a perfectly capable model for implementation (there are many in this class--deepseek flash, spark1.3, luna, etc). I find the usage to be very generous w/ supergrok. I find the model to be just fine for 90% of what I want to do, but I use a smarter model to plan complicated things.

swalsh 9 hours ago|||
After the cursor aquisition it's become a quite capable coding model. If you take cost into account, it's close to the top. OpenAI is maybe still #1, but I'd put Grok at #2 (again, including cost as a factor).
grim_io 9 hours ago|||
It's probably the most aligned (to a single person) model out there!
puszczyk 9 hours ago||
For me it works well for agentic coding tasks and terminal/unix/bash (in cursor and grok build); it's also token efficient and cheaper than gpt 5.6. It's def not as good as Fable for me (I haven't used Astra much, can't comment). So it's not the cheapest, not the most capable, but it has a good mix of it for my backend, go, infra work.

The voice is the same AI slop as the others imho.

(This is about Grok 4.6, I didn't test 4.7 yet).

edit: clarified I mean agentic coding tasks

Shekelphile 2 hours ago|||
> it's also token efficient and cheaper than gpt 5.6.

Deepswe results show that grok 4.6 is more expensive per-task and consistently scores worse than: luna xhigh, glm 5.3, astra low, sol high/xhigh, opus 5 medium.

Grok also used almost 3x as many tokens/turns to complete tasks than all of those models (besides luna), so it takes way more time to complete a task.

There isn't much reason to use Grok at all, it's gotten better but it's still worse than every other player in the field, which shouldn't be a surprise considering until about a year ago they were just buying tokens from other providers and pretending it was their own model.

With gpt-6 luna and sol coming tomorrow it's going to look even worse too, especially if new luna retains the same dirt cheap pricing that 5.6 luna has.

svachalek 8 hours ago|||
The voice is the weird part. The early Grok 4 models had a very distinct presentation unlike anything else out there. Then suddenly it made a big jump in coding ability and started sounding just like every other model.
puszczyk 7 hours ago||
[dead]