Grok 4.3

170 points207 comments5 hours ago
sundarurfriend

As an English-as-second-language speaker and writer, one thing Grok really shines at is capturing the tone and level of "formality" of a piece of text and the replicating it correctly. It seems to understand the little human subtleties of language in a way the other major providers don't. Chatgpt goes overly stiff and formal sounding, or ends up in a weird "aye guvnor" type informal language (Claude is sometimes better but not always).

Grok seems in general better at being "human" in ways that are hard to define: for eg. if I ask it "does this message roughly convey things correctly, to the level it can given this length", it will likely answer like a human would (either a yes or a change suggestion that sticks to the tone and length), while Chatgpt would write a dissertation on the message that still doesn't clear anything up.

Recently I've noticed that Grok seems to have gotten really good at dictation too (that feature where you click the mic to ask it something). Chatgpt has like 90-95% accuracy with my accent, the speech input on Android's Gboard something like 75%, Grok surprisingly gets something like 98% of my words correct.

show comments
artdigital

Grok is my favorite model for chatting, and my favorite voice mode. It seems to be the only voice mode that isn't routing to a extremely cheap model (like Haiku), and has been the highest quality out of all the frontier ones. When you subscribe to SuperGrok you can also create a "council" of agents, each with their own system prompt and when you ask something, they will all get asked in parallel to come to a conclusion. Good stuff!

Just wish they would finally put some work into their apps, it's the only thing keeping me from actually subscribing to SuperGrok:

- No MCP / connected apps support. It's been teased but here we are, still not available. I can't connect Grok to anything, so I can't use it for serious work

- Projects are still not available in the app so as soon as you move something into a project, it's gone from all the native apps

- No way to add artifacts (like generated markdown docs) directly to a project, we have to export to PDF/markdown and re-import. And there isn't even a way to export artifacts. This makes serious project work hard because we can't dynamically evolve projects with new information

- No memory, no ability to look up other chats, each chat is completely new

- No voice mode in projects at all

If someone from xAI is reading this, please consider adding some of these.

show comments
tornikeo

So, we have: - claude for corps and gov - codex for devs - grok for what, roleplay, racism? Those are the two things I've ever heard grok associated with around me.

show comments
Barbing

Grok 4.3 was completed ahead of its CEO’s lesson on this common safety resource:

  Asked if he knew anything about OpenAI's "safety card," Musk smiled and replied: "Safety card? Why would it be a card?"
https://www.axios.com/2026/04/30/musk-openai-safety-grok

Low relevancy in spite of cluster size and musical chair gas generators for time being:

  Later in his testimony, Musk was asked about a claim he made last summer that xAI would soon be far beyond any company besides Google. In response, he ranked the world’s leading AI providers, saying Anthropic held the top spot, followed by OpenAI, Google, and Chinese open source models. He characterized xAI as a much smaller company with just a few hundred employees.
https://techcrunch.com/2026/04/30/elon-musk-testifies-that-x...

(Affiliated with no AI company, just surprised to read this yesterday - how could Elon miss model cards…concerning…, & the fact money can’t buy success every time.)

show comments
xiphias2

It's just at the Chinese levels for coding, so right now it's just a money earing thing for investors.

I hope the Cursor guys help them catch up to be closer to frontier models because they badly need help in it.

show comments
ezoe

While the tread is swapping between "OMG Claude good. OpenAI was done for" and "OMG Codex good. Anthropic was done for". I've never heard about Gemini and Grok. It works mostly similar performance, but people don't mention that much.

Still, my impression is, Gemini hallucinate too much while Grok is always less capable than competitors so it's not worth using it.

show comments
maz1b

I still wish they named it something else, but congratulations to the team on what seems to be a good release!

Pricing is also quite surprising, compared to comparable competitors. I guess they have tons of capacity or really want to bring over more people.

show comments
netdur

In court vs openai, Musk said Grok is partly trained on openai models, so it should be somehow similar to Chinese models in terms of performance and cost!

mythz

Ok speed (202.7 tok/s) and value (1.25 -> 2.50) look great, with pretty decent intelligence.

show comments
ragchronos

When looking at the benchmarks, this model seems to be really close to Kimi K2.6 in terms of intelligence and pricing, hitting that sweet spot. It does also have a higher AA-Omniscience index, which is something kimi and other open models lack in. Curious to see how pleasant it is to use.

show comments
samagragune

Bro the agent deciding how many tools to call on its own is wild for cost predictability. Who's approving that bill?

alyxya

Despite their attrition, this combined with their cursor partnership is likely going to make them competitive in coding agents soon.

mirekrusin

All those plans from providers should be sliders – prepay more, get more in return.

agunapal

Very competitive price for the speed and intelligence being offered!

kilroy123

People are going to hate on Grok because of Musk. However, I do hope they're successful in making a powerful model. We desperately need more competition. I want cheap subsidized AI plans.

I hope Meta finally comes around, too. I want those sweet, sweet billionaire subsidized tokens.

show comments
OtherShrezzing

The tok/s stat is interesting. Since the dominant constraint on inference speed is hardware, it suggests X purchased far more compute than was really needed to serve the demand for their models.

Expensive miscalculation.

show comments
BoredPositron

Yay, free tokens. I don't know why but grok always seems good fast in the free token phase and after that degrades.

Imustaskforhelp

Pelican riding a bike here: https://gist.github.com/SerJaimeLannister/f6de26bd0d0817e056...

(ran this on arena.ai direct chat and also tried to write this gist inspired by how simon writes his gists about pelicans)

Edit: just realized that I made pelican riding a bike instead of bicycle, which now makes sense as to why it hardened the bicycle to look tankier, going to compare this with pelican riding a bicycle if anybody else shares the pelican riding a bicycle.

show comments
sexylinux

Is this now a reliable product or will it still produce errors?

happosai

I lost the trust in them when they added the racist "what about killing of Boers in south Africa" thing to their system prompt.

No way am I going to use a model where the backing has such blatantly obvious brain washing goals.

show comments
alfiedotwtf

If there was any model I wouldn’t trust, it wouldn’t be the ones from China, it would be the one from Elon Musk

show comments
khalic

This project is a gigantic waste of resources, it’s fine tuned on politics of the CEO, was used for CSAM generation and just sucks overall

show comments
gigatexal

How do the grok models fare in coding challenges to say gpt 5.5 and opus 4.6/4.7?

I hate giving Elon any money. The man is a net negative to society but … if the models are objectively better then logically I must no?

show comments