# Grok 4.7 costs the same, but bills double

Published: 2026-09-27

Grok 4.7 costs exactly what Grok 4.6 did per token, $2 in and $6 out. So why did the independent bill per task double, and is it really "twice as fast"? A one-week-later follow-up. Requested by @therealhaas. Verdict: NEEDS REVIEW.

Canonical: https://thedailydiff.dev/video/2026-09-26-grok-follow-up/

## What this video covers

- Same price, double the bill?
- Why does $2 / $6 now cost $3.74 a task?
- Twice as fast? The clock says 57 tokens a second
- Lazy or hard-working? 81,000 tokens a task
- Where did the Copilot+ PC go?

## Chapters

- 0:00 Same price, double the bill?
- 0:48 Why does $2 / $6 now cost $3.74 a task?
- 1:40 Twice as fast? The clock says 57 tokens a second
- 2:19 Lazy or hard-working? 81,000 tokens a task
- 3:01 Where did the Copilot+ PC go?
- 3:29 A month with zero AI: $30 of tokens for nothing?
- 3:54 The perfect 5 out of 5: how humble is Grok?
- 4:33 Verdict: would I ship Grok 4.7?

## Transcript

### Same price, double the bill?

0:00 You'd think the same price means the same bill. On Monday I told you Grok four point seven costs exactly what Grok four point six did. Per token, it does. Per task, it costs double. In this video, three questions. Why does the same price cost double? Is it really twice as fast, like the launch post says? And is it the laziest model around, or the hardest working?

0:20 Also, this one is by request. The real Haas asked under Friday's video, can you do a vid on Grok four point seven, colon D. We replied, easy peasy, later today. By the time you watch this, it's probably tomorrow, so we shipped late. Relax, so did Grok, about two weeks late, if you believe Hacker News. And one test gave Grok a perfect five out of five. It's my favourite number of the week, and I'll get to it at the end.

0:44 It's Saturday, September twenty-sixth, and this is The Daily Diff.

### Why does $2 / $6 now cost $3.74 a task?

0:48 Here's the trick. The price per token didn't move, two dollars in and six out, per million. What moved is how much it talks. Artificial Analysis ran its whole test suite, and Grok four point seven wrote about two hundred forty million tokens. That's two and a half times its predecessor, and almost three times a typical model. So the average cost per task went from about a dollar eighty-six to three

1:10 seventy-four. Same price tag, double the receipt. It's a taxi with the same rate per mile that takes the scenic route through two other cities. Hacker News spotted it in the first hour. The biggest thread under the launch says forty percent more weights, same price, and almost two weeks late, so xAI can't have loved the results. Another commenter asked about the chart on top, which compares against GPT five point six and quietly leaves out Astra. That, he wrote, can't have been an oversight.

1:39 Now, the launch post.

### Twice as fast? The clock says 57 tokens a second

1:41 The headline says twice as fast, at half the price of comparable models. A few lines down, it says served at the same price and speed as Grok four point six. So it's twice as fast as somebody else's model. Artificial Analysis clocked it at about fifty-seven tokens a second, slower than the old Grok, and summed it up in two words. Notably slow. Meanwhile, xAI's own account, which now goes by SpaceX AI, posted the calmer version. A notable improvement over Grok four point six, at the same price and speed.

2:12 Twenty-eight thousand likes. So the tweet and the blog headline disagree, and for once the tweet is the careful one. Which brings us to the strangest part.

### Lazy or hard-working? 81,000 tokens a task

2:20 Artificial Analysis says it works harder than the last Grok, about eighty-one thousand output tokens per task, more than double the last one. And on long office work and in its own coding harness, it really did improve. It's fourth among coding agents now, behind two Claudes and GPT six Astra, and that puts xAI in the top four labs. The people using it describe a different model. One Hacker News commenter says Grok ends tasks almost immediately and claims

2:46 done, and calls it the laziest of them all. Another watched it loop in thinking mode, from fix one all the way to fix eighty-one. So it's lazy and verbose at the same time. It's the coworker who writes an eighty-one step plan, then says done and goes home.

### Where did the Copilot+ PC go?

3:01 One more straight diff before the fun number. Microsoft has quietly retired the Copilot plus PC brand. The Surface boss told Windows Central the new Surface PCs are not called Copilot plus PCs, even though they meet every requirement. Two years after a launch whose headline feature, Recall, had to be delayed for security, the name lives only on spec sheets. Windows Central's own take is shorter. These PCs have nothing to do with Copilot.

3:28 And one developer tried the opposite of Grok, a month with zero AI.

### A month with zero AI: $30 of tokens for nothing?

3:32 His post hit the Hacker News front page. The low point before he quit. An agent sat stalled for thirty minutes, he told it off, it apologised and delivered in twenty seconds, and the bill for that half hour was thirty dollars of tokens for absolutely nothing. A month later he says the joy of programming is back, and he isn't scared of being fired. Grok would call thirty dollars a warm up.

3:53 Which brings me to the perfect score I promised.

### The perfect 5 out of 5: how humble is Grok?

3:56 Personality Bench runs every new model through a stack of personality tests, and on honesty and humility, Grok four point seven maxed out, five out of five. Its profile is literally called the humble type. To be fair, thirty-two other models also maxed it out. The same page says it believes powerful others control what happens to it, and flags that as unusual for a frontier model. I wouldn't call that a training artifact.

4:20 I'd call it reading the org chart. And according to the birth chart on the same page, it's a Virgo. If you'd rather read this than hear me say it, the diff lands in your inbox every morning, free at the daily diff dot dev, link below. So, today's verdict on Grok four point seven.

### Verdict: would I ship Grok 4.7?

4:35 NEEDS REVIEW. I'd use it for long office work and inside its own harness, but I'd put a spending cap on it first, because the price didn't change and the bill did. Subscribe, hit the bell, and tell me in the comments if you'd have stamped it differently. And the real Haas, this one was yours. Requests are open. And that's the diff for today. I'm Niko from Axrisi.

4:53 Merge responsibly.

## Sources

- [SpaceXAI — Introducing Grok 4.7 (archived)](https://web.archive.org/web/20260923060447/https://x.ai/news/grok-4-7) — web.archive.org
- [Hacker News (609 pts)](https://news.ycombinator.com/item?id=49788838) — news.ycombinator.com
- [SpaceXAI on X](https://x.com/SpaceXAI/status/2102069815225586149) — x.com
- [Artificial Analysis — Benchmarking Grok 4.7](https://artificialanalysis.ai/articles/benchmarking-grok-4-7) — artificialanalysis.ai
- [Artificial Analysis — Grok 4.7 model page](https://artificialanalysis.ai/models/grok-4-7) — artificialanalysis.ai
- [Artificial Analysis — Grok 4.6 model page](https://artificialanalysis.ai/models/grok-4-6) — artificialanalysis.ai
- [Personality Bench — Grok 4.7](https://persona.earthpilot.ai/models/x-ai/grok-4.7) — persona.earthpilot.ai
- [Windows Central — The Copilot+ PC brand is dead](https://www.windowscentral.com/microsoft/windows-11/the-copilot-pc-brand-is-dead-microsoft-and-pc-makers-quietly-pull-back-on-tarnished-windows-11-ai-pc-branding) — www.windowscentral.com
- [Hacker News (90 pts)](https://news.ycombinator.com/item?id=49854945) — news.ycombinator.com
- [bustikiller — One month without AI](https://blog.bustikiller.com/2026/09/25/one-month-without-ai.html) — blog.bustikiller.com
- [Hacker News (159 pts)](https://news.ycombinator.com/item?id=49855018) — news.ycombinator.com
