Ranked #6 Everyday Ecosystem — The Leading AI Assistants
xAI

Grok 4.5

Grok is back in the frontier conversation—not by winning every trophy, but by making near-frontier agent work cheap enough to run all day. Grok 4.5 pairs a serious comeback in intelligence with a $2/$6 API price, fast output, Grok Build, Cursor, and Office work. It is the practical ‘use the good model more often’ play.

Updated July 10, 2026 AgenticKnowledge WorkOffice
9.5out of 10
Official Website
Best for

Grok is back in the frontier conversation—not by winning every trophy, but by making near-frontier agent work cheap enough to run all day. Grok 4.5 pairs a serious comeback in intelligence with a $2/$6 API price, fast output, Grok Build, Cursor, and Office work. It is the practical ‘use the good model more often’ play.

Why It Wins

Artificial Analysis scores Grok 4.5 at 54, fourth on its Intelligence Index; it rates the model #3 as a coding agent in Grok Build, on par with GPT-5.5 in Codex. $2/$6 per 1M input/output tokens, $0.50 cached input, roughly 91 tokens/sec, and a $2.49 Coding Agent Index task cost make its performance-per-dollar case unusually strong. Grok Build is the default home, with Cursor and Microsoft Office integrations.

Watch out

This is a cost-adjusted comeback, not an unconditional frontier crown. Grok 4.5 trails the highest raw leaders on several coding tests, is not yet available in the EU, and Artificial Analysis found its hallucination rate rose versus Grok 4.3 even as accuracy improved. Use the value advantage to run verification—not to skip it.

01

What It Actually Is

For a while, Grok’s place in the model pecking order was easy to describe: interesting, fast, occasionally spicy, rarely the tool you reached for when the work had teeth. Grok 4.5 changes that sentence. It does not return holding every benchmark trophy. It returns holding a calculator.

Artificial Analysis puts the model at 54 on its Intelligence Index, fourth behind Fable 5, GPT-5.5, and Opus 4.8. That is frontier-adjacent in the useful sense: strong enough for serious agentic work, not strong enough to pretend the mountain has no summit. Its standout is the deal attached to the intelligence. The API costs $2 per million input tokens and $6 per million output tokens, with a $0.50 cache-hit price. Artificial Analysis estimates $0.31 per Intelligence Index task and measures 91.3 tokens per second. This is the rare release where “cheaper” does not mean “use it only for labeling support tickets.”

The ecosystem also gives the model somewhere to earn its keep. Grok 4.5 is the default in Grok Build, available in Cursor on every plan, and accessible from the xAI API. xAI is pushing it into Office work too: research-backed Excel models, PowerPoint diagrams, and Word drafts. That makes Grok 4.5 a plausible daily driver for the messy middle of knowledge work—research, spreadsheet iteration, analysis, slides, product drafts, and the agent loops that would become too expensive on a premium flagship.

The API and consumer-subscription maths are different. SuperGrok starts at $30 per month: more expensive than the usual $20-ish individual AI plan, but still a fully useful daily product for people who want Grok Build, higher limits, and the consumer experience rather than an API bill. SuperGrok Heavy lists at $300 per month, which is poor value for normal daily work. If the checkout offers the temporary $99/month Heavy promotion, that is the more interesting consumer sweet spot for someone who will genuinely use the larger allowance and heavier agent workflows every day. It is a promotion, not the permanent list price—check the live plan page before treating it as your price.

Grok 4.5’s value is simple: it makes repeated serious work affordable. Research, spreadsheet iteration, analysis, slides, product drafts, Cursor tasks, and longer agent loops all benefit when a capable model is cheap enough to take another pass. That is not a special mode or a narrow use case; it is the practical advantage of getting near-frontier work without treating every iteration like a premium purchase.

There is a lovely trap in all this: taking a good cost curve as proof that the model is reliable enough to stop checking. Do the opposite. Use the savings to buy another test run, another source check, or another agent pass. For a one-shot decision with high consequences, the raw-capability leaders still have the stronger case—but Grok 4.5’s intelligence-per-dollar is what makes disciplined verification affordable.

The honest catch is availability and calibration. It is not yet in the EU at launch, and Artificial Analysis saw accuracy rise while hallucinations rose too. So Grok 4.5 is not the universal champion. It is the smart, fast second opinion you can afford to ask repeatedly—which, for much of real work, is a very powerful place to be.

02

Strengths and honest limitations

Key Strengths

  • A real return to the frontier: Artificial Analysis gives Grok 4.5 an Intelligence Index score of 54, placing it fourth behind Fable 5, GPT-5.5, and Opus 4.8 in its July 8 analysis. That is a 16-point jump over Grok 4.3, not a cosmetic version number.
  • The daily-driver economics are excellent: The API costs $2 input / $6 output per 1M tokens, with cached input at $0.50. Artificial Analysis puts its Intelligence Index cost at $0.31 per task and places it on the cost-versus-performance Pareto frontier.
  • Fast enough to feel like an assistant, not a meeting: Artificial Analysis measured 91.3 output tokens per second on the first-party API. xAI says the model also uses roughly half as many steps as comparable leaders on its selected tasks. Speed is not intelligence, but it changes whether you will actually use the model for a tenth iteration.
  • A useful work ecosystem, not just a chat tab: Grok 4.5 is the default in Grok Build, works in Cursor on all plans, and is available through the API. xAI also highlights Microsoft Word, PowerPoint, and Excel integrations for research-backed spreadsheets, presentations, and documents.
  • Agentic knowledge work has independent signal: Artificial Analysis reports #4 on GDPval-AA v2 at 1543 Elo and a top 33% result on τ³-Banking. Add a 500K-token context window, image input, and configurable reasoning, and this is much more than a fast code completer.

Honest Limitations

  • It is frontier-adjacent, not the raw-capability king: On the same independent Intelligence Index, Grok 4.5 follows Fable 5, GPT-5.5, and Opus 4.8. That earns it a strong value position here, not a fictional claim that it beats every flagship.
  • Accuracy still needs a seatbelt: Artificial Analysis found Grok 4.5’s AA-Omniscience accuracy rose from 35% to 52% versus Grok 4.3, but its hallucination rate also rose from 25% to 54%. Research, finance, and office output still need source checks.
  • EU access is still missing at launch: xAI says Grok 4.5 is not yet available in the EU through its products or API console, with availability expected in mid-July. A great price is not much help in a region where you cannot buy the model.
  • The office story is promising, not magic: Grok Build can make complex Excel models and Office artifacts, but formula logic, slide facts, and permissions remain your responsibility. Treat a first draft as a capable analyst’s draft, not a signed report.
  • Efficiency claims need task-level validation: xAI reports roughly 2× token efficiency and fewer steps on comparable tasks. Those are valuable signals, but your prompts, tools, cache hit rate, and review loops decide the real invoice.
03

Benchmark Snapshot

Artificial Analysis Intelligence Index — 54 (#4)

Independent composite result placing Grok 4.5 behind Fable 5, GPT-5.5, and Opus 4.8 in Artificial Analysis's July 8 release analysis.

GDPval-AA v2 — 1543 Elo (#4)

Agentic professional-work evaluation: Grok 4.5 lands between Opus 4.8 and GLM-5.2 in the cited Artificial Analysis comparison.

τ³-Banking — 33%

Artificial Analysis reports the top result on this agentic banking evaluation, ahead of GPT-5.5's 31% in its release analysis.

Intelligence Index cost — $0.31 per task

Artificial Analysis's cost calculation combines the model's token use and API prices. It supports the value case, but does not predict every workload's bill.

04

The Verdict

Grok 4.5 is the comeback daily driver: not the most universally capable model on every chart, but one of the smartest places to spend an agent budget. It has enough independent intelligence and agentic-work credibility to belong beside the frontier models, then wins the practical argument with speed, $2/$6 pricing, and real homes in Grok Build, Cursor, and Office. Use it for iterative research, business artifacts, and ordinary agent work where you will verify the result. Escalate to the raw-capability leaders for the rare job that truly needs them.

05

Frequently Asked Questions