← Back to dashboard
SpaceXAI·LLM — weighted competency·Tue 25 Aug 2026

Grok 4.6

Grok 4.6 scores 98/100 competency for LLM at $3.00.

Moderate confidence
Competency
98/100
category rank #1
Cost
$3.00/1M tok
blended
Speed
27/100
Slow
Accuracy
no data yet

Score history.

Weighted LLM competency · last 12 weeks · weekly snapshot.

98/100
— flat
94
16 Aug19 Aug22 Aug25 Aug

What it's actually good at.

Score on each business-task subset that goes into the weighted competency.

Customer Service20% of category weight
97
Content Writing15% of category weight
97
Research & Analysis16% of category weight
98
Summarization12% of category weight
98
Translation7% of category weight
99
Reasoning16% of category weight
96
Knowledge14% of category weight
76

Pricing.

Published usage pricing. Figures refresh daily from our pricing sources.

$3.00

Usage pricing

$3.00/1M tok
Input
$2.00 / 1M tokens
Output
$6.00 / 1M tokens

Effort modes.

Same model, different reasoning effort — the competency/cost trade-off across modes.

ModeCompetencyCostSpeed
Default98 / 100$3.0027 / 100
Xhigh97 / 100$3.0026 / 100
Medium96 / 100$3.0029 / 100
Low86 / 100$3.0024 / 100

Where it sits.

Position against the rest of the category. Ranked by weighted competency.

RankModelScoreCostSpeedΔ vs Grok
01Grok 4.698$3.0027 / 100
02Claude Opus 597$10.0027 / 100-1
03GPT-5.6 Sol96$8.0031 / 100-2
04Claude Fable 595$20.0033 / 100-3
05Gemini 3.7 Flash94$1.5060 / 100-4
06Kimi K394$6.0018 / 100-5
07Qwen3.8 2.4T A95B92$3.0024 / 100-6
08Qwen3.8 Max89$3.0017 / 100-9
09GPT-5.589$11.2540 / 100-9
10GPT-5.6 Terra88$4.5057 / 100-10
11DeepSeek V4 Pro 081387$1.9849 / 100-11
12DeepSeek V4 Flash 073184$0.6678 / 100-15
13Claude Sonnet 583$4.0042 / 100-15
14Gemini 3.1 Pro Preview81$4.500 / 100-18
15GLM-5.280$2.1540 / 100-18