Grok 4.3
xAI released 2026-04-30 proprietary
Gemini 3.7 Flash is both better and cheaper.
It scores +18.1 higher and costs 52% less ($0.750/M against $1.56/M) on this workload — and it does everything this model does.
DeepSeek V4 Flash 0731 is cheaper still (93% less) but drops no image, file input.
Every price dimension
| Input | $1.25/M |
|---|---|
| Output | $2.50/M |
| Cached input | $0.200/M |
| Web search | $0.005/call |
Past 200,000 tokens the price changes. Input goes to $2.50/M (2×) and output to $5.00/M. The headline rate does not apply to a long-context workload.
Independent scores
| Intelligence | 37.9 |
|---|---|
| Coding | 42.2 |
| Agentic | 24.2 |
| LMArena Elo | 1397.5default |
Sources are listed separately rather than averaged. Across the 62 models both have scored they correlate at r = 0.806 — close agreement overall, but four models rank very differently between them, and a blended score would hide exactly those.
Best rankings by task
- pptxslides #9of 14 1071
- agentichtmlslides #10of 16 1066
- agenticslides(html) #10of 16 1066
- agenticslides #10of 16 1072
- agenticslides(python-pptx) #10of 16 1068
- agenticgamedev #20of 34 1010
- python-pptxslides #21of 35 1072
- htmlslides #22of 34 1024
Capability
| Context window | 1M |
|---|---|
| Input modes | text, image, file |
| Tool use | yes |
| Reasoning | optional |
| Open weights | no |
Provenance
| Price source | openrouter.ai |
|---|---|
| Fetched | 2026-08-24 |
| Quality data | verified |
| Cross-checked | vendor page |
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...