head-to-head

Anthropic: Claude Haiku Latest vs SpaceXAI: Grok 4.20

Side-by-side comparison of specs, pricing, benchmark scores, and task rankings. Updated 2026-09-23.

Anthropic: Claude Haiku Latest SpaceXAI: Grok 4.20
Vendor~anthropicx-ai
Quality Score100100
Benchmark Score-50.6
Input Price$1.00/M$1.25/M
Output Price$5.00/M$2.50/M
Context Window200,0002,000,000
Max Output64,0001,800,000
Tool Calling
Structured Output
Reasoning Mode
Vision
Audio--
Benchmark Scores
ai_index-42.3
eqbench-55.8

Who wins by task?

TaskAnthropic: Claude Haiku LatestSpaceXAI: Grok 4.20
SQL Generation 129 142
Code Review 124 146
Code Completion 114 121
Code Refactoring 124 150
Bug Fixing 128 150
Unit Test Generation 120 133
Code Documentation 122 139
Regex Writing 118 125
CI/CD Pipelines 116 129
Frontend Component Design 122 129
Data Analysis 124 133
CSV / Spreadsheet Cleanup 125 138
ETL Scripting 120 139
JSON Extraction 122 123
Bulk Data Labeling 120 120
OCR / Document Parsing 127 134
Table Extraction from PDFs 127 134
Long-Document Summarization 125 151
Short-Form Summarization 114 118
Blog Post Writing 117 130

Scores reflect capability match + benchmark data + pricing for each task. Methodology →

Related comparisons

Anthropic: Claude Haiku Latest vs Anthropic: Claude Sonnet Latest Anthropic: Claude Haiku Latest vs ByteDance Seed: Seed-2.0-Lite Anthropic: Claude Haiku Latest vs ByteDance Seed: Seed-2.0-Mini Anthropic: Claude Haiku Latest vs Google: Gemini 3.1 Flash Lite Preview Anthropic: Claude Haiku Latest vs Google: Gemini 3.1 Pro Preview Anthropic: Claude Haiku Latest vs Google: Gemini 3.1 Pro Preview Custom Tools Anthropic: Claude Haiku Latest vs Google: Gemini Flash Latest Anthropic: Claude Haiku Latest vs Google: Gemini Pro Latest

The Model Movers Report

One email every Friday, built from this site's own rankings: the current top five by benchmark score, every model released in the last seven days, and one note worked out from that week's numbers. You can unsubscribe from any issue with one click.