head-to-head

OpenAI: GPT-5.3-Codex vs Qwen: Qwen3.5 Plus 2026-02-15

Side-by-side comparison of specs, pricing, benchmark scores, and task rankings. Updated 2026-09-23.

OpenAI: GPT-5.3-Codex Qwen: Qwen3.5 Plus 2026-02-15
Vendoropenaiqwen
Quality Score100100
Benchmark Score53.3-
Input Price$1.75/M$0.26/M
Output Price$14.00/M$1.56/M
Context Window400,0001,000,000
Max Output128,00065,536
Tool Calling
Structured Output
Reasoning Mode
Vision
Audio--
Benchmark Scores
ai_index53.6-

Who wins by task?

TaskOpenAI: GPT-5.3-CodexQwen: Qwen3.5 Plus 2026-02-15
SQL Generation 138 133
Code Review 142 132
Code Completion 117 131
Code Refactoring 144 136
Bug Fixing 146 136
Unit Test Generation 130 124
Code Documentation 133 131
Regex Writing 122 119
CI/CD Pipelines 126 120
Frontend Component Design 128 122
Data Analysis 132 124
CSV / Spreadsheet Cleanup 134 133
ETL Scripting 135 128
JSON Extraction 120 131
Bulk Data Labeling 117 129
OCR / Document Parsing 133 131
Table Extraction from PDFs 133 131
Long-Document Summarization 145 137
Short-Form Summarization 115 123
Blog Post Writing 126 121

Scores reflect capability match + benchmark data + pricing for each task. Methodology →

Related comparisons

Anthropic: Claude Haiku Latest vs OpenAI: GPT-5.3-Codex Anthropic: Claude Sonnet Latest vs OpenAI: GPT-5.3-Codex ByteDance Seed: Seed-2.0-Lite vs OpenAI: GPT-5.3-Codex ByteDance Seed: Seed-2.0-Mini vs OpenAI: GPT-5.3-Codex Google: Gemini 3.1 Flash Lite Preview vs OpenAI: GPT-5.3-Codex Google: Gemini 3.1 Flash Lite Preview vs Qwen: Qwen3.5 Plus 2026-02-15 Google: Gemini 3.1 Flash Lite vs OpenAI: GPT-5.3-Codex Google: Gemini 3.1 Pro Preview Custom Tools vs OpenAI: GPT-5.3-Codex

The Model Movers Report

One email every Friday, built from this site's own rankings: the current top five by benchmark score, every model released in the last seven days, and one note worked out from that week's numbers. You can unsubscribe from any issue with one click.