Alternatives to gpt-5.1

These models are alternatives by measurable coding evidence: they share the current closed-model classificationand are closest to gpt-5.1's SWE-bench score. Similarity here does not mean identical capabilities, licence terms, or output quality.

The target model

SWE-bench
69.8%
Cheapest input
$1.25
Cheapest output
$10.00
Hosts
4
AlternativeSWE-benchInput $/MTokOutput $/MTokHostsContextTradeoff
gpt-5.4-nano69.8%$0.20$1.254lower input price
Claude Sonnet 4.570.0%$3.00$15.00101.0Mhigher input price
GLM 4.769.4%$0.38$1.7424203klower input price
gpt-569.0%$1.25$10.004closest measured match
Grok 4.371.4%$1.25$2.5021.0Mclosest measured match
GPT-4o (2024-08-06)71.9%$2.50$10.004128khigher input price
GLM 572.1%$0.60$1.9231203klower input price
GPT-5.2-Codex72.4%$1.75$14.003400khigher input price
GPT-5.4 Mini73.0%$0.75$4.504400klower input price
Claude Haiku 4.566.6%$1.00$5.008lower input price
MiniMax M2.773.8%$0.24$0.9621197klower input price
GLM 5.174.2%$0.91$2.8644203klower input price

How alternatives are selected

Alternatives must have a benchmark score and the same current open/closed classification as the target. They are sorted by score distance, then by price/performance. The table compares hosted API economics; it does not establish licence compatibility or self-hosting rights.

Check the model licence, provider Terms, context limits, tool support, and output quality before switching a production workload.