Alternatives to Grok 4.6
These models are alternatives by measurable coding evidence: they share the current closed-model classificationand are closest to Grok 4.6's SWE-bench score. Similarity here does not mean identical capabilities, licence terms, or output quality.
The target model
SWE-bench
95.6%
Cheapest input
$2.00
Cheapest output
$6.00
Hosts
4
| Alternative | SWE-bench | Input $/MTok | Output $/MTok | Hosts | Context | Tradeoff |
|---|---|---|---|---|---|---|
| GLM 5.3 | 95.4% | $1.40 | $4.40 | 2 | 1.0M | lower input price |
| GPT-5.6 Terra | 95.4% | $2.00 | $12.00 | 6 | 1.1M | closest measured match |
| Claude Fable 5 | 95.0% | $10.00 | $50.00 | 10 | 1.0M | higher input price |
| gpt-5.6-sol | 96.2% | $2.00 | $10.00 | 6 | — | closest measured match |
| Claude Opus 5 | 97.0% | $5.00 | $25.00 | 10 | 1.0M | higher input price |
| GPT-5.6 Luna | 93.0% | $0.20 | $1.20 | 6 | 1.1M | lower input price |
| Claude Opus 4.8 | 88.6% | $5.00 | $25.00 | 10 | 1.0M | higher input price |
| Muse Spark 1.2 | 86.6% | $1.25 | $4.25 | 2 | 1.0M | cheaper, lower coding score |
| Grok 4.5 | 86.6% | $2.00 | $6.00 | 2 | 500k | closest measured match |
| Claude Opus 4.7 | 83.5% | $5.00 | $25.00 | 10 | 1.0M | higher input price |
| gpt-5.5 (<272K context length) | 82.6% | $5.00 | $30.00 | 6 | — | higher input price |
| Muse Spark 1.1 | 82.0% | $1.25 | $4.25 | 2 | 1.0M | cheaper, lower coding score |
How alternatives are selected
Alternatives must have a benchmark score and the same current open/closed classification as the target. They are sorted by score distance, then by price/performance. The table compares hosted API economics; it does not establish licence compatibility or self-hosting rights.
Check the model licence, provider Terms, context limits, tool support, and output quality before switching a production workload.