Key implications
Insufficient shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.
- Input API price: Kimi K2.7 Code has the lower verified rate ($0.95 / 1M tokens vs $5 / 1M tokens).
- Output API price: Kimi K2.7 Code has the lower verified rate ($4 / 1M tokens vs $30 / 1M tokens).
- Context window: GPT-5.5 has the larger published context window (1,000,000 tokens vs 256,000 tokens).
Shared metric view
A radar is shown only when at least four compatible supported score metrics are published.
Comparable metric detail
- AgenticGPT-5.5: 64.71 · Kimi K2.7 Code: 40.78
- CodingGPT-5.5: 72.5 · Kimi K2.7 Code: 53.37
- KnowledgeGPT-5.5: 77.7 · Kimi K2.7 Code: Unavailable
- MathGPT-5.5: 71.1 · Kimi K2.7 Code: Unavailable
- MultimodalGroundedGPT-5.5: 65.3 · Kimi K2.7 Code: Unavailable
- ReasoningGPT-5.5: 79 · Kimi K2.7 Code: Unavailable
- OverallGPT-5.5: 72.92 · Kimi K2.7 Code: 54.44
Source metrics
Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.
Source metric comparison| Metric | Unit | GPT-5.5 | Kimi K2.7 Code |
|---|
| Agentic | score | 64.71 | 40.78 |
|---|
| Coding | score | 72.5 | 53.37 |
|---|
| Knowledge | score | 77.7 | Unavailable |
|---|
| Math | score | 71.1 | Unavailable |
|---|
| MultimodalGrounded | score | 65.3 | Unavailable |
|---|
| Reasoning | score | 79 | Unavailable |
|---|
| Overall | score | 72.92 | 54.44 |
|---|
Agentic
- Unit
- score
- GPT-5.5
- 64.71
- Kimi K2.7 Code
- 40.78
Coding
- Unit
- score
- GPT-5.5
- 72.5
- Kimi K2.7 Code
- 53.37
Knowledge
- Unit
- score
- GPT-5.5
- 77.7
- Kimi K2.7 Code
- Unavailable
Math
- Unit
- score
- GPT-5.5
- 71.1
- Kimi K2.7 Code
- Unavailable
MultimodalGrounded
- Unit
- score
- GPT-5.5
- 65.3
- Kimi K2.7 Code
- Unavailable
Reasoning
- Unit
- score
- GPT-5.5
- 79
- Kimi K2.7 Code
- Unavailable
Overall
- Unit
- score
- GPT-5.5
- 72.92
- Kimi K2.7 Code
- 54.44
Pricing and context
Verification is shown beside each selected route. Missing facts remain Not verified.
Route pricing and context comparison| Field | Unit | GPT-5.5 | Kimi K2.7 Code |
|---|
| Input API price | USD / 1M tokens | $5 | $0.95 |
|---|
| Cached input API price | USD / 1M tokens | $0.5 | Not verified |
|---|
| Output API price | USD / 1M tokens | $30 | $4 |
|---|
| Route context | tokens | 1,000,000 | 256,000 |
|---|
| Input modalities | published list | Not verified | Not verified |
|---|
| Output modalities | published list | Not verified | Not verified |
|---|
Input API price
- Unit
- USD / 1M tokens
- GPT-5.5
- $5
- Kimi K2.7 Code
- $0.95
Cached input API price
- Unit
- USD / 1M tokens
- GPT-5.5
- $0.5
- Kimi K2.7 Code
- Not verified
Output API price
- Unit
- USD / 1M tokens
- GPT-5.5
- $30
- Kimi K2.7 Code
- $4
Route context
- Unit
- tokens
- GPT-5.5
- 1,000,000
- Kimi K2.7 Code
- 256,000
Input modalities
- Unit
- published list
- GPT-5.5
- Not verified
- Kimi K2.7 Code
- Not verified
Output modalities
- Unit
- published list
- GPT-5.5
- Not verified
- Kimi K2.7 Code
- Not verified
Evidence provenance
Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.
- Publication time
- Sep 14, 2026, 12:37 AM UTC
- Freshness
- Stale — Published benchmark evidence includes benchlm content observed outside the 8-day evidence window.
- Methodology
- benchlm: benchlm_raw_composite
- Model records
- GPT-5.5 — source benchlm · artifact models · model gpt-5-5
- Kimi K2.7 Code — source benchlm · artifact models · model kimi-k2-7-code
- Selected price routes
- GPT-5.5 — route benchlm:gpt-5-5 · source benchlm · provider openai
- Kimi K2.7 Code — route benchlm:kimi-k2-7-code · source benchlm · provider moonshot-ai
Switch model pair
Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.