Claude Opus 5
Agentic knowledge work · Complex reasoning
EVIDENCE PROFILE ↗Compare leading proprietary intelligence with open-weight control and deployment flexibility.
OPEN CUSTOMIZABLE VERSION ↗Selection is stored in the URL so this exact comparison can be bookmarked or shared.
Agentic knowledge work · Complex reasoning
EVIDENCE PROFILE ↗Open deployment · Agentic workflows
EVIDENCE PROFILE ↗| DECISION FACTOR | Claude Opus 5Anthropic | Kimi K3Moonshot AI |
|---|---|---|
| QUALITY EVIDENCE | ||
| Intelligence Index | 61 | 57 |
| Coding Agent Index | NOT RECORDED | NOT RECORDED |
| GDPval-AA v2 | 1861 Elo | 1668 Elo |
| Output speed | 55.7 tok/s | NOT RECORDED |
| COST + SCALE | ||
| Input / 1M tokens | $5 | NOT RECORDED |
| Output / 1M tokens | $25 | NOT RECORDED |
| Context window | 1M | Not disclosed |
| Access | Proprietary | Open weights |
| CAPABILITIES | ||
| Reasoning | YES | YES |
| Vision input | YES | YES |
| Tool calling | YES | YES |
| Structured output | YES | YES |
| Fine-tuning path | NO | YES |
| Self-hosting | NO | YES |
| DEPLOYMENT | ||
| Hosting | Anthropic API / cloud partners | Open weights / hosted APIs |
| Privacy posture | Provider API; enterprise controls | Self-host or provider API |
| Strongest fit | Agentic knowledge work · Complex reasoning · Professional deliverables | Open deployment · Agentic workflows · Knowledge work |
Benchmark values are evaluation-specific snapshots, not universal quality scores. Prices can change and long-context or tool use may add fees. Validate provider documentation and run your own workload before production commitment.