Kimi K3
Widely considered the strongest open-weights model for general quality and complex agentic tasks.
- active
- 2026-07-01
- unknown
- 1M
- text,code
- tool use, API, local use
- 2026-07-20
Reviewed use cases
- complex agentic tasks
- coding
- reasoning
Recorded limitations
- Smaller ecosystem than Llama/Qwen; limited third-party evaluation
Benchmark Performance
64.4
| Benchmark | Score | Provenance | Date | Source |
|---|---|---|---|---|
| SWE-bench Verified | ~70% | Vendor | 01/07/2026 | Source |
| LMSYS Chatbot Arena | ~1300 | Independent | 01/07/2026 | Source |