Rank #6 / 13
Kimi K3
Kimi K3 is Moonshot's Kimi model (open weights). It is ranked 6 of 13 models with enough coverage on the Visual Intelligence Index.
- Moonshot
- Open weights
- Kimi
- Video input
- Structured output
- Tool calling
- Streaming
- Self-hostable
Comparison Summary
Kimi K3 ranks 6 of 13 on the perception index. Immediately ahead is Gemini 3.8 Flash. Immediately behind is Claude Fable 5.1. List price is $0.60 per million input tokens and $2.50 per million output. 7 of 13 listed models cost less to prompt. Latency is 8.2s with a 256K context window and — tokens/s. Prices and most scores in this snapshot are placeholders, not sourced figures. Strongest capability in this set is Spatial Reasoning (85%). Weakest is Action / Event Understanding (68%).
- Visual Intelligence
- 81.5%
- Availability
- Open weights
- Input / 1M tokens
- $0.60
- Output / 1M tokens
- $2.50
- Latency
- 8.2s
- Context Window
- 256K
- Speed (tokens/s)
- —
- Parameters
- 3T
- Time to first token
- —
All benchmark scores
| Benchmark | Metric | Score | 95% CI | Setting | Source |
|---|---|---|---|---|---|
| Video-MME v2 | Accuracy, no subtitles | — | — | — | — |
| LVBench | Test accuracy | — | — | — | — |
| Perception Test | Overall accuracy | 82.4% | 78.9–85.9 | 1 fps | Reported |
| NExT-QA | Accuracy (hard split) | 86.3% | 85.3–87.2 | 1 fps | Reported |
| Q-Bench Video | Overall accuracy | — | — | — | — |
| EgoSchema | Accuracy (fullset) | 78.6% | 75.0–82.2 | 1 fps | Reported |
| UCF101-AD | Accuracy | 39.4% | 37.9–40.8 | 1 fps | Reported |
Benchmark Scores
Compare reported model scores across each available benchmark or capability index.
13 of 13 models
Display
Capabilities Index Scores
Compare reported model scores across each available benchmark or capability index.
13 of 13 models
Display
Usability
4.0/ 5
What it's like to build against this model, scored out of five from the developer-facing capabilities in the dataset.
- Reachable1.0
There is an endpoint you can call without hosting anything.
- First-party API — supported
- Third-party API — supported
- Portable1.0
You can run it yourself, and are not tied to one vendor.
- Open weights — supported
- Self-hostable — supported
- Multimodal input0.5
It takes the footage directly, rather than frames you extracted.
- Video ingestion — supported
- Audio — not supported
- Programmable1.0
Output you can parse, and tools it can call on its own.
- Structured output — supported
- Tool calling — supported
- Operable0.5
Usable interactively and in bulk, not only one call at a time.
- Streaming — supported
- Batch API — not supported
Other Moonshot models
Other models from the Moonshot family.
| Date | Model | Visual Intelligence | Latency |
|---|---|---|---|
| Apr 15, 2026 | Kimi K3 | 81.5% | 8.2s |