← All matchupsMinecraft — five models build it from a two-line prompt
Kimi K3 (high)vsLuna 5.6 (max reasoning)
1 prompt where both models ran the exact same instructions. Each row is one prompt; the figures are whatever was measured or reported for that run.
Kimi K3 (high) and Luna 5.6 (max reasoning) ran the same prompt on 1 task, side by side. Cost, duration and outcome for each — 0 measured locally, no aggregate score.
- Shared prompts
- 1
- Compared
- models
- Measured runs
- 0
No winner is declared. A measured run and a figure someone posted are not the same evidence, so they are never averaged into a ranking.
GamesReferenced
Kimi K3 (high)
$2.87not measuredCompletedreported
Luna 5.6 (max reasoning)
$0.35not measuredCompletedreported