Grok 4.5 beat Kimi K3 on the same prompt at 13x lower cost.
Very interesting experiment by AI/ML API (@aimlapi)
Same prompt, same one-shot task
Cost per figure: Grok 4.5 - $0.15 GPT 5.6 Sol - $0.60 Qwen 3.8 Max - $0.67 Kimi K3 - $1.98
Kimi spent ~19 minutes thinking and billed 13x more. Grok just shipped.
For production agents, cost per completed task is starting to matter a lot more than cost per token.