Kimi K3, and what we can still learn from the pelican benchmark
3 experts across 2 network communities independently surfaced this.
“My notes on Kimi K3, plus some thoughts on what we can still learn from the pelican benchmark even while it becomes further detached from how good the models are at the things that matter (like agentic tool calling across longer conversations) simonwillison…”
“simonwillison.net/2026/Jul/16/... >The new model is notable for the pricing: $3/million input tokens and $15/million output tokens, putting it at the same level as Anthropic’s Claude Sonnet series and making it the most expensive model released by a Chinese…”