Kimi K3 Benchmarks: Every Score, Every Comparison, Every Surprise (July 2026)
Kimi K3 benchmark results are live. 2.8T MoE, #1 on Frontend Code Arena, 76.8% SWE-bench Verified — full comparison tables vs GPT-5.6, Fable 5, Opus 4.8, and GLM-5.2.
Kimi K3 benchmark results are live. 2.8T MoE, #1 on Frontend Code Arena, 76.8% SWE-bench Verified — full comparison tables vs GPT-5.6, Fable 5, Opus 4.8, and GLM-5.2.
Kimi K3 open weights arrive on Hugging Face July 27, 2026. Current status, expected file sizes (594 GB), hardware tiers, vLLM staging, and what K2.6 tells us about the K3 model card.
Kimi K3 open weights arrive July 27, 2026 under a Modified MIT license. Full breakdown: what Moonshot has already open-sourced, K3 self-hosting hardware reality, and whether you should wait for weights or use the API now.
Kimi K3 is live on OpenRouter as moonshotai/kimi-k3 at $3/$15 per million tokens. Complete setup guide: model ID, API key, pricing comparison vs direct API and AWS Marketplace, and code examples.

Hunyuan 3 (Hy3) is Tencent's 295B MoE model with 21B active parameters, Apache 2.0 license, and single-GPU GGUF support. Architecture, benchmarks, pricing, how to run it, and honest limitations.

Inkling is a 975B open-weight MoE model (41B active) from Thinking Machines Lab. Specs, benchmarks, pricing, and how to use it.