Kimi K3
Same categoryMoonshot's 2.8T open-weight MoE with always-on thinking and native video understanding
Operating as a massive 2.8T-parameter mixture-of-experts architecture activating 104B parameters per token, Kimi K3 provides a 1,048,576-token context window alongside native video and image ingestion via its 401M-parameter MoonViT-V2 encoder. The model card documents notable agentic evaluations, including an 88.3 score on Terminal-Bench 2.1 and 91.2 on BrowseComp.
Best for: Teams handling multimodal agent workflows involving long video analysis and terminal-based tasks who can support hosting a large MoE or using Moonshot's API.
Consider: Thinking is always on across low, high, or max effort levels, meaning every single request generates reasoning tokens billed at the $15.00 per million output token rate with no option to disable deliberation entirely.
From $3/1M input tokens · Product API available