Model Detail
Kimi-K3
—Kimi-K3 is a code generation model with 1390.0B parameters released by moonshotai. The model is registered under the image-text-to-text pipeline tag on Hugging Face, and supports text+image->text inputs, distributed under a other license.
Kimi-K3 reports a Chatbot Arena ELO of 1,485 across 3,569 votes. Other benchmark slots are still empty in our dataset, so this single figure is best read as a partial picture rather than a full evaluation.
Kimi-K3 is priced at $3/M input tokens and $15/M output tokens. Operationally the model offers a 1049K-token context window, which matters when sizing it for prompt-heavy or latency-sensitive workloads. Pricing in this range is the working middle of the API market — neither the cheapest nor the most expensive option per token, so cost-fit is usually a function of how much output you generate.
Kimi-K3 ships with 1390.0B parameters. Total weight footprint is approximately 2779.9 GB, which is the relevant figure when planning local-inference VRAM. Distribution is governed by the other license — review the exact terms before commercial deployment.
Kimi-K3 is best fit for code completion, repository-scale Q&A, and pair-programming integrations, and long-context tasks such as full-codebase analysis or book-length summarization (1049K tokens). It is a less obvious choice for one-shot generation of security-critical code without review. Treat this as a starting matrix rather than a benchmark verdict — the right deployment usually depends on the specific evaluation suite that mirrors your workload.
Kimi K3: Open Frontier Intelligence
arXiv:2607.24653v1 Announce Type: new Abstract: We introduce Kimi K3, a 2.8T parameter Mixture-of-Experts model with 104 billion activated parameters, native vision capabilities, and a 1-million-token context window. Kimi K3 is built on Kimi Delta Attention and Attention Residuals, which improve inf
‘AI communism’, rogue models, and the why Kimi K3 spooked Wall Street
Chinese AI lab Moonshot’s open model Kimi went viral this week for reasons that had less to do with the model itself and more to do with how the U.S. AI industry reacted to it. Meanwhile, an unreleased OpenAI model wandered outside its test environment and ended up connected to a real security breac
Experts say exploiting Anthropic’s Fable isn’t how Kimi K3 got so good
"I don't think you get a model this strong and this quickly on the heels of Fable doing strictly distillation," one expert told TechCrunch.
Kimi: Threat or menace?
Chinese company Moonshot AI released a new version of its Kimi model this week, prompting concern about "full AI communism."
Moonshot’s upcoming Kimi 3 is expected to close the gap with Anthropic’s Opus 4.8
The FT reports Kimi K3 will be the largest open AI model from China, with a parameter count between 2 trillion and 3 trillion.
An Independent Safety Evaluation of Kimi K2.5
arXiv:2604.03121v1 Announce Type: cross Abstract: Kimi K2.5 is an open-weight LLM that rivals closed models across coding, multimodal, and agentic benchmarks, but was released without an accompanying safety evaluation. In this work, we conduct a preliminary safety assessment of Kimi K2.5 focusing on