·
DataBubble
  • Home
  • Models
  • News
  • Compare
  • Boards
  • Pricing
  • About
  • Newsletter
  • Methodology
  • Contact
Latest
Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning6h◆Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks6h◆Distribution-Consistent Inference for Dynamic Sparse Mixture-of-Experts6h◆In RAG We Trust? Measuring Robustness of Retrieval-Augmented Generation Under Document Poisoning6h◆FrontierChallenge: Evaluating Scientific Workflow Completion6h◆IBIB: A Protocol for Measuring Enterprise AI Systems by Serving Route, Not Model Identifier6h◆AgenticGen: Reward-Guided Agentic Video Generation for Advertising6h◆Strangers to Themselves: What Language Models Say About Themselves Is Generic6h◆Omni Interaction Agent Technical Report6h◆CoGReV: A Confidence-Gated Post-Hoc Non-Monotonic Belief Revision Framework for Phishing Website Classification6h◆Bit-Flip Attacks on Vision-Language-Action Models: Action-Decoding Architecture Shapes the Vulnerability6h◆Can Foundation Models Moderate Online Content? Evaluating Instruction- vs. Example-Driven Policy Operationalization6h◆RiLM: Parameter-Efficient Language Modeling via Geodesic Decoding6h◆Tracing Computation Density in LLMs6h◆Cultural Binding Heads in Language Models6h◆Direct Diversity Optimization for Diverse Successful Trajectories in Preference Post-Training6h◆Left-Branching Transformers Excel at Right-Branching Languages: Data Shapes Word Order Preferences in Language Models6h◆Palmyra x6 Technical Report: An Agentic, Tool-Use Model Post-Trained via Anchored Supervised Fine-Tuning6h◆'Ghaib in Translation' aka Unseen Harm: Measuring Cross-Script Safety Inconsistency with 'Missed-in-Urdu' Scores in LLM Hate Speech Detection6h◆DexterSQL: Deep Schema Exploration and Rule-based Correction for Text-to-SQL Generation6h◆Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning6h◆Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks6h◆Distribution-Consistent Inference for Dynamic Sparse Mixture-of-Experts6h◆In RAG We Trust? Measuring Robustness of Retrieval-Augmented Generation Under Document Poisoning6h◆FrontierChallenge: Evaluating Scientific Workflow Completion6h◆IBIB: A Protocol for Measuring Enterprise AI Systems by Serving Route, Not Model Identifier6h◆AgenticGen: Reward-Guided Agentic Video Generation for Advertising6h◆Strangers to Themselves: What Language Models Say About Themselves Is Generic6h◆Omni Interaction Agent Technical Report6h◆CoGReV: A Confidence-Gated Post-Hoc Non-Monotonic Belief Revision Framework for Phishing Website Classification6h◆Bit-Flip Attacks on Vision-Language-Action Models: Action-Decoding Architecture Shapes the Vulnerability6h◆Can Foundation Models Moderate Online Content? Evaluating Instruction- vs. Example-Driven Policy Operationalization6h◆RiLM: Parameter-Efficient Language Modeling via Geodesic Decoding6h◆Tracing Computation Density in LLMs6h◆Cultural Binding Heads in Language Models6h◆Direct Diversity Optimization for Diverse Successful Trajectories in Preference Post-Training6h◆Left-Branching Transformers Excel at Right-Branching Languages: Data Shapes Word Order Preferences in Language Models6h◆Palmyra x6 Technical Report: An Agentic, Tool-Use Model Post-Trained via Anchored Supervised Fine-Tuning6h◆'Ghaib in Translation' aka Unseen Harm: Measuring Cross-Script Safety Inconsistency with 'Missed-in-Urdu' Scores in LLM Hate Speech Detection6h◆DexterSQL: Deep Schema Exploration and Rule-based Correction for Text-to-SQL Generation6h◆
DataBubble·

Model Detail

moonshotai logo

Kimi-K3

▼ 2.5%
Provider: moonshotaiCategory: codePipeline: image-text-to-text
DB Score
38.3
Downloads
2.3M
Likes
11K
Day
-2.5%
Week
+0.0%
Month
+0.0%
Overview

Kimi-K3 is a code generation model with 1390.0B parameters released by moonshotai. The model is registered under the image-text-to-text pipeline tag on Hugging Face, and supports text+image+video->text inputs, distributed under a other license.

Performance

Kimi-K3 reports a Chatbot Arena ELO of 1,485 across 3,569 votes. Other benchmark slots are still empty in our dataset, so this single figure is best read as a partial picture rather than a full evaluation.

How we score this →
Pricing & Throughput

Kimi-K3 is priced at $3/M input tokens and $15/M output tokens. Operationally the model offers a 1024K-token context window, which matters when sizing it for prompt-heavy or latency-sensitive workloads. Pricing in this range is the working middle of the API market — neither the cheapest nor the most expensive option per token, so cost-fit is usually a function of how much output you generate.

Technical

Kimi-K3 ships with 1390.0B parameters. Total weight footprint is approximately 2779.9 GB, which is the relevant figure when planning local-inference VRAM. Distribution is governed by the other license — review the exact terms before commercial deployment.

Trending Signal

Downloads of Kimi-K3 have moved -2.5% over the past 24 hours. That is a slight downtrend, consistent with normal cooling as newer models compete for the same workloads. These numbers are signal, not guarantee — week-over-week download counts on Hugging Face also reflect mirror traffic, CI scrapes, and one-off benchmarking runs.

Read about databubble_score →
Use Cases

Kimi-K3 is best fit for code completion, repository-scale Q&A, and pair-programming integrations, and long-context tasks such as full-codebase analysis or book-length summarization (1024K tokens). It is a less obvious choice for one-shot generation of security-critical code without review. Treat this as a starting matrix rather than a benchmark verdict — the right deployment usually depends on the specific evaluation suite that mirrors your workload.

Download History
Pricing
Input ($/M tokens)
$3
Output ($/M tokens)
$15
Context Window
1024K
Arena & Community
Arena ELO
1,485
Arena Votes
3,569
Model Info
Licenseother
Modalitytext+image+video->text
Recent newsView all news →
Related News
arxivneutral44d ago

Kimi K3: Open Frontier Intelligence

arXiv:2607.24653v1 Announce Type: cross Abstract: We introduce Kimi K3, a 2.8T parameter Mixture-of-Experts model with 104 billion activated parameters, native vision capabilities, and a 1-million-token context window. Kimi K3 is built on Kimi Delta Attention and Attention Residuals, which improve i

techcrunchneutral48d ago

‘AI communism’, rogue models, and the why Kimi K3 spooked Wall Street

Chinese AI lab Moonshot’s open model Kimi went viral this week for reasons that had less to do with the model itself and more to do with how the U.S. AI industry reacted to it. Meanwhile, an unreleased OpenAI model wandered outside its test environment and ended up connected to a real security breac

techcrunchneutral49d ago

Experts say exploiting Anthropic’s Fable isn’t how Kimi K3 got so good

"I don't think you get a model this strong and this quickly on the heels of Fable doing strictly distillation," one expert told TechCrunch.

techcrunchneutral54d ago

Kimi: Threat or menace?

Chinese company Moonshot AI released a new version of its Kimi model this week, prompting concern about "full AI communism."

techcrunchneutral56d ago

Moonshot’s upcoming Kimi 3 is expected to close the gap with Anthropic’s Opus 4.8

The FT reports Kimi K3 will be the largest open AI model from China, with a parameter count between 2 trillion and 3 trillion.

arxiv158d ago

An Independent Safety Evaluation of Kimi K2.5

arXiv:2604.03121v1 Announce Type: cross Abstract: Kimi K2.5 is an open-weight LLM that rivals closed models across coding, multimodal, and agentic benchmarks, but was released without an accompanying safety evaluation. In this work, we conduct a preliminary safety assessment of Kimi K2.5 focusing on

Related Models
moonshotai logo
Kimi-K2.5
moonshotai · 1.5M downloads
moonshotai logo
Kimi-K2.6
moonshotai · 857K downloads
sentence-transformers logo
all-MiniLM-L6-v2
SBERT · 254.3M downloads
nomic-ai logo
nomic-embed-text-v1.5
nomic-ai · 17.1M downloads
HomeModelsNews