Kimi K3
kimi-k3moonshotai/kimi-k3- 🧠 Ultra-large-scale open-weight reasoning model from Moonshot AI
- 📏 Massive 1M-token context window
- 👁️ Multimodal — reasons over images, logs, and tests
- 🔧 Strong tool use, function calling, and web search
- 🎯 Built for complex coding and long-horizon agentic work
- 🌍 Open weights for self-hosting and inspection
Moonshot is an AI research lab known for developing the Kimi family of large language models. The organization has gained recognition for building capable reasoning-oriented models, with the Kimi line representing its flagship series of text generation systems.
Explore 6 more models by Moonshot →Kimi K3 is Moonshot AI's ultra-large-scale, open-weight multimodal reasoning model, built for complex coding, knowledge work, and long-horizon agentic workflows. Released in July 2026, it pairs a one-million-token context window with vision, tool use, function calling, and web search, letting it iterate against images, logs, tests, and runtime feedback rather than reasoning in isolation. Its design leans heavily toward navigating large repositories, debugging, and multi-step problem solving where a model must plan, act, and revise over extended sessions.
Within Moonshot's Kimi K lineup, K3 sits at the top of the general-purpose reasoning tier, alongside Kimi K2.6 and Kimi K2.5, while Kimi K2.7 Code serves as the dedicated coding-specialized branch. As an open-weight release, it can be inspected and self-hosted, an appealing trait for teams wanting transparency alongside frontier-scale capability.
K3 is best suited to demanding agentic and engineering tasks — large-codebase reasoning, automated debugging, tool-driven research, and workflows that combine text and visual inputs across very long contexts.
This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.
| Seller | Reputation↓ | Routing | Input $/M | Cached $/M | Output $/M | Categories | API |
|---|---|---|---|---|---|---|---|
| Phala 0x88c8…15d6 | 81.21 | #7 | $3.00 | $0.30 | $15.00 | chat,confidential,reasoning,multimodal | openai-chat-completions |
| ▲ Apex Ant 0x73b4…e736 | 71.90 | #1 | $1.67 | $0.17 | $8.35 | chat,coding,vision,multimodal,reasoning,long-context,agents | openai-chat-completions |
| Venice.ai Proxy 0x1f22…18c9 | 70.55 | #4 | $1.875 | $0.1875 | $9.375 | chat,reasoning,coding,vision,multimodal,web-search | openai-chat-completions |
| D5V1N2 0xd5e7…7be0 | 67.45 | #2 | $1.80 | $0.30 | $9.00 | chat,coding,reasoning,agent,vision,large-context,router,fallback | openai-chat-completions |
| Prime Seed 0x71e2…fc07 | 65.07 | #5 | $2.10 | $0.2436 | $11.76 | chat,coding,reasoning,premium | openai-chat-completions |
| Super Seeder 0xd19f…41f3 | 63.77 | #3 | $1.8563 | $0.1856 | $9.2813 | chat,coding,reasoning,vision,multimodal,tools,long-context | openai-chat-completions |
| Edith AI 0xb269…b1a6 | 61.93 | #6 | $1.481 | $0.2962 | $14.80 | chat,coding | openai-chat-completions |
| Open Forge 0x1d90…b0aa | 58.60 | gated | $2.85 | $0.285 | $14.25 | chat,multimodal | openai-chat-completions |
| ZLKPro-Api 0x0b0b…f446 | 58.35 | gated | $0.098 | $0.091 | $0.40 | agent,chat,text,reasoning,research,smart,long-context,multimodal,visual | openai-chat-completions |
| Chutes 0xded6…657c | 55.36 | gated | $3.30 | $0.33 | $16.50 | chat,reasoning,coding,vision,tee | openai-chat-completions |
| bartly.eth64.de 0x666e…4666 | 53.43 | gated | $0.0306 | $0.0031 | $0.1529 | chat,coding,reasoning,long-context | openai-chat-completions |
| Open Ant 0xe4f6…5bc4 | 50.00 | gated | $3.75 | $0.375 | $18.75 | chat,coding,reasoning,vision,multimodal,tools,long-context | openai-chat-completions |
| Fire Ant 🔥🐜 0xbe05…bc5d | 50.00 | gated | $2.9862 | $0.2986 | $14.9312 | chat,coding,reasoning,tools | — |
| Open Bird 0xc0f1…8183 | 50.00 | gated | $1.05 | $0.105 | $5.25 | chat,long-context | openai-chat-completions |
| NovaRoute AI 0xc50d…ed7b | 50.00 | gated | $0.099 | $0.099 | $0.4455 | chat,coding,code,reasoning,agent,tasks,frontier,kimi,value,surplus,openai-compatible,low-cost,verified,github,response-auth,base-usdc,monitored | openai-chat-completions |
| Best 0x472f…69fd | 49.76 | gated | $0.0305 | $0.0031 | $0.1528 | chat,coding | openai-chat-completions |
| Hana Gateway ✅ 0x4ae1…117b | 45.41 | gated | $0.10 | $0.10 | $0.50 | chat,coding,reasoning,tools | openai-chat-completions |
| incharge-austin 0x5938…ac97 | 30.79 | gated | $2.85 | $0.285 | $14.25 | chat,coding | openai-chat-completions |
| Xerxes Inference 0xdc73…65ce | 18.10 | gated | $0.16 | $0.10 | $0.75 | chat,code,reasoning | openai-responses |
| Apex TEE Test 0xe672…7955 | 15.22 | gated | $1.6875 | $1.6875 | $8.4375 | chat,coding,vision,multimodal,reasoning,long-context,agents | openai-chat-completions |
| Cooper 0xacb4…00ae | 12.07 | gated | $0.35 | $0.0525 | $1.25 | chat,coding,reasoning | openai-chat-completions |
"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = buyer trust score (0-100, the AntSeed SDK's own formula). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.