MoonshotMoonshot·💬 Text Generation·New

Kimi K2.6🔒Private

ReasoningVisionCodeFunction CallingWeb SearchE2EEprivate
🧠 Try in Intelligence →Try on Venice.ai ↗
Quick reference
Kimi K2.6 — TLDR
  • 🏢 Moonshot AI's open-weight Kimi K2.6, served inside a Trusted Execution Environment.
  • 🔒 Hardware attestation lets you verify enclave identity and configuration independently.
  • 🧠 Trillion-parameter Mixture-of-Experts, roughly 32B parameters active per token.
  • 📏 262,144-token context, the setting used in Moonshot's own evaluations.
  • 👁️ Native multimodal input: text plus images, via a MoonViT vision encoder.
  • 🔧 Built for long-horizon coding, tool calling and swarm-style agent orchestration.
  • 🎯 Moonshot reports 36.4% on the HLE text-only subset without tools.
💰 Pricing
$0.870 / $4.12
per 1M · input / output
📏 Context
262K tokens
📅 On Venice since
Sep 6, 2026
1 day ago
Provider

Moonshot is an AI research lab known for developing the Kimi family of large language models. The organization has gained recognition for building capable reasoning-oriented models, with the Kimi line representing its flagship series of text generation…

Read full profile →
7 models on Venice
7 text
Since Jan 27, 2026

About this model

Kimi K2.6 is Moonshot AI's open-weight, natively multimodal agentic model, offered here in a confidential-computing configuration: the model runs inside a Trusted Execution Environment, and hardware attestation evidence is published so users can verify which weights and configuration the enclave is actually running. That makes it aimed at workloads where prompt and output confidentiality matter as much as capability.

Architecturally it is a Mixture-of-Experts design with about one trillion total parameters and roughly 32 billion active per token, paired with a MoonViT vision encoder that accepts images alongside text. Moonshot positions it for long-horizon coding, proactive autonomous execution and swarm-based task orchestration, with tool calling and structured outputs supported for agent frameworks. Evaluations in the model card were run at a 262,144-token context length.

It follows Kimi K2.5 in Moonshot's K2 line, with the vendor's published evaluations emphasising agentic and long-horizon behaviour. On Moonshot's own reported figures, K2.6 reaches 36.4% on the text-only subset of Humanity's Last Exam without tools and 55.5% with tools, with thinking mode enabled.

Related models in the wider Kimi lineup include Kimi K2.7 Code, which narrows the focus to coding, along with Kimi K3 and its latency-oriented variant Kimi K3 Fast. K2.6 is the choice when open weights, multimodal input and verifiable private execution are the priorities.

This About section is AI-generated from public sources (Claude Opus 5), with no human editing. It may contain inaccuracies — verify critical details against the sources listed above.

Research & Papers

Primary reference paper for this model family, sourced from the HuggingFace model card.

Data sources: Venice API · HuggingFace · Wikipedia · arXiv — enrichment updated 1d ago