Kimi K2 Thinking

Kimi K2 Thinking is the latest, most capable version of open-source thinking model. Starting with Kimi K2, we built it as a thinking agent that reasons step-by-step while dynamically invoking tools. It sets a new state-of-the-art on Humanity's Last Exam (HLE), BrowseComp, and other benchmarks by dramatically scaling multi-step reasoning depth and maintaining stable tool-use across 200–300 sequential calls. At the same time, K2 Thinking is a native INT4 quantization model with 256k context window, achieving lossless reductions in inference latency and GPU memory usage.

Contexte
262K
Sortie maximale
100K
Entrée
0,60 $US
Sortie
2,5 $US
DECISION SUMMARY

Cas d’usage recommandés

Points forts dans ces données

  • 262,144-token context window
  • text input
  • 17 supported API parameters listed

Limites et réserves

  • Provider behavior and pricing can change; verify the linked sources before production use.
CAPABILITIES

Capacité

Faits propres au modèle

Modèle
Raisonnement
Pris en charge
Poids ouverts
Inconnu

Faits de l’endpoint fournisseur

Endpoint fournisseur
Appels d’outils
Pris en charge
Sortie structurée
Pris en charge
Streaming
Inconnu
Cache de prompt
Pris en charge
Traitement par lots
Inconnu
Ajustement fin
Inconnu
PROVIDER PRICING

Kimi K2 Thinking Tarifs du fournisseur

Endpoint fournisseur: moonshotai/kimi-k2-thinking

Entrée
0,60 $US
par million de tokens
Sortie
2,5 $US
par million de tokens
Entrée en cache
0,15 $US
par million de tokens
Sortie image
Inconnu
par million de tokens
SOURCE RECORDS

Sources et vérification

Hugging Face model card

Champs: tags, gated, license, summary, library name, pipeline tag

OpenRouter Models API

Champs: identity, description, modalities, context window, maximum output, pricing, supported parameters

MODEL FAQ

Questions fréquentes

Les réponses utilisent les mêmes données sourcées sur le modèle et les fournisseurs que celles affichées ci-dessus.

Kimi K2 Thinking est un modèle d’IA de MoonshotAI, appartenant à la famille Other. Sa fenêtre de contexte répertoriée est de 262K.

Les spécifications et tarifs peuvent varier selon le fournisseur et évoluer. Vérifiez les sources et dates de validation avant toute utilisation en production.