kimi-k2-turbo-preview

by Moonshot AI

The kimi-k2-turbo-preview model is a high-speed version of kimi-k2, with the same model parameters as kimi-k2, but the output speed has been increased from 10 tokens per second to 40 tokens per second.

API Pricing

Input$1.2 / 1M tokens
Output$4.8 / 1M tokens
Cache read$0.3 / 1M tokens

Specifications

Context window262,144 tokens
Modalitiestext
Featurestools, function_calling, structured_outputs

More from Moonshot AI

Use kimi-k2-turbo-preview via the AIHubMix unified API — one interface for every major LLM.