HomeModelsMoonshot Kimi K3.7 Turbo (512k)

Moonshot Kimi K3.7 Turbo (512k)

Moonshot AI 512K context Released: 2026-07

The low-latency hosted tier of Kimi K3.7. Same weights, dedicated high-throughput serving, roughly double the price and a large jump in tokens per second.

Visit Model Website

Pricing Information

Input Pricing

Standard:$1.2000
Per 1,000 tokens
Cached:$0.1200
Per 1,000 tokens (cached requests)

Output Pricing

Standard:$5.0000
Per 1,000 tokens

Example Costs

Short Conversation
1K input + 500 output tokens
$3.7000
Book Analysis
50K input + 2K output tokens
$70.00

Token Calculator

Tokens:0
Words:0
Characters:0
Input Cost:$0.00
Estimated based on current token count as input
Use our main calculator for more detailed estimates including input/output combinations.

Key Features

  • Turbo serving tier
  • 512K token context window
  • Very high tokens per second
  • Same weights as K3.7
  • Priority capacity

Common Use Cases

  • Latency-sensitive agents
  • Interactive coding
  • Real-time assistants

Ratings & Feedback

0.0 / 5 · 0 votes

Comments

Frequently Asked Questions

What is a token in the context of Large Language Models (LLMs)?

Why is understanding token count important for using LLMs?

What factors affect LLM pricing?

What is a context window in LLMs?

How can I optimize my prompts to use fewer tokens?