Moonshot · Active

kimi-k2.7-code-highspeed

kimi-k2.7-code-highspeed
Context window
262,144 tokens
Maximum output
Not documented
Input price
$1.9 per 1M tokens
Output price
$8 per 1M tokens
Modalities
text

Last verified · Source ↗

Official description

The official Kimi inference catalog lists this model with separate cached input, uncached input, and output billing.

Read the official model reference ↗.

Where this model fits

kimi-k2.7-code-highspeed is part of the Moonshot API catalog. Its documented modalities are text. Choose it against the request format, output requirements and operational limits of your application.

For a repeatable evaluation, use representative requests, inspect structured output or tool calls, and log latency and usage in your own account. Catalog metadata establishes compatibility; it does not measure answer quality on your workload. Start from the provider guide for authentication and endpoint setup.

Current pricing

Verified model prices
ModelPrice typeUSDUnitTierRegionSource
kimi-k2.7-code-highspeedCached Input$0.38per 1M tokensStandardSee sourceOfficial source ↗
kimi-k2.7-code-highspeedInput$1.9per 1M tokensStandardSee sourceOfficial source ↗
kimi-k2.7-code-highspeedOutput$8per 1M tokensStandardSee sourceOfficial source ↗

Last verified · Source ↗

Estimate request, daily and monthly usage in the API cost calculator.

Rate limits and availability

Verified rate limits
Model or scopeTierMetricLimitNotesSource
All models (provider scope)Tier0concurrency1Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier0RPM3Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier0TPD1,500,000Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier0TPM500,000Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier1concurrency15Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier1RPM100Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier1TPM2,000,000Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier2concurrency40Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier2RPM100Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier2TPM3,000,000Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier3concurrency50Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier3RPM200Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier3TPM3,000,000Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier4concurrency60Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier4RPM200Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier4TPM4,000,000Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier5concurrency100Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier5RPM300Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗
All models (provider scope)Tier5TPM5,000,000Published account tier. Unlimited quotas are not converted into a numeric cap.Official source ↗

Last verified · Source ↗

Account-specific capacity may differ. Check the provider rate-limit guide and the free-tier conditions.

Three alternatives to compare

These alternatives share a documented modality and have the closest combination of input price and context size in the current catalog. This is a specification comparison, not a quality benchmark.

Computed model alternatives
ModelProviderInput / 1M tokensContext
Mistral: Mistral Medium 3.5OpenRouter$1.5262,144
Cohere: Command AOpenRouter$2.5256,000
OpenAI: o3OpenRouter$2200,000

Use the exact identifier in Python

This identifier fragment belongs in the authenticated request described in the provider Python tutorial. The tutorial covers the correct SDK, endpoint, environment variable and response shape.

model = "kimi-k2.7-code-highspeed"
# Pass model to the provider-specific request in the tutorial.

Recent changes

  1. Moonshot · kimi-k2.7-code-highspeed — Amount UsdNot previously recorded → Price category: output · USD: 8 · Unit: per 1M tokensSource ↗
  2. Moonshot · kimi-k2.7-code-highspeed — Amount UsdNot previously recorded → Price category: input · USD: 1.9 · Unit: per 1M tokensSource ↗
  3. Moonshot · kimi-k2.7-code-highspeed — Amount UsdNot previously recorded → Price category: cached_input · USD: 0.38 · Unit: per 1M tokensSource ↗
  4. Moonshot · kimi-k2.7-code-highspeed — MetadataRecord updated; consult the linked source for details. → Record updated; consult the linked source for details.Source ↗
  5. Moonshot · kimi-k2.7-code-highspeed — First SeenNot previously recorded → Model: kimi-k2.7-code-highspeedSource ↗

Subscribe to the changelog RSS feed

Frequently asked questions

What is the exact API model identifier?
Use kimi-k2.7-code-highspeed exactly as shown in the provider model reference. Display names are not interchangeable with API identifiers.
How much does this model cost?
The current verified prices and billing units are listed in the pricing table above. Cache, batch, tier and region conditions remain separate rows when the provider documents them.
How much input can the model accept?
The model card shows the documented context window and maximum output when available. Your request must leave room for both input and generated output, including any provider-specific reasoning allocation.
Can I use a free tier?
Check the linked provider free-tier page for eligibility and account conditions. An account credit or temporary trial does not establish a permanent free rate.
Where do these model facts come from?
Official provider references are linked beside the data and in Sources. Changes are recorded with observation times; unresolved or undocumented fields are left unavailable.

Sources

Last verified · Source ↗