Pricing

Per 1M token pricing for models available through the Kimchi serverless API. For current plans, billing details, and the latest model availability, see kimchi.dev/pricing.

ModelProviderCapabilitiesInput / 1M tokensCached input / 1M tokensOutput / 1M tokens
kimi-k2.7Deployed by Cast AItext, image$0.95$0.19$4.00
nemotron-3-ultra-fp4Deployed by Cast AItext$0.60$3.60
deepseek-v4-flashDeployed by Cast AItext$0.14$0.07$0.28
deepseek-v4-flash-0731Deployed by Cast AItext$0.14$0.07$0.28
glm-5.2-fp8Deployed by Cast AItext$1.40$0.26$4.40
minimax-m3Deployed by Cast AItext, image$0.30$0.06$1.20

External providers

Requests routed through Kimchi to external providers incur a flat per-token surcharge regardless of model. This surcharge is added on top of the provider's own token pricing. See Supported Providers for the list of external providers you can connect.

Input / 1M tokensCached input / 1M tokensOutput / 1M tokens
$0.20Free$0.30


Did this page help you?