EidosStack Journal

DeepSeek API pricing per million tokens

Understand input and output token pricing, request cost, and the checks that matter before choosing a DeepSeek API.

Input and output are priced separately

A language-model call sends input tokens and receives output tokens. Providers normally publish a different price per one million tokens for each direction, so both numbers matter.

A long context with a short answer can be input-heavy. A short prompt that generates code or a long article can be output-heavy. Estimate both instead of multiplying only the visible answer length.

Use the live selling price

EidosStack publishes the current IDR selling price from an active price card. The card is locked before a request runs, so a price update cannot change a request already in flight.

Availability remains best-effort because upstream supply can change. Check the live catalog and use per-key budgets to bound spend.

Explore the related solution