No free snack is open right now.
When the next one opens, its claim terms will appear here.
Compare live model pricing and availability, then call leading AI models through one OpenAI-compatible API.
Free tokens and paid model packs live here. Each paid pack is reserved for the exact model on its card.
Enjoy a small token snack for the free models that are available today.
When the next one opens, its claim terms will appear here.
Use a paid pack only with its named model. It never silently falls back to PAYG.
A pack appears here only after its model and price are published for sale.
Prices are shown per one million input, output, or cache tokens. Model availability is refreshed every 30 seconds.
claude-fable-5ClaudeRpĀ 70.850$3.9555Official claude-fable-5.1ClaudeRpĀ 71.500$3.9918Official claude-haiku-4.5ClaudeRpĀ 7.150$0.3992Official claude-opus-4.6ClaudeRpĀ 35.750$1.9959Official claude-opus-4.6-thinkingClaudeRpĀ 35.750$1.9959Official claude-opus-4.7ClaudeRpĀ 35.750$1.9959Official claude-opus-4.8ClaudeRpĀ 35.750$1.9959Official claude-opus-5ClaudeRpĀ 35.750$1.9959Official claude-sonnet-4.5ClaudeRpĀ 21.450$1.1975Official claude-sonnet-4.6ClaudeRpĀ 21.450$1.1975Official Every model has separate live rates for input, output, and cache tokens. The model price is confirmed before the request runs.
Enter a realistic token count. This estimate reads the same public price data as the model catalog above.
Add credit through EidosStack checkout. Your balance becomes available after payment is confirmed.
For a paid model with an active package, the package quota is used first. If there is no active package for that model, the request uses your PAYG wallet. An exhausted active package returns an insufficient-package-quota error and does not silently charge PAYG.
Prompt text and completion text are not shown in usage history or placed in cache keys. Eligible requests can use a short-lived response cache that is isolated to the same account.
The live catalog shows the models that are currently available, their token prices, and their cache prices. It refreshes availability every 30 seconds.
See endpoints, authentication, limits, and request examples.
Use input, output, cache, and a realistic request scenario before adding credit.
Learn why input and output tokens are priced separately.
Check live availability, token units, integration, and subscription boundaries.
Choose a billing approach based on your workload.
See how request content, usage records, and service providers are handled.
Use maintained examples that start from the live model catalog and keep credentials in environment variables.