Billing
FIM Gate bills by token usage with per-model rates; some multimedia APIs are billed per call. Live prices for all models are listed on the pricing page.
FIM Gate prices in US dollars and charges based on actual usage. Live prices for all models are listed on the pricing page — this page only explains how costs are calculated and does not maintain any price snapshots.
Token-based billing
Most text models are billed by token usage: input and output are priced separately, and rates vary by model.
Cost = input tokens × input rate + output tokens × output rateThe pricing page lists each model's input and output rates and pricing units. Tokens generated by features such as caching and thinking also count toward usage; the usage field in each response reports the actual consumption of that request.
Example: suppose a model is priced at $0.40 per million input tokens and $1.60 per million output tokens, and a request consumes 2,000 input tokens and 1,000 output tokens:
Cost = 2000 / 1,000,000 × 0.40 + 1000 / 1,000,000 × 1.60
= 0.0008 + 0.0016
= $0.0024Per-call billing
Some multimedia APIs are billed per call: each successful call incurs one charge, regardless of the length of the generated content. The applicable APIs and per-call prices are also listed on the pricing page.
Checking prices and billing details
- Live model list and prices: https://gate.fim.ai/pricing
- Usage and charge details: sign in to the Console and view per-request consumption in the usage records.
Whenever you're unsure what a model costs, just open the pricing page for the latest rates.