jargon

Applied AI·The API layer

you compute cost per request on your real traffic shape and find the output tokens dominate the number the pricing page quoted.

Pricing model

Draft summary, pending review

Per-token pricing with separate rates for input, output and cached input; output usually costs several times input. Know your provider's numbers and compute cost per request for your real traffic shape, not the marketing example.