Applied AI·The API layer
you compute cost per request on your real traffic shape and find the output tokens dominate the number the pricing page quoted.
Pricing model
Draft summary, pending review
Per-token pricing with separate rates for input, output and cached input; output usually costs several times input. Know your provider's numbers and compute cost per request for your real traffic shape, not the marketing example.