Model rates and metered billing
Rates are published per million tokens for every model on the platform, input and output listed separately.
Cached input reads are charged at a fraction of the standard input rate, which matters most for long-running agent sessions that resend context.
Billing draws down a prepaid balance: nothing renews, nothing resets, and there is no seat count to manage.