AI by Hand ✍️

AI by Hand ✍️

Asymmetric Pricing

Two rates in one call

Prof. Tom Yeh's avatar
Prof. Tom Yeh
May 28, 2026
∙ Paid

Library › Token Problems

  1. Tokenization

  2. Subword Tokens

  3. Non-Word Tokens

  4. Token Ratio

  5. Headroom

  6. Token Pricing

  7. Prompt and Completion

  8. The Billing Line

  9. Asymmetric Pricing

  10. Words to Cents

  11. Token Mix

  12. Usage Forecast

  13. Standing Instructions

  14. Fixed Overhead

  15. Budget Ceiling

  16. Token Allowance

  17. Two Speeds

  18. Output Bound

  19. Rate Cards

  20. Usage Audit

Why does the same token count sometimes cost more? Because input and output are billed separately, at different rates. Input is read in a single pass; output is generated one token at a time, which is the expensive part. An agent that writes long replies will always cost more than one that reads the same amount and answers briefly.

Paid members: the worksheet and its printable PDF are below ↓

This post is for paid subscribers

Already a paid subscriber? Sign in
© 2026 Tom Yeh · Privacy ∙ Terms ∙ Collection notice
Start your SubstackGet the app
Substack is the home for great culture