Cost per answer
Every reply carries its own cost line, right under the answer: input tokens, cached tokens, output tokens, and the price. The figures update while the answer streams.
The line under a reply
Answer
Input tokens
Cached tokens
Output tokens
Price
Cached tokens
Cached input is cheaper than fresh input. Long threads benefit from that automatically.
The full ledger
The line under a reply covers that one answer. Usage holds the full ledger, per chat, per model and per day.
The same entries, grouped three ways
Per chat
Per model
Per day