Do these system prompts count against your token usage?
No, because these ones affect the consumer chat products and not the API or Claude Code.
(Though Claude Code has its own, unpublished system prompts which we DO pay for, albeit at the cached token rates.)
Technically these do count against your token usage if you happen to use claude.ai web chat alongside Claude Code - both use the same allowance. Makes me appreciate OpenAI/ChatGPT giving you unlimited chat that doesn't drain your Codex allowance.
Yeah, Claude chat does show little in-app messages occasionally warning that Opus or Fable will burn through your rates faster.
So YES if you're using Cowork or Chat
No, because they are cached, the inference cost is paid once per model, does not scale linearly per user or use.
No, because these ones affect the consumer chat products and not the API or Claude Code.
(Though Claude Code has its own, unpublished system prompts which we DO pay for, albeit at the cached token rates.)
Technically these do count against your token usage if you happen to use claude.ai web chat alongside Claude Code - both use the same allowance. Makes me appreciate OpenAI/ChatGPT giving you unlimited chat that doesn't drain your Codex allowance.
Yeah, Claude chat does show little in-app messages occasionally warning that Opus or Fable will burn through your rates faster.
So YES if you're using Cowork or Chat
No, because they are cached, the inference cost is paid once per model, does not scale linearly per user or use.