← Back to context Comment by alberth 12 hours ago Do these system prompts count against your token usage? 5 comments alberth Reply simonw 12 hours ago No, because these ones affect the consumer chat products and not the API or Claude Code.(Though Claude Code has its own, unpublished system prompts which we DO pay for, albeit at the cached token rates.) virtujoel 12 hours ago Technically these do count against your token usage if you happen to use claude.ai web chat alongside Claude Code - both use the same allowance. Makes me appreciate OpenAI/ChatGPT giving you unlimited chat that doesn't drain your Codex allowance. simonw 11 hours ago Yeah, Claude chat does show little in-app messages occasionally warning that Opus or Fable will burn through your rates faster. Dfol 12 hours ago So YES if you're using Cowork or Chat TZubiri 12 hours ago No, because they are cached, the inference cost is paid once per model, does not scale linearly per user or use.
simonw 12 hours ago No, because these ones affect the consumer chat products and not the API or Claude Code.(Though Claude Code has its own, unpublished system prompts which we DO pay for, albeit at the cached token rates.) virtujoel 12 hours ago Technically these do count against your token usage if you happen to use claude.ai web chat alongside Claude Code - both use the same allowance. Makes me appreciate OpenAI/ChatGPT giving you unlimited chat that doesn't drain your Codex allowance. simonw 11 hours ago Yeah, Claude chat does show little in-app messages occasionally warning that Opus or Fable will burn through your rates faster. Dfol 12 hours ago So YES if you're using Cowork or Chat TZubiri 12 hours ago No, because they are cached, the inference cost is paid once per model, does not scale linearly per user or use.
virtujoel 12 hours ago Technically these do count against your token usage if you happen to use claude.ai web chat alongside Claude Code - both use the same allowance. Makes me appreciate OpenAI/ChatGPT giving you unlimited chat that doesn't drain your Codex allowance. simonw 11 hours ago Yeah, Claude chat does show little in-app messages occasionally warning that Opus or Fable will burn through your rates faster.
simonw 11 hours ago Yeah, Claude chat does show little in-app messages occasionally warning that Opus or Fable will burn through your rates faster.
Dfol 12 hours ago So YES if you're using Cowork or Chat TZubiri 12 hours ago No, because they are cached, the inference cost is paid once per model, does not scale linearly per user or use.
TZubiri 12 hours ago No, because they are cached, the inference cost is paid once per model, does not scale linearly per user or use.
No, because these ones affect the consumer chat products and not the API or Claude Code.
(Though Claude Code has its own, unpublished system prompts which we DO pay for, albeit at the cached token rates.)
Technically these do count against your token usage if you happen to use claude.ai web chat alongside Claude Code - both use the same allowance. Makes me appreciate OpenAI/ChatGPT giving you unlimited chat that doesn't drain your Codex allowance.
Yeah, Claude chat does show little in-app messages occasionally warning that Opus or Fable will burn through your rates faster.
So YES if you're using Cowork or Chat
No, because they are cached, the inference cost is paid once per model, does not scale linearly per user or use.