20 ms·
Save Claude Code Tokens with Smart Routing
- deleted 3mo ago[deleted]
- nithiink 3mo agoHow do you handle prompt caching? A lot of cost savings for a single model chat come from cache hits on the conversation context, and switching models invalidates that cache — the new model has to reprocess everything at full input price.
- FrancescoMassa 2mo ago[flagged]
- patch_dev 3mo agoWhat does this solve that well used subagents doesn't solve already?
- FrancescoMassa 3mo agoOn our tests subagents & well used workflows are 20-30% more expensive for context & token efficiency