<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Prompt Caching on</title><link>https://harryzhang.cn/tags/prompt-caching/</link><description>Recent content in Prompt Caching on</description><generator>Hugo</generator><language>zh-CN</language><lastBuildDate>Mon, 17 Aug 2026 10:00:00 +0800</lastBuildDate><atom:link href="https://harryzhang.cn/tags/prompt-caching/index.xml" rel="self" type="application/rss+xml"/><item><title>把 Token 花在刀刃上：Claude Code 的 Context 管理与效率实践</title><link>https://harryzhang.cn/2026-08-17/claude-code-context-token-efficiency/</link><pubDate>Mon, 17 Aug 2026 10:00:00 +0800</pubDate><guid>https://harryzhang.cn/2026-08-17/claude-code-context-token-efficiency/</guid><description>&lt;h2 id="引言同一个任务为什么账单差好几倍"&gt;引言：同一个任务，为什么账单差好几倍&lt;/h2&gt;
&lt;p&gt;两个工程师用 Claude Code 修同一个 bug：一个新开会话、@-mention 相关文件、改完就 &lt;code&gt;/clear&lt;/code&gt;；另一个在一个开了一整天的会话里连续处理十几个不相关的任务，中途还切换过两次模型。最终改动可能一模一样，但后者消耗的 token 往往是前者的数倍。&lt;/p&gt;</description></item></channel></rss>