<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>Agent 成本 on 智识脉搏</title>
    <link>https://cuigh.com/zh/tags/agent-%E6%88%90%E6%9C%AC/</link>
    <description>Recent content in Agent 成本 on 智识脉搏</description>
    <generator>Hugo</generator>
    <language>zh-CN</language>
    <lastBuildDate>Mon, 22 Jun 2026 10:51:00 +0800</lastBuildDate>
    <atom:link href="https://cuigh.com/zh/tags/agent-%E6%88%90%E6%9C%AC/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>把 90% 的 token 烧在错的地方：headroom 帮我把 Agent 成本压了 7 倍</title>
      <link>https://cuigh.com/zh/posts/headroom-token-90-percent-burned-2026/</link>
      <pubDate>Mon, 22 Jun 2026 10:51:00 +0800</pubDate>
      <guid>https://cuigh.com/zh/posts/headroom-token-90-percent-burned-2026/</guid>
      <description>GitHub 月榜 #1 的 headroom 主张 token 减 60–95%。我拿一组真实的 Claude Code 调用记录喂进去，结果是 7.3 倍成本下降、响应时间砍掉一半、答案质量肉眼可见没掉。这不是新模型的故事，是上下文工程第一次站到 C 位。</description>
    </item>
  </channel>
</rss>
