Skip to content

#Context/memory

0 items today
8/30Sun
  1. Ryan Lopopolo38

    Ryan Lopopolo 如何为 Homelab 做 Harness Engineering

    Ryan Lopopolo 把 Homelab 的 agent 质量保障主要落在文档上:文档站点按关注点组织清单、拓扑、工作负载、runbook 和维护约定,AGENTS.md 只提供索引,agent 按需加载详细操作模型。当在线指令遵循不够时,这些文档会被接入一个很薄的每周自动化来收敛仓库,让正确的上下文持久化并闭环。

    Awaiting translation

8/29Sat
  1. DevAgentStack · Field Notes60

    用 ccusage 追踪 Claude Code 的 token 成本与缓存命中

    ccusage 通过读取本地会话日志,在终端直接输出 Claude Code 的 token 用量和费用明细,无需 API key 或代理配置。它把未缓存输入、缓存写入和折扣后的缓存读取分开统计,支持按会话、日、周、月查看,可用 npx ccusage@latest 直接运行,缓存过定价数据后加 --offline 可避免联网。

    Awaiting translation

8/28Fri
8/27Thu
  1. Addy Osmani · Blog71

    如何审计你的 Agent 配置文件:CLAUDE.md、Skills 与 Hooks 的定期清理

    Addy Osmani 建议每隔几周运行一次 Claude Code 的 /doctor,单独用 /memory 检查记忆,并让每条指令重新证明自己的价值,因为模型、harness 和代码库都在变,旧配置会留下。

    Awaiting translation

    Why it matters: 作者结合自身配置审计经验与近期研究,说明 Agent 配置文件为何会腐化,以及如何按节奏清理。

8/24Mon
8/20Thu
  1. Permission Protocol · AI Agent Incident Tracker65

    加密上下文注入绕过模型过滤并窃取 Grok 聊天数据

    Adversa AI 披露一种加密上下文注入手法:攻击者提供密文、密钥和解密指令,输入过滤只能看到加密内容,模型在初始安全边界之后还原出明文指令,进而访问私有对话上下文或把数据外传,演示了 Grok 聊天数据泄露和 Gemini 的护栏绕过。

    Awaiting translation

8/19Wed
  1. Cursor · Changelog66

    Cursor Updates Cloud Agents and Harness: Event Subscriptions, Independent Sub-Agent Runs, and /goal

    Cursor has updated its cloud agents and Cursor harness so cloud agents can subscribe to event sources, resume when there's new activity in a PR, Slack thread, or scheduled task, and keep going until the work is done—fixing CI failures and handling bot comments.

    Why it matters: Cloud agents are moving from one-shot runs to subscribing to events and following up continuously on PRs and Slack threads, which gives readers a way to judge how the boundaries of automation are shifting.

8/18Tue
  1. Permission Protocol · AI Agent Incident Tracker78

    Context7 MCP custom AI instruction prompt injection can leak credentials and delete files

    Context7 MCP's custom AI instruction feature returns unsanitized attacker content alongside normal document queries, carrying injected instructions into the coding agent's trusted context and tricking it into reading keys, exfiltrating data, or deleting files.

    Why it matters: The material breaks down how Context7 MCP injects prompts through custom instructions, and offers a mitigation approach: adding an authorization gate at the tool invocation boundary.

8/15Sat
  1. Drew Breunig62

    Drew Breunig:编码 harness 是「情境化智能体」

    Drew Breunig 提出用「情境化智能体」来定义 harness:Harrison Chase 所说的系统提示词、规划工具、文件系统和子智能体构成开发者控制的核心循环,harness 则管理循环之外的世界,包括会话、环境、仓库、记忆、Skills、团队、组织和模型这些由内向外、使用人数递增而变动递减的层次。

    Awaiting translation

8/14Fri
  1. InfoQ · AI Coding Presentations80

    InfoQ 演讲:300 个精准 token 胜过 10 万个噪声 token,上下文工程的架构

    Baruch Sadogursky 与 Patrick Debois 在 InfoQ 演讲中用 Claude Code 现场演示:把全部项目文档塞进 CLAUDE.md 后,给接口加错误处理会因约定冲突返回 500,改用按描述懒加载的 Skill 后同一提示词通过测试。

    Awaiting translation

    Why it matters: 两位作者用现场演示拆解上下文工程的四类反模式,并给出 Skill、检索通道、外部记忆与评测的对应做法。

8/13Thu
8/11Tue
  1. Paper Compute · Engineering Blog80

    如何降低 AI Agent 成本:审计 500 个 Claude Code 会话后的三类成本泄漏

    作者用 tapes 记录并导出团队最近 500 个 Claude Code 会话,按工具名和参数做指纹统计,发现同一会话内完全重复的工具调用只占 3.6%(932/25587),真正的开销在别处。

    Awaiting translation

    Why it matters: 作者审计 500 个 Claude Code 会话,把成本拆成会话内、会话间与会话周边三类,给出可复用的排查方法。

8/9Sun
  1. Kondasamy Jayaraman · Engineering Blog74

    Hermes Agent 复盘:为什么我的个人智能体跑在 VPS 上

    作者把 Hermes Agent 装在 VPS 上,因为只在本地跑的工作流会随 Mac 休眠而中断,首个可用版本花了一天调通。他把 Hermes 设计成网关加四个循环,用 ~/.hermes/memories/ 下 2,200 字符的 MEMORY.md 和 1,375 字符的 USER.md 做有界记忆,会话存 SQLite 并用 FTS5 搜索,Skill 按需加载。

    Awaiting translation

8/7Fri
  1. V2EX · Vibe Coding36

    用大模型写代码的自查陷阱:同会话查不出 Bug,新开会话才能发现问题

    有开发者发现,业务代码逻辑稍复杂后,让刚写完代码的 AI 在同一会话中自查漏洞,它大多判定代码不存在 Bug;新建空白会话重新提交审查,AI 往往能找出各类问题,更换另一款大模型检测还能发现更多细节问题,其中不少是影响不大的次要瑕疵。

    Awaiting translation

8/6Thu
8/5Wed
  1. Chen Dahuang · AI Coding 实录66

    DeepSeek V4 Flash 正式版深度体验:便宜、快、1M 上下文、内置搜索

    作者深度体验几天 DeepSeek V4 Flash 0731 正式版后总结:便宜到跑批处理、Agent 循环和几十轮对话账单基本无感,速度快到配合 Agent 工具循环每步几秒内完成,1M 上下文可容纳整个仓库和完整对话历史、无需频繁 compact,结合 Cache 打折长上下文成本还能再降。

    Awaiting translation

8/4Tue
8/2Sun
8/1Sat
7/29Wed
  1. Vibe Built · Blog48

    如何为编程写好 AI 提示词:来自一线开发者的四个步骤

    一位用 Cursor 和 Claude Code 构建并运营真实产品的开发者总结出编程提示词的四个要点:先给上下文再派任务、只交办一个窄任务、写明硬性约束、让 AI 自证结果。他以给 POST /api/submit 路由加限流为例,对比了"给 API 加限流"这类模糊提示与点名文件、复用现有 Redis 客户端、禁止新增依赖、限定 10 次/分钟并只返回 diff 不提交的写法。

    Awaiting translation

7/28Tue
  1. Cursor Forum · Guides62

    一份智能体编程 token 去向图谱:200+ 工具、模式与配置

    作者在 Claude Max 额度 30 分钟内被烧光后做了系统梳理,指出智能体编程的 token 成本主要不在代码生成,而在重复读取同一批文件、智能体重新发现已有上下文、每个任务再派生新任务,以及每一步都默认用最贵的模型。他把成本追踪工具、缓存与路由模式、上下文管理实践、研究和基准整理成一个开源目录,共 200+ 条,每条附一手来源和核验日期,采用 CC-BY 许可,并邀请社区补充修正。

    Awaiting translation

  2. Paper Compute · Engineering Blog48

    如何为 AI 智能体构建基础设施:从 tokenmaxxing 转向 valuemaxxing

    面对 GitHub 等基础设施被 AI 智能体流量压垮的现状,作者主张从 tokenmaxxing 转向 valuemaxxing,用任务完成数、节省时间和避免返工来衡量价值,而非 token 消耗量。他指出 Claude Code 会话默认 30 天后删除,导致已付费的上下文白白流失,并认为 Skill 是比 markdown 文件更好的上下文路由方式,但大规模管理 Skill 仍无解。

    Awaiting translation

7/27Mon
7/26Sun
  1. Hacker News · Context Engineering 讨论82

    Anthropic publishes new context engineering rules for Claude 5

    Thariq Shihipar, a member of Anthropic's engineering team, wrote up the new context engineering rules for Claude 5, saying the team has cut over 80% of the system prompt from Claude Code for models like Claude Opus 5 and Claude Fable 5, with no measurable loss on coding evals.

    Why it matters: Anthropic lays out the new context engineering rules for Claude 5 and explains how to trim the system prompt, CLAUDE.md, and Skills.

7/24Fri
7/20Mon
  1. Addy Osmani · Blog74

    Addy Osmani on software factories: the visible factory and the hidden factory, where validation is the bottleneck

    Addy Osmani proposes that a software factory has three layers—loop, harness, and factory. The factory isn’t a smarter agent; it’s multiple loops with harnesses feeding into a single review gate, with humans controlling the outer loop.

    Why it matters: The author breaks the software factory into three layers—loop, harness, and factory—and points out that validation, not generation, is the real bottleneck.

7/19Sun
7/14Tue
7/12Sun