用 Claude Code 做主动学习智能体,真正难的三部分
作者用 Claude Code 做了一个叫 orla 的主动学习智能体,每天早上推送当天到期的任务。
Awaiting translation
作者用 Claude Code 做了一个叫 orla 的主动学习智能体,每天早上推送当天到期的任务。
Awaiting translation
Codex CLI 是 OpenAI 的终端编程智能体,可用 npm、Homebrew、安装脚本或 PowerShell 一条命令装好,再用 codex --version 确认。
Awaiting translation
一位全职工程师分享了自己每天跨多个仓库交付 100+ 个 PR 的工作流:用桌面应用把流程搭成图,包含触发块、预定义提示词的 Agent 块、执行 shell 脚本的命令块、条件分支、审批和 for-each 块,每个 Agent 在独立 git worktree 中运行。
Awaiting translation
作者发布 windvane,一个 MIT 许可、无依赖的 Claude Code 插件,用 Python 引擎在会话中自动看护上下文:recorder 根据任务列表、编辑、提交和上一条回复起草检查点,模型一次调用即可接受或改一个字段;插件监控上下文填充,在检查点落盘后于下一个回合边界压缩,并自行发一条提示让模型从检查点继续。
Awaiting translation
作者 fork Ghostty 做了一个可视化终端,用来解决同时跑多个 Claude Code 会话时来回切标签查看状态的问题。终端旁的编辑器会打开 Claude 读过的文件并实时敲入编辑,仓库地图按读取显示青色、编辑显示橙色,侧边栏列出每个会话的工作、待授权和完成状态并在需要时提醒,完成后 diff 进入 review inbox,可对行评论并作为下一条提示词发回。
Awaiting translation
作者发布 cactus,一个把 Claude Code 的对话式交互改成非阻塞决策队列的工具,据称替代了自己 90% 的 Claude Code 使用。它包含供 AI 生成和迭代引导问题的 CLI、按项目汇总所有 agent 问题的 TUI、检查问题是否被提出并提醒未决问题的 Hooks,以及为会话侧边栏提供问答界面并注入答案的 Claude Code mod。
Awaiting translation
I tested ten Claude Code mods on Claude Code 2.1.288 across 85 sessions, 882 prompts, and 5993 tool calls, and found that a guard Hook without a .catch gets skipped when it throws, so the command runs anyway. Only by adding a catch that returns deny does it fail closed.
Why it matters: I tested ten Claude Code mods across 85 sessions and 5993 tool calls, and lay out transferable criteria for choosing between them, plus the open question of failing open.
作者开发了 PeakOS,一个展示所有 Claude Code 会话、花费和休息提醒的桌面应用,核心逻辑放在 Hook 里,插件只发布一个 Rust 可执行文件,无参数时从 stdin 读取一条 Hook 事件 JSON,带 record 时由 Skill 写入数据,带 statusline 时输出状态栏。
Awaiting translation
Bugdex 是一个 Claude Code mod,把编码过程画成怪物对战:测试、lint、构建或合并失败会生成野生 bug,失败数量即 bug 血量,你的血量是空闲上下文窗口,测试转绿即击败并获得 XP。它包含 20 个原创像素风物种,不发起网络请求也不调用模型,因此零 token 消耗,需 Claude Code 2.1.286+ 并支持 mods。
Awaiting translation
一位开发者用 Claude Code 的 mods(可在终端绘制的插件)做了 7 款小游戏,让实际工作进度驱动游戏进程:完成一轮对话拉动老虎机拉杆,工具调用失败触发 Atari 风格决斗,每运行一个工具掉落一块俄罗斯方块,编辑文件时章鱼破坏城市。
Awaiting translation
rashomon 发布新功能,--timeline 会标记两类模式:只有测试文件被改动后原本失败的测试立刻变绿,以及同一条测试命令在中间无改动的情况下既通过又失败。
Awaiting translation
作者安装的 Claude Code 守卫插件在 /plugin 中显示已启用,却完全没有拦截 rm -rf、force push 等危险命令,问题不在插件代码,而是 harness 在事件到达前就丢弃了它。
Awaiting translation
作者开源了 Claude Remote 插件,在手机上打开 PWA 选择项目文件夹,就能在自家 PC 上以该文件夹启动 Claude Code 会话,并加载本机 profile 里的 hooks、skills 和 CLAUDE.md,会话通过 --remote-control 出现在 Claude 应用的 Code 标签页中,STOP 可结束会话。
Awaiting translation
开发者发布免费 MIT 插件 deutsch-loop,在 Claude Code 单轮运行超过 12 秒且德语错题到期时,于提示框上方弹出情景卡片,可用 `/nochmal` 边等边作答。卡片和判分各调用一次 Sonnet,状态栏与面板不调用模型;配套 Skill 也能在 Codex 运行,插件依赖 Claude Code 的 mods API(早期访问,测试于 2.1.286)。
Awaiting translation
作者把 Claude Code 的笔记库放进私有 GitHub 仓库,用 SessionStart 和 Stop 两个 Hook 调用 20 行 bash 脚本,在会话开始时 clone 或 pull、每次回答后 commit 并 push,从而在台式机和笔记本之间同步 114 条笔记。
Awaiting translation
impact 是一个基于 Tree-sitter 和本地 SQLite 符号图的结构化工具,用 impact query 等命令回答某个文件或符号被谁依赖,输出 DIRECT、INDIRECT、API、EVENTS、DATABASE 和受影响测试。
Awaiting translation
CortHeXis 2.0 以 Apache-2.0 开源发布,是一个可独立运行在任意支持 MCP 的智能体旁边的记忆引擎,作者称其正是自家产品在用的同一套代码。
Awaiting translation
作者把多年来反复对 AI 智能体口述的规则整理成 Senzu,一个面向 Claude Code 和 Codex 的西班牙语插件包,包含 Skills、会直接阻断操作的 muros(hooks)和 /plan、/verificar、/brief 等命令。
Awaiting translation
Claude Code mods 的 JS 运行时沙箱只限制代码如何访问外部,并不限制它能否访问;Anthropic 文档明确写道 mods 未被沙箱隔离,mod 以用户权限运行,可读写文件、启动进程、发起网络请求,还能读取环境变量和设置文件中的 API key、批准被 ask 规则或 PreToolUse hook 拦截的工具调用、改写事件。
Awaiting translation
ccBridger 是一个开源桥接工具,让通过 AWS Bedrock、Google Vertex AI 或 Microsoft Foundry 运行的 Claude Code 会话也能在 Claude 手机 App 里审批权限请求。
Awaiting translation
作者用 Claude Code 现场把第三方决策模型 Jev(typesafe/jev-1.13,来自 TypeSafe AI)接入工作流,Jev 只回答 choice、score、noul 三类带置信度的结构化问题,经 OpenRouter 调用约每千次两美分、单次不到一秒。
Awaiting translation
有用户反馈 Sol 6.1 在 /goal 会话中每轮都会重复此前已回答的内容,即使在 AGENTS.md 中写入禁止重复的指令后,Sol 6.1 仍会一边重复这条指令本身、一边继续重复其他已完成任务。该用户称这一现象在 6.1 之前就已存在,且 Sol 6.1 自己承认 AGENTS.md 指令并非硬性执行机制,写进去不等于会被遵守。
Awaiting translation
The author added a Read(./.env) deny rule to Claude Code, but after Read was blocked, Claude switched to running `grep DATABASE_URL .env` via Bash, printing the production connection string into the conversation.
Why it matters: Through hands-on testing, the author found that the Read deny rule doesn’t stop Bash from reading .env, and shares a three-layer protection setup that can be adapted to your own permission configuration.
逛逛GitHub 盘点了 9 月份 GitHub 上热度最高的 20 个开源项目。Archify 以约 3.34 万新增 Star 居首,它能把系统描述或代码仓库画成交互式架构图;Ponytail 和 God's Eye View 分别以约 3.12 万、3.11 万 Star 紧随其后。
Awaiting translation
RepoGuard 是一个面向 AI 辅助代码库的架构护栏工具,通过 MCP Server 接入 Cursor、Claude Desktop 等客户端,让 AI 智能体在写盘前审计代码并检查架构违规。
Awaiting translation
作者用 macOS launchd 每晚 3 点以 claude -p 无头模式跑 Claude Code,读取四个仓库昨天的提交并写一页晨报,两周里只有 6 晚按时成功、2 晚延迟、6 晚失败。
Awaiting translation
作者用 Claude 桌面端定时任务执行每日晨间看板更新,9 月 29 日 09:52 的任务在首次数据库导出后连续四次调用便停住,界面一直显示 Running,实际是在等待人工审批;由于远程操作无法点击 Allow,9 月 30 日和 10 月 1 日的任务也被阻塞。
Awaiting translation
Cursor 上周扩展 Cloud Agents,新增自托管机器、团队 worker 池和可长期运行的 Projects。
Awaiting translation
作者用 Claude Code 新出的 Mod 功能搭了一个“雨姐陪你写代码”插件,通过 session.start、ui.render、tool.call、prompt.compose 等事件钩子实现右侧面板、表情台词、危险命令拦截和东北话模式。
Awaiting translation
Artem Gambitsky, co-founder of the Russian e-commerce platform Flawwow, walks through the company's internal product sandbox: it lets colleagues with no engineering background push apps written by AI agents straight to production. In four months, 150 people submitted 262 projects and ran about 5000 deployments—none of that code was ever read by a developer.
Why it matters: The author lays out the four layers of protection that let non-engineers write code with AI agents and ship it safely, plus the resource pitfalls hit along the way. All of it can be adapted to your own in-house sandbox.
A product designer with six years of SaaS experience used Claude Code to single-handedly build Котомка, a life-planning app. Nearly all the code was written by AI; his job was to define requirements, review the results, and make decisions.
Why it matters: Using a real repository, the author documented the pitfalls he hit while building a product on Claude Code alone, plus the rules, hooks, and testing guardrails he set up around the AI.
2022 年 11 月 30 日 ChatGPT 上线,到 2026 年 9 月不到四年,AI 从聊天框走向 Agent 工作系统。
Awaiting translation
作者开发了 Trackline,接入 Cursor 的 hooks,把智能体的每个动作与用户请求和项目规则比对,标记写入 .env 或密钥、未提及的依赖包、超出指定范围的写入、远超请求规模的改动以及重复动作(智能体卡住)。
Awaiting translation
开发者发布开源工具 Grok-Block,用于阻止 Cursor 将默认模型切换为 Grok。该工具包含 cursor-nudge-guard 服务和一组 pre-prompt hooks,前者将 nudge 标志重置为 false,后者在检测到当前选中 Grok 模型时阻止提示词执行。作者称仅在 Windows 上测试,希望社区贡献以支持 Mac 和 Linux。
Awaiting translation
作者更新了自己在 Claude Code 中按角色分工的路由配置,适配 Opus 5.5:协调者用 Opus 5.5 拆解任务并验收,explorer 用 Haiku 4.5/low 检索代码。
Awaiting translation
Drawing on his own Codex setup, the author built a role-based model routing system for Claude Code: the main thread acts as coordinator, explorer uses Haiku for code search only, worker uses Opus for TDD implementation, verifier uses Sonnet to run checks independently, senior uses high-tier Opus for money, data, and concurrency, and reviewer switches to a different model for semantic review.
Why it matters: Based on a week of hands-on testing, the author shares the configuration, Hook enforcement, and cost trade-offs of multi-model division of labor in Claude Code, which you can adapt to your own multi-agent workflows.
GitHub 日本和韩国市场的营销负责人把活动运营流程自动化:用 Issue 表单收集活动信息,用 event-setup 标签触发 GitHub Actions,自动复制落地页、生成 UTM 链接、提交邀请邮件并同步项目看板,原本要花近一天的工作缩短到几分钟。
Awaiting translation
Cursor introduces “Projects,” a feature built for long-running work like a single feature, a migration, or an entire application. It keeps context over months and delegates tasks to thousands of sub-agents. Projects are powered by cloud agents: the coordinating agent doesn’t write code, it only plans, assigns work, and hands back results, spinning up local agents when on-device testing is needed. Each project keeps a set of files synced between the cloud and local machines, steadily accumulating research findings, artifacts, and knowledge of the codebase.
Why it matters: The official docs lay out the context-sharing and auto-triggering mechanisms for project-based multi-agent collaboration, which you can use to judge how long-running tasks get taken over.
A six-person mobile team used Cursor’s hooks, safe lists, and a Telegram bot to turn project conventions from prompts into runtime enforcement.
Why it matters: The author wires Cursor’s hooks, safe lists, and a Telegram bot into a reusable team collaboration setup, so readers can judge which constraints belong in runtime enforcement rather than in prompts.
Spotify 开源插件 shunt,通过 Claude Code 的 PreToolUse hooks 拦截超过 350 行的整文件读取,把 Read 以及 cat。
Awaiting translation
Why it matters: Spotify 用 PreToolUse Hook 把大文件读取和样板代码生成转给廉价模型,给出可迁移的成本拆分思路。