在 GPT-6.1 Sol 发布后账号用量重置、VS Code 尚未提供该模型的情况下,一位开发者改用 Codex CLI,并行运行三个终端:GPT-6.1 Sol Max Fast 负责实现与最终决策,GPT-6.1 Sol Max 负责审查,Astra Max Fast 负责深度推理与升级。三个会话通过 MCP 工具协调,作者用 /status 复制状态上下文在主实现者会话中继续推进。
有用户在组织内用 Claude Code 自动化工作流时遇到一致性问题:Claude 常搜索所有项目找参考,导致每次实现结果不同,还会绕过预校验模板和渲染函数,直接手写组件。即使用 skills 文件要求它使用组件库,也只是偶尔生效。该用户希望找到办法,让 Claude 稳定判断组件是否匹配用户需求,并改用组件库 SDK。
In a product project where an AI agent writes the code and the author doesn't read it, automated checks repeatedly reached the wrong conclusion. The author found 86 checks that no workflow had ever triggered, a secret scan that missed 438 of 1413 files because Git escapes Russian filenames by default, a new check that mistook WHERE for a table alias and let an injection slip through, and three false alarms from the test dashboard and the agent's replica.
Why it matters: The author walks through five real cases to show why automated checks produce false greens or false reds, and lays out validation rules that carry over to other projects.
For coding agents running unattended, the author built a checkpoint mechanism based on hidden git refs. Before each task starts, it snapshots the entire working tree—including untracked files—and rolls back automatically when validation fails. The restore operation itself can also be undone.
Why it matters: With roughly 40 lines of shell, the author turned git checkpoints into rollback-capable infrastructure, laying out the concrete approach and the limits of running coding agents unattended.
The author has open-sourced the Project Athena v9.9.9 kernel (MIT License), a local-first memory and governance framework built to solve the problems of Claude Code needing repeated corrections, sub-agents overstepping their bounds and modifying files, and losing state when it hits the quota limit mid-task.
Why it matters: The author validated a CLAUDE.md slimming and local-memory approach across 1900 sessions, and provides a directory structure and verification rules you can reuse directly.
AI Hero Skills v1.3 is out, adding three skills—/implement-spec, /pr, and /retro—and extending the main flow from /grill-with-docs → /to-spec → /to-tickets into implementation, PRs, and retros.
Why it matters: The author rounds out the skill set into a full path from writing specs to PRs and retros, and lays out the trade-offs at each step along with the rough edges they already know about.
作者分享自己把 Claude Code 接到免费 LLM API 上的做法,通过设置 ANTHROPIC_BASE_URL 和 ANTHROPIC_AUTH_TOKEN 两个环境变量,把请求路由到 Google AI Studio 的 Gemini 2.5 Flash 或 Groq,Cursor 和 Codex CLI 也有对应的 base URL 配置方式。