用 Codex CLI 通过 Tailscale 连接 NVIDIA DGX Spark 上的 gpt-oss:120b
Simon Willison 在 Mac 上通过 Tailscale 网络,让 OpenAI Codex CLI 调用运行于 NVIDIA DGX Spark 上的 gpt-oss:120b 模型。
Awaiting translation
Simon Willison 在 Mac 上通过 Tailscale 网络,让 OpenAI Codex CLI 调用运行于 NVIDIA DGX Spark 上的 gpt-oss:120b 模型。
Awaiting translation
继 Shortcut 和 Microsoft 的 Agent Mode 之后,Claude for Excel 的推出让 Excel 智能体成为新热点。作者估算,美国约 7090 万管理及专业岗位从业者中,38% 的工作时间花在 Excel 上,对应约 2.4 万亿美元年人力成本,即使只提升 50% 效率,也意味着至少 1 万亿美元被浪费的工时。
Awaiting translation
Author Jesse Vincent spent an afternoon porting Superpowers and the whole SKILL.md system to the OpenAI Codex CLI, shipping it with Superpowers 3.3.0.
Why it matters: The author ported Claude's SKILL.md system to the Codex CLI, with tool mappings and install instructions, so you can judge whether reusing Skills across models is feasible.
Dagster Labs 分享了用 OpenAI Codex 加速技术文档写作、跨媒介内容转换和文档覆盖度评估的实践。他们重写了 CONTRIBUTING.md,明确文档层级、结构和最佳实践,让 Codex 能据此生成符合规范的文档;还借助 gh 命令让 Codex 解读 PR 的 diff 和描述,并让 Codex 把教程改写成 YouTube 视频脚本。
Awaiting translation
Why it matters: Dagster 团队把 Codex 用于文档写作、PR 解读和内容跨媒介转换,其中用文档生成代码来反向衡量文档覆盖度的做法可以迁移。
作者用 ChatGPT Atlas 的 Agent 模式自动清理 Facebook 信息流,通过一段提示词让智能体持续滚动页面、隐藏赞助帖并对含 Follow/Join 链接的帖子点“不感兴趣”,最终信息流只剩自己关注的人发布的真实帖子。作者对 AI 浏览器的提示注入风险仍持警惕,表示目前不会把银行或邮箱凭据交给这类浏览器。
Awaiting translation
Jesse Vincent 为 Claude Code 创建新插件 superpowers-lab,用作开发中 Skill 的孵化器。
Awaiting translation
一位非技术 CFO 借助 Claude Code 搭建出整合多个系统的内部运营看板,此前用 AirTable、低代码工具和 retool 外包机构都因规模或业务知识传递问题失败。
Awaiting translation
Anthropic rolled out its first-party Skills system simultaneously on Claude Code, Claude.ai, and the Claude API, and author Jesse Vincent quickly followed with a new version of Superpowers built on the official Skills.
Why it matters: Drawing on nearly a month of hands-on use, the author compares the official Skills with his own setup and lays out the trade-offs involved in migrating.
Martin Fowler 试用 Kiro、spec-kit 和 Tessl 三款自称实现 spec-driven development(SDD)的工具,把 SDD 归纳为 spec-first、spec-anchored、spec-as-source 三个层次,并指出目前所有方案都停留在 spec-first。
Awaiting translation
Why it matters: 作者亲手试用 Kiro、spec-kit 和 Tessl 三款 SDD 工具,给出 spec-first、spec-anchored、spec-as-source 三层划分,并指出小任务被过度规格化的问题。
开发者 Jesse Vincent 发现,让 Claude 先读 Strunk 1920 年版《The Elements of Style》再写 README,成稿比原来短约 30%,文风也更合他意。他把该书 HTML 转成约 12,000 词的 Markdown,因 Anthropic 的版权过滤机制拒绝处理这本已进入公有领域的书,最终改用 GPT-5 Codex 完成删减。
Awaiting translation
Jesse Vincent 把工作流打包成 Superpowers 的 Skill 后,将提示词中对用户的称呼从“Jesse”改为“your human partner”,结果至少在一个案例中,Claude 的沟通方式发生变化,开始以“Human”称呼 John Dimatos 并给他派发任务。
Awaiting translation
OpenAI 第三届 DevDay 首次全程用 Codex 参与搭建,从舞台演示、社区大厅街机到产品本身都由它协助完成。Codex 自行实现 VISCA 协议控制网络摄像头,并搭建灯光 MCP server;Codex CLI 让 Romain Huet 一个下午就完成初版。
Awaiting translation
Author Jesse Vincent released Superpowers, a set of Skills built on Claude Code's new plugin system. Once installed, it injects a guiding prompt through the session-start hook, prompting Claude to proactively search for and use these Skills.
Why it matters: The author packaged his own coding-agent workflow into an installable Skill plugin, so readers can directly reuse his implementation flow from brainstorming to TDD.
The author walks through their full workflow with Claude Code: first isolating tasks with git worktree, then using a brainstorming prompt to make Claude ask only one question at a time and confirm the design in stages, and finally using a planning prompt to break the plan into small tasks and write them into docs/plans/.
Why it matters: The author splits Claude Code into two sessions—an architect and an implementer—and shares reusable prompts plus a git worktree approach for isolating tasks.
Anthropic's applied AI team argues that context engineering is a continuation of prompt engineering, and the core idea is picking the smallest set of high-signal tokens within a limited attention budget.
Why it matters: Anthropic lays out a systematic approach to context engineering, covering the trade-offs among three long-task strategies: compression, note-taking, and sub-agents.
Harper Reed 团队为 Claude Code 等编程智能体搭建了社交媒体服务器 botboard.biz,让它们在工作时发帖、回复并互相吐槽。团队随后测试了这些社交工具对智能体表现的影响,结果发现它反而成了性能增强器,相关论文已发布。botboard.biz 即将开放试用。
Awaiting translation
The author rewrote a long block of CLAUDE.md rules as a GraphViz dot flowchart, using quoted strings as node names, different shapes to distinguish decisions, commands, and warnings, and giving each flow an explicit trigger condition.
Why it matters: After rewriting the CLAUDE.md rules as a GraphViz dot flowchart, Claude followed the rules better, and this approach can be carried over to your own projects.
作者认为 GitHub Actions 默认 runner 的 2vCPU 实际只是共享物理核的一个线程,实测单线程性能约为 Ryzen 9950X3D 的一半,磁盘读写约 200MB/s、IOPS 约 1 万,远低于 PCIe5 NVMe 的 6000MB/s 和百万级 IOPS。
Awaiting translation
Martin Fowler 分享用参考应用(reference application)给编码助手提供可编译、一致的代码样例,做法是建一个 MCP server 让助手访问 Spring Boot 的 repository、service、controller 等典型模式样例,替代在 markdown 里手写代码片段。
Awaiting translation
The Manus team shares context engineering lessons from building AI agents, centered on designing around the KV cache, managing tools by masking rather than removing them, treating the file system as context, steering attention by restating to-do items, keeping errors in context, and avoiding getting stuck on few-shot examples.
Why it matters: The Manus team distilled lessons from rewriting their agent framework four times into six context engineering principles—ready to apply directly to your own agent implementation.
OpenAI Codex Cookbook 给出把 Codex CLI 接入 GitLab CI/CD 的完整做法,用于生成 CodeClimate JSON 代码质量报告、把 SAST 结果整理成 security_priority.md,并让 Codex 输出可 git apply 的补丁。
Awaiting translation
Why it matters: 官方 Cookbook 给出把 Codex CLI 接入 GitLab CI 的完整配置,含提示词约束、JSON 标记提取与 diff 校验,可直接照搬。
ECMAScript 2025 的 Iterator 类型让 JavaScript 向 Ruby Enumerable 和 Rust Iterator 的惰性链式迭代靠拢。
Awaiting translation
Ryan Lopopolo 将 hyperbo.la 博客的构建系统从 Bazel 迁回 Vite 和 Tailwind CSS,此前该博客仓库超过 25% 是 Starlark 代码。他称改用现代前端技术后,切换到 Tailwind CSS 只需几小时,而过去可能要几天,并保留了浅色/深色模式和原有品牌外观。
Awaiting translation
服务网格的价值在于以难以绕开的方式注入策略,包括连接池与 keep alive、重试、超时、数据本地性约束和 AZ 本地路由等。产品工程追求快速迭代与可靠性之间存在一定冲突,而服务网格难以被绕过,能确保默认情况下做正确的事。这也是部署服务网格最简洁的理由:它实现了职责分离,又不需要产品团队额外记住步骤。
Awaiting translation
Google 的 gemini-embedding-001 正被 Box、re:cap、Everlaw、Roo Code、Mindlid、Poke 等产品用于 RAG 与上下文工程。
Awaiting translation
作者以自己用 n8n 搭建的多智能体深度研究应用为例,说明上下文工程不只是写提示词,而是设计并优化提供给 LLM 的完整上下文。他拆解了搜索规划智能体的系统提示词,涵盖指令、用户输入分隔符、子任务字段定义、结构化输出示例,以及注入当前日期时间让模型推断 start_date 和 end_date 的做法。
Awaiting translation
作者认为 AI 领域正从提示词工程转向上下文工程,即设计动态系统,在合适的时间以合适的格式为 LLM 提供正确的信息和工具。他把上下文拆成系统提示词、用户提示词、短期状态与历史、长期记忆、RAG 检索信息、可用工具和结构化输出七类,并指出多数智能体失败已不是模型失败而是上下文失败。
Awaiting translation
作者分享了自己 2025 年 6 月在 Claude Code 中写代码的完整流程:先让模型反复提问并写出计划草稿,确认后把计划存成 docs/plans/somefeature.plan.md,再 /clear 清空上下文,让模型读计划、提问、更新计划后开始写代码。
Awaiting translation
Martin Fowler 给 OpenAI Codex 布置了一个前端标签格式化的化妆类小任务,并完整公开了 Codex 的日志和生成的 PR。日志显示 Codex 主要靠 grep 反复文本搜索定位代码,中途因把 AGENTS.md 误写成 AGENT.md 来回折腾,还因删掉 .yarnrc 导致测试无法运行,最终 PR 里有两个回归测试失败。
Awaiting translation
Why it matters: 作者完整记录 Codex 自主完成一次前端小任务的日志,并对比 6 次运行结果,展示后台编码智能体在环境配置和代码复用上的真实短板。
Jesse Vincent 用 Claude Code 设计了一个名为 process_feelings 的 MCP 工具,让 Claude 在与用户交互后把内心想法写入工作区 .private-journal 目录下按日期命名的 markdown 文件。
Awaiting translation
Jesse Vincent 用 Raspberry Pi Pico 和一只廉价 Staples Easy Button 改装出专为 AI 编程助手设计的单键键盘,按下即向 Claude Code、Cursor、Cline 等工具发送 "continue",让 AI 接着干活。他还在按钮里保留了原装扬声器播放励志语音,但指出音效很快就让人厌烦,取出电池后键盘功能仍可正常使用。
Awaiting translation
作者用 LLM 构建了 PlantUMLSteps,为 PlantUML 时序图增加按步骤播放功能,把单张复杂图拆成登录、认证、仪表盘等逐步展示的步骤。开发中 Claude 在 Cursor Agent 模式下生成 StepParser 解析逻辑、Gradle 任务和 HTML 查看器,解析器最初在 newPage 属性、首个步骤标记前的声明处理上出错,经测试反馈修正后通过。
Awaiting translation
作者分享了自己迁移到 Claude Code 后的完整开发工作流:先用 gpt-4o 打磨想法,再用 o1-pro 或 o3 生成 spec.md 和 prompt_plan.md,然后让 Claude Code 逐条执行未完成的 prompt、跑测试、提交 git 并更新计划文件。
Awaiting translation
作者借助 Claude 等 LLM,在毫无 D3.js 和 UI 开发经验的情况下,为 Thirty Meter Telescope 主镜搭建了一个交互式可视化原型,用 492 块六边形镜面拼出对称的圆形蜂窝结构。
Awaiting translation
Harper Reed 结合自己从 2023 年起的实践,把采用 AI 辅助编程拆成一条九步路径:先用 IDE 自动补全,再把代码复制粘贴到 Claude 或 ChatGPT,接着用 Cursor 等 AI IDE,然后先写规格再写代码,再用 aider 加快循环,最后进入全自动智能体编码。
Awaiting translation
Martin Fowler 团队以 Java ByteBuffer 读写页头为例,展示开发者如何引导 LLM 把功能正确但不安全的代码逐步改造成健壮组件。
Awaiting translation
Harper Reed 认为 AI 代码生成正把开发推向"15 分钟瀑布":先写清规格、让模型生成、再审查,并可同时跑多个智能体分别写功能、文档和测试。他建议团队先用低风险内部项目做小范围试点、轮换成员参与,并把规格和架构文档沉淀到共享仓库。他还提醒 AI 容易过度测试基础逻辑,需频繁提交以便回滚。
Awaiting translation
Jesse Vincent 借助 Claude 在几天内做出一个定制博客客户端 Post Through It,可直接与 GitHub 交互,把文本文件推送到 git 仓库并触发 Eleventy 编译成网站。这是他首次用该工具发布文章,此前他因不熟悉 Swift/Cocoa 一直没能做出图形化博客客户端。
Awaiting translation
Vibinex 创始人指出,团队采购 CodeRabbit、CodeAnt、Greptile、Ellipsis 等 AI 代码审查工具,主要优化的是作者写代码的质量,而非减少审查者逐行读代码的时间。审查者仍需自行理解改动、识别业务逻辑与架构影响,原有流程一步没少。作者工具与审查者工具应互补,而非互相替代。
Awaiting translation
Harper Reed 用 Claude、Aider 和 ChatGPT 为自己的网站新增了一个媒体板块,自动追踪并展示他在 Goodreads 上的已读书籍、Spotify 上最近保存的曲目,以及通过 feedbin/NetNewsWire 收藏的链接。数据由几个脚本抓取后存为 YAML 文件,再生成 Hugo 博客条目,并提供了全量、书籍、音乐、链接四个 RSS 订阅源。
Awaiting translation