用 Mac Mini 自建语音助手:摆脱 Alexa 和 Google
作者因 Amazon 在 Echo 上频繁推销 Alexa Plus 订阅,用一台 Mac Mini 自建语音助手,架构为麦克风→唤醒词→语音转文字→LLM+工具→文字转语音→扬声器,各组件可替换。
Awaiting translation
作者因 Amazon 在 Echo 上频繁推销 Alexa Plus 订阅,用一台 Mac Mini 自建语音助手,架构为麦克风→唤醒词→语音转文字→LLM+工具→文字转语音→扬声器,各组件可替换。
Awaiting translation
Skyscanner 工程师把 OpenAI 的 Codex CLI 接入 JetBrains IDE 的 MCP server,让 Codex 能调用 IDE 的 get_file_problems 检查文件错误、执行预设的 run configurations 跑测试和 lint。
Awaiting translation
Why it matters: Skyscanner 工程师把 Codex CLI 接入 JetBrains MCP,让 AI 直接读取 IDE 报错并跑测试,读者可借鉴这套反馈闭环。
作者介绍如何用 Gemini 2.5 Flash 做图像目标检测,让 AI 工作流不仅能描述图片,还能返回物体所在坐标。做法是给模型一张图和自然语言提示词,要求返回含 label 与 box_2d 的 JSON 数组,坐标按 0-1000 归一化,再用 normalize_to_pixels 换算回真实像素。
Awaiting translation
作者用 RosettaCode 数据集中 19 种语言都有的编程任务,配合 Hugging Face 上 Xenova/gpt-4 分词器,比较各语言的 token 效率。
Awaiting translation
作者介绍自己从 iPhone 远程运行 Claude Code 的完整方案,需要解决网络、终端客户端、工作站和工具四件事。网络用 Tailscale 打通任意设备到工作站的连接,终端客户端选 blink,工作站是一台持续供电、网络良好的 Mac;工具层面用 SSH 密钥、Mosh 保持断线后连接可恢复、TMUX 让多个 Claude Code 会话长期运行并随时重连。
Awaiting translation
作者为 Discord 游戏项目搭建了一条 AI 内容流水线,用 Claude Code Skill 把图像生成、Discord 表情包转换和 Sora 视频动画串成可复用工具。
Awaiting translation
作者把 1990 年的 Photoshop 1.0 源码(约 10 万行 Pascal 加约 2 万行 68k 汇编)交给 Claude Code,让它先探索代码库并制定移植计划,再让多个并行子智能体分头实现,30 分钟后得到一个能在现代 macOS 上运行的 C# 版本。
Awaiting translation
Dragos’ investigation shows that attackers used Claude Code and OpenAI GPT-4.1 to target the OT environment of a Mexican water company. Claude Code handled broad discovery, identifying vNode industrial gateways, researching vendor credentials, generating password lists, and executing password spraying, while GPT-4.1 handled structured data analysis and Spanish-language output.
Why it matters: Dragos reconstructed the full chain of how attackers used Claude Code and GPT-4.1 to conduct reconnaissance and password spraying against a Mexican water utility’s OT environment, showing how AI was actually divided across the intrusion lifecycle.
Zenity Labs 披露 Microsoft Copilot Studio 的 Connected Agents 功能默认在所有新建智能体上开启,恶意智能体可静默调用同一环境中受信任智能体的工具,且目标智能体的活动标签页不产生任何记录。
Awaiting translation
作者 Martin Alderson 认为 MCP 的核心问题在于 token 消耗过大,转而用自建 CLI 替代。他举例 Playwright MCP 开箱即用就要占用近 15000 token,超过 Claude Code 上下文窗口的 10%,而 Linear MCP 也因工具定义过多频繁触发上下文上限。
Awaiting translation
美国旅行社数量从 2000 年的 124,000 家降至 2012 年的 65,000 家,历时十年;而 LLM 在软件工程领域的采用率从 2022 年的 0% 升至 2025 年的 84%,速度远超当年互联网对旅行社的冲击。作者认为,只会把需求手动翻译成代码的通用型开发者,处境类似当年被 OTA 淘汰的普通旅行社代理。
Awaiting translation
作者 Matthew Fontana 分享自己不再逐行审查 AI 生成的代码,而是用 Playwright MCP 让智能体截图证明功能可用。他给出的做法是直接提示智能体导航到页面、截图、点击并截图结果,例如让智能体截取空表单、填入非法数据截取校验错误、再正确填写截取成功状态,三张图即可判断表单是否可用。
Awaiting translation
作者把 Wordiest 最后一个 Android APK 交给 Codex + ChatGPT 5.2,半小时内得到可玩的核心游戏,几小时后完成 Android 与 iOS 版本,全程未写也未读一行代码。
Awaiting translation
Why it matters: 作者用 Codex 反编译 Android 游戏并移植到 iOS,展示了智能体开发中难易直觉失效的真实体验。
作者 Martin Alderson 用 acorn 生成两个 npm 包版本的 AST,让 Claude Code 对 AST 做 diff 并派出 10 个子智能体分头分析,不到 10 分钟就产出一份报告,涵盖功能开关、未发布功能、日志与遥测细节以及内部架构。
Awaiting translation
Superpowers 4.0 发布,核心改动是把实现步骤后的代码评审拆成两个智能体:先由 spec review 智能体确认实现符合计划,通过后 code review 智能体再检查代码质量,两步都改为循环执行,协调智能体知道实现者修复后要重跑评审。
Awaiting translation
据 Financial Times 报道,中国区一次持续 13 小时的 AWS 服务中断被归因于使用 Amazon Kiro AI 编码智能体时的用户操作失误。Amazon 据称将该事件描述为影响极其有限。报道指出,生产环境的删除、重建或发布变更本应经过签名审批流程,但公开信息未披露具体控制边界,因此无法确认智能体是否触及部署代码、内部工具或发布操作。
Awaiting translation
DeepSource(YC W20)团队发布 Autofix Bot,一个把静态分析与前沿 AI 智能体结合的代码审查智能体,面向 AI 编码工作流。其混合架构分三步:5000+ 确定性检查器建立高精度基线并由子智能体抑制误报,AI 审查以静态发现为锚点并调用 AST、数据流图、控制流、导入图等工具,最后由子智能体生成修复、静态校验后输出干净的 git patch。
Awaiting translation
Jesse Vincent has released an open-source tool called packnplay. With a single command — `packnplay run claude --dangerously-skip-permissions` — it spins up a pre-configured throwaway container to run a coding agent.
Why it matters: The author wraps all the tedious setup for running a coding agent in a container into one command, and also shares exactly how he handles credential conflicts with Claude Code.
When developing web apps with the coding agent, the author often runs into client-side JavaScript bugs. If the agent can't fix them by reading the code, it fires up browser MCP for interactive debugging just to see the browser console logs—burning tokens and slowing things down.
Why it matters: The author shares a reusable front-end/back-end log bridge approach that lets the coding agent see front-end logs without browser MCP.
作者 Martin Alderson 认为当前模型进步是一次更微妙的 GPT-4 时刻,但现有基准测不出来。他指出 Gemini 3 Pro Preview 在设计网页和落地页上明显强于其他模型,并给出流程:上传产品 CSS 让模型提取设计系统,再结合产品截图生成 HTML 原型。
Awaiting translation
作者把自己的软件开发工作流工具集 Superpowers 移植到了开源智能体编程工具 OpenCode。
Awaiting translation
作者发布 Claude Code 插件 Double Shot Latte(DSL),用 Stop hook 在 Claude 想停下来请求人工确认时,把最近几条消息交给另一个 Claude 实例判断是否真的需要人介入,倾向于让它继续工作;若 Claude 在五分钟内三次尝试停止则放弃接管。
Awaiting translation
In the Codex Cookbook, OpenAI lays out a complete workflow for modernizing a legacy codebase with Codex CLI, using a COBOL portfolio system as the example and moving through five phases built around an ExecPlan design document.
Why it matters: Using a COBOL portfolio system as the example, it offers reusable documents and a validation workflow for modernizing legacy code in phases with Codex CLI.
Terminal-Bench has released version 2.0 and the Harbor package. The former is a more rigorously validated, harder benchmark for evaluating agents; the latter is for evaluating and optimizing agents. Harbor rewrites Terminal-Bench's test harness, supports deploying containers in the cloud, provides rollout interfaces for RL and SFT, and works with any agent you can put in a container.
Why it matters: Terminal-Bench 2.0 and Harbor are released together, so readers can see how the agent evaluation benchmark is validated and how to scale it in the cloud.
作者 chrisloy 提出,随着 LLM 从聊天机器人变成复杂系统的决策组件,提示词工程正让位于上下文工程,即动态、有针对性地设计喂给模型的每一个 token。
Awaiting translation
Author Jesse Vincent spent an afternoon porting Superpowers and the whole SKILL.md system to the OpenAI Codex CLI, shipping it with Superpowers 3.3.0.
Why it matters: The author ported Claude's SKILL.md system to the Codex CLI, with tool mappings and install instructions, so you can judge whether reusing Skills across models is feasible.
作者用 ChatGPT Atlas 的 Agent 模式自动清理 Facebook 信息流,通过一段提示词让智能体持续滚动页面、隐藏赞助帖并对含 Follow/Join 链接的帖子点“不感兴趣”,最终信息流只剩自己关注的人发布的真实帖子。作者对 AI 浏览器的提示注入风险仍持警惕,表示目前不会把银行或邮箱凭据交给这类浏览器。
Awaiting translation
作者 mbleigh 认为上下文工程普遍忽视了超链接这一手段,并提出只需一个接受 URI 列表的 read_resources 工具加一个入口 URI,就能让模型按需递归加载上下文。
Awaiting translation
一位非技术 CFO 借助 Claude Code 搭建出整合多个系统的内部运营看板,此前用 AirTable、低代码工具和 retool 外包机构都因规模或业务知识传递问题失败。
Awaiting translation
Martin Fowler 试用 Kiro、spec-kit 和 Tessl 三款自称实现 spec-driven development(SDD)的工具,把 SDD 归纳为 spec-first、spec-anchored、spec-as-source 三个层次,并指出目前所有方案都停留在 spec-first。
Awaiting translation
Why it matters: 作者亲手试用 Kiro、spec-kit 和 Tessl 三款 SDD 工具,给出 spec-first、spec-anchored、spec-as-source 三层划分,并指出小任务被过度规格化的问题。
作者发布 Superpowers 2.0,把 skills 抽成可 fork 和本地管理的独立 git 仓库,方便添加、定制和分享 skills。随后他在 claude --debug 日志中发现 Claude Code 会把 SKILL.md 的 YAML name 字段转成 SlashCommand,并认为可以自动激活它们,这比他设计的引导方式更可靠。
Awaiting translation
作者参加伦敦 MCP Dev Summit Europe,记录 Anthropic 的 David Soria Parra 介绍 MCP 标准进展,其重点方向是未来一年的 Agentic Discovery 愿景,即由 LLM 自行发现并安装 MCP。
Awaiting translation
Anthropic's applied AI team argues that context engineering is a continuation of prompt engineering, and the core idea is picking the smallest set of high-signal tokens within a limited attention budget.
Why it matters: Anthropic lays out a systematic approach to context engineering, covering the trade-offs among three long-task strategies: compression, note-taking, and sub-agents.
The author rewrote a long block of CLAUDE.md rules as a GraphViz dot flowchart, using quoted strings as node names, different shapes to distinguish decisions, commands, and warnings, and giving each flow an explicit trigger condition.
Why it matters: After rewriting the CLAUDE.md rules as a GraphViz dot flowchart, Claude followed the rules better, and this approach can be carried over to your own projects.
作者认为 GitHub Actions 默认 runner 的 2vCPU 实际只是共享物理核的一个线程,实测单线程性能约为 Ryzen 9950X3D 的一半,磁盘读写约 200MB/s、IOPS 约 1 万,远低于 PCIe5 NVMe 的 6000MB/s 和百万级 IOPS。
Awaiting translation
Martin Fowler 分享用参考应用(reference application)给编码助手提供可编译、一致的代码样例,做法是建一个 MCP server 让助手访问 Spring Boot 的 repository、service、controller 等典型模式样例,替代在 markdown 里手写代码片段。
Awaiting translation
作者 Martin Alderson 发现 Google AI Studio 的 Gemini API 近两周频繁出现 503 模型过载错误,而官方状态页没有报告。
Awaiting translation
The Manus team shares context engineering lessons from building AI agents, centered on designing around the KV cache, managing tools by masking rather than removing them, treating the file system as context, steering attention by restating to-do items, keeping errors in context, and avoiding getting stuck on few-shot examples.
Why it matters: The Manus team distilled lessons from rewriting their agent framework four times into six context engineering principles—ready to apply directly to your own agent implementation.
OpenAI Codex Cookbook 给出把 Codex CLI 接入 GitLab CI/CD 的完整做法,用于生成 CodeClimate JSON 代码质量报告、把 SAST 结果整理成 security_priority.md,并让 Codex 输出可 git apply 的补丁。
Awaiting translation
Why it matters: 官方 Cookbook 给出把 Codex CLI 接入 GitLab CI 的完整配置,含提示词约束、JSON 标记提取与 diff 校验,可直接照搬。
Kilocode 作者认为 AGENTS.md 这类规则文件正让开发者更愿意写文档,因为写一次就能立刻让 AI 编程助手更懂项目。
Awaiting translation