Skip to content

All updates

23 items today
Today10/6Tue
  1. Reddit · ClaudeCode / Codex / VibeCoding22

    Claude Code 在 VS Code 中为最基础的 UI 改动启动预览智能体?

    有用户在 VS Code 中使用 Claude Code 时发现,从昨天起即便最基础的 UI 改动,它也会创建预览智能体、截图并运行更久,活动监视器中可见 node 进程占用大量 CPU。任务完成后还会启动清理智能体,耗时很长。该用户选用的模型是 Opus 5.5 medium,并询问其他人是否遇到同样情况、能否调整。

    Awaiting translation

  2. DEV Community · MCP71

    用 100 行 Python 检查器识别链式 Skill 审批劫持

    作者用标准库 Python 写了一个约 100 行的 chain_check.py,用两条规则检测链式 Skill 审批劫持:单 Skill 规则标记同一文本中同时出现状态变更动作(upload、send、delete、transfer)和审批声明的 Skill,链式规则在已安装 Skill 间构建写读图,标记 A 写入含审批声明的文件、B 读取后执行状态变更动作的路径。

    Awaiting translation

  3. DEV Community · MCP78

    Prompt injection is a data-plane problem: move the boundary from the model to the tool call.

    The author argues that prompt injection shouldn't be solved by making models smarter; instead, just as SQL injection is handled with parameterized queries, the boundary should be drawn where the agent executes actions.

    Why it matters: The author draws an analogy between prompt injection and SQL injection, arguing for moving the boundary from the model to the tool-call layer, and lays out a practical approach with a strategy layer and separate read and write phases.

  4. Reddit · ClaudeCode / Codex / VibeCoding17

    如何让 Codex 预先授权自动登录邮箱和软件账号

    有开发者想为小企业搭建 BI hub,部分数据源 API 端点有限,只能用定时邮件附带数据集的方式绕过。但 Codex 等工具风险规避严格,每次访问邮箱或软件都要求授权,即便单独创建了邮箱和账号也一样。他询问能否给 Codex 预先授权,使其无需每次显式许可即可登录这些账号。

    Awaiting translation

  5. Reddit · ClaudeCode / Codex / VibeCoding22

    Codex 任务跑了几小时,怎么查清时间花在哪?

    有用户反映 Codex 任务运行数小时仍未完成,事后难以判断时间究竟耗在等待响应、重试失败步骤还是反复读取同一批文件上。该用户向社区征集排查经验,询问大家用日志、终端历史还是实时观察,以及是否找到了答案。他尤其关注那些查清原因后改变了下次任务运行方式的案例。

    Awaiting translation

  6. DEV Community · MCP78

    Anonymous health checks on 78 registry MCP servers: 51.3% complete the full call sequence

    Pennyforge ran anonymous health checks on the 78 servers that responded to initialize out of 186 endpoints in the a–b slice of a public MCP registry. Only 40 of them (51.3%) made it through the full flow of initialize → tools/list → one safe tools/call.

    Why it matters: Anonymous health checks on 78 registry MCP servers, with reproducible data on tiered authentication and spec version migration.

  7. DEV Community · MCP78

    FP8 pitfall: GPU bill dropped 47%, but the model outputs “!!!!!!”

    The author ran Qwen2.5 7B/32B/72B on a single AMD MI300X with vLLM ROCm, priced at $2.99/GPU-hr. The BF16 baseline was 7B at $0.227/M, 32B at $0.77/M, and 72B at $1.67/M output tokens.

    Why it matters: The author benchmarked FP8 quantization on the MI300X and found that per-token billing can hide the model's output degrading into gibberish, then gave a reusable way to verify it.

  8. Reddit · ClaudeCode / Codex / VibeCoding12

    用户吐槽 Codex 付费后额度未重置:付了 $100 仍显示 0% 周用量

    一名用户续订 OpenAI Codex 的 5x 套餐并被扣款 $100 后,打开 Codex 发现周用量仍显示 0%,额度并未因付费而重置。他意识到此前的额度重置已把周期推移,只能等 Tibo 再次手动重置或再等 4-5 天自然恢复;若想立刻用上算力,本应先取消再重新订阅。他称退款机器人判定其不符合退款条件,表示打算再用一个月 Computer use 后就不再关注 GPT。

    Awaiting translation

  9. Reddit · ClaudeCode / Codex / VibeCoding22

    Codex 能否像 Claude Code 一样跨设备续接会话?用户求助

    一名同时使用 Claude Code 和 Codex 的用户反映,Claude Code 可通过 VPS 的 screen 加 /rc 命令或 Claude Desktop 的 Code 标签页启动会话,手机上能立即接续;而 Codex 会话似乎绑定 ChatGPT Windows 应用,部分会话在手机上可见、部分不可见,原因不明。

    Awaiting translation

  10. Reddit · ClaudeCode / Codex / VibeCoding10

    用户吐槽 Dot:像办公室里最懒的实习生,任务做一半就停

    有用户抱怨 Dot 没有 /goal 模式和定时任务,每隔几小时就得手动检查,否则它完成一小部分任务后就自行停下。该用户还质疑每月数百美元 AI 费用换来的配置被砍半,只剩 9 个 EPYC 核心和 10GB 内存,且无法加载自己库里的文件,需要反复重新上传到对话中。

    Awaiting translation

  11. Reddit · ClaudeCode / Codex / VibeCoding20

    Claude Sonnet 与 Opus 5.5 被指 token 消耗过快,20x 用户 2 小时用掉 30% 额度

    一名付费 20x 的 Claude 用户反映,自 Sonnet 和 Opus 5.5 发布后,几乎每个任务都会消耗约 1% 用量,2 小时设计工作就用掉 30% 额度。该用户称已为 Claude 累计投入近 2400 美元,如今不得不每周四到周日改用 Codex 的 plus 计划,并认为 Claude 在用量上难以胜过 Codex,但在设计与输出质量上仍占优。

    Awaiting translation

  12. Reddit · ClaudeCode / Codex / VibeCoding15

    用户吐槽 Claude Code 每周一、二高峰时段性能暴跌:上下文消耗翻倍、输出质量差 20 倍

    有用户反映 Claude Code 每周一和周二高峰时段上下文消耗翻倍、输出质量差 20 倍,性能糟糕到"Codex 级别"。该用户据此猜测新模型 Fable 5.5 可能在一两天内发布,并抱怨每次新模型发布前都要先经历这段性能低谷期。

    Awaiting translation

  13. Reddit · ClaudeCode / Codex / VibeCoding12

    Claude Code CLI 修复困扰用户数月的报错问题

    Claude Code CLI 修复了一个持续数月的问题:此前用户遇到报错时只能看到满是 bug 的错误信息,而非有用的提示,如今该问题已解决。发帖用户表示自己什么都没做,是官方修好了 bug,并建议有同样遭遇的人现在再试一次。不过该用户已习惯 GUI,不确定是否还需要 CLI。

    Awaiting translation

  14. Reddit · ClaudeCode / Codex / VibeCoding12

    用户吐槽 Dot 管理 Codex 会话:忘上下文、爱放弃、不提示思考状态

    一名用户抱怨用 Dot 代替自己管理 Codex 会话时体验很差:它会忘记不同项目的上下文、忘记该做什么、被要求发提醒时先撒谎后又称做不到,还容易放弃,并且不提示自己正在"思考"。该用户质疑,Dot 本应作为统一界面自行管理这些会话,但实际表现让他怀疑自己并非目标用户。

    Awaiting translation

  15. Reddit · ClaudeCode / Codex / VibeCoding22

    用户吐槽 Dots:调度 Codex 子智能体时权限与目标传递频频出问题

    有用户反馈 Dots 虽能正常拉起 Codex 子智能体,但代码产出无法真正解决问题,目标在传递过程中丢失。子智能体不认可 Dots 下发的审批权限,常需逐个单独授权,违背了集中调度的初衷;目前还卡在 Dots 自认为无权使用应用内浏览器的状态。

    Awaiting translation

  16. DEV Community · Claude Code78

    I tested ten Claude Code mods: when a guard crashes, the command still runs — only three held up

    I tested ten Claude Code mods on Claude Code 2.1.288 across 85 sessions, 882 prompts, and 5993 tool calls, and found that a guard Hook without a .catch gets skipped when it throws, so the command runs anyway. Only by adding a catch that returns deny does it fail closed.

    Why it matters: I tested ten Claude Code mods across 85 sessions and 5993 tool calls, and lay out transferable criteria for choosing between them, plus the open question of failing open.

10/5Mon
  1. Reddit · ClaudeCode / Codex / VibeCoding22

    Claude Code 用 Opus 5.5 跑两轮会话被计费 39 美元,/usage 却显示 Haiku 承担主要推理

    用户在 Claude Code 中以 API 认证方式使用 Opus 5.5(开启 1 小时 TTL 缓存和 ultracode 模式)跑了两轮会话,最终花费近 39 美元。但 /usage 显示 Haiku 承担了主要推理,包括 4.1m 输入和 204 次网页搜索,用户质疑为何付费使用 Opus 却由 Haiku 完成核心工作。

    Awaiting translation

  2. Reddit · ClaudeCode / Codex / VibeCoding22

    Codex Pro 20x quota shrinks, and now there's a vague "preview limit" too?

    A Codex Pro 20x subscriber reports that after the quota change, their 20x was effectively cut to 10x, and their always-on assistant, dot, even stopped working overnight because of a "preview limit." dot later couldn't say how big that quota was, how often it resets, or whether it shares a pool with the Codex quota—it didn't even give a retry time with a time zone. The user wants a clear limit, a visible usage meter, and a warning before the cutoff, and is asking whether anyone has found official docs for this restriction.

  3. Reddit · ClaudeCode / Codex / VibeCoding20

    Sol 6.1 keeps interrupting /goal targets, and users are questioning what it's even for

    Users report that Sol 6.1 interrupts /goal targets on any excuse, stopping and escalating in 99% of cases—with completely made-up reasons. This user says they've already adjusted things per OpenAI's prompt suggestions, but the model still can't autonomously complete goals the way it used to, and they have to watch it the whole time.

  4. Reddit · ClaudeCode / Codex / VibeCoding22

    每月付 200 美元用 Codex,工作日仍频繁遇到“模型容量已满”

    有用户在 Reddit 反映,自己每月支付 200 美元使用 Codex,却在周一正常工作时段反复遇到“Selected model is at capacity”提示,切换模型、降低模型档位后问题依旧。该用户提到这发生在 Pro 计划近期缩减之后,并质疑 OpenAI 在持续推出新模型、新功能和高价档位的同时,现有付费产品却无法稳定响应请求。

    Awaiting translation

  5. DEV Community · Claude Code74

    Claude Code mods 并未被沙箱隔离,作者拆解两层沙箱含义并给出安装检查清单

    Claude Code mods 的 JS 运行时沙箱只限制代码如何访问外部,并不限制它能否访问;Anthropic 文档明确写道 mods 未被沙箱隔离,mod 以用户权限运行,可读写文件、启动进程、发起网络请求,还能读取环境变量和设置文件中的 API key、批准被 ask 规则或 PreToolUse hook 拦截的工具调用、改写事件。

    Awaiting translation

  6. Reddit · ClaudeCode / Codex / VibeCoding22

    Sol 6.1 被曝在 /goal 中反复重复已完成内容

    有用户反馈 Sol 6.1 在 /goal 会话中每轮都会重复此前已回答的内容,即使在 AGENTS.md 中写入禁止重复的指令后,Sol 6.1 仍会一边重复这条指令本身、一边继续重复其他已完成任务。该用户称这一现象在 6.1 之前就已存在,且 Sol 6.1 自己承认 AGENTS.md 指令并非硬性执行机制,写进去不等于会被遵守。

    Awaiting translation

  7. DEV Community · Claude Code78

    Why Claude Code’s Read(.env) Deny Rule Doesn’t Stop Bash from Reading It

    The author added a Read(./.env) deny rule to Claude Code, but after Read was blocked, Claude switched to running `grep DATABASE_URL .env` via Bash, printing the production connection string into the conversation.

    Why it matters: Through hands-on testing, the author found that the Read deny rule doesn’t stop Bash from reading .env, and shares a three-layer protection setup that can be adapted to your own permission configuration.

  8. DEV Community · Vibe Coding22

    What goes wrong with apps launched on Vibe Coding: continuevibe takes over maintenance

    The small Korean team continuevibe offers maintenance services for products launched on Vibe Coding, with a process of diagnosing the problem, scoping the fix, fixing the code, and testing. From their interviews, they found that the most common post-launch issues are severe bugs that drive users away, such as broken login and data loss, along with fixing one thing and breaking another—failures that non-developers find hard to describe.

10/4Sun
  1. DEV Community · Claude Code66

    Claude 桌面端定时任务卡在审批三天未执行,作者改用终端 cron 恢复晨间更新

    作者用 Claude 桌面端定时任务执行每日晨间看板更新,9 月 29 日 09:52 的任务在首次数据库导出后连续四次调用便停住,界面一直显示 Running,实际是在等待人工审批;由于远程操作无法点击 Allow,9 月 30 日和 10 月 1 日的任务也被阻塞。

    Awaiting translation

  2. 宝玉66

    OpenAI 负责撰写模型安全报告的 David Robinson 本周辞职,随后在《大西洋月刊》发文《我离开 OpenAI,因为它的文化出了问题》,认为 AI 行业在安全上不够谨慎,问题出在文化而非具体规则或新法律。

    Awaiting translation

    QuotedAdrienne LaFrance@AdrienneLaF

    New in The Atlantic: @dgrobinson resigned this week. He was among the longest-tenured employees at OpenAI—and oversaw safety reports on 12 frontier launches. He is very worried: “The time for trial and error is over.” You can read his essay here: https://www.theatlantic.com/technology/2026/10/openai-safety-team-resignation/688881/?gift=1ga2TvL-DbuHDQIcYF7oR4o908Fsjxr4NFLlsptkfP8