OpenAI 在欧盟为 ChatGPT 和 Codex 文本加入 textGrain 隐藏水印
OpenAI 表示欧盟符合条件的 ChatGPT 和 Codex 文本将携带名为 textGrain 的隐藏水印,未来几周内面向所有套餐推出,目前仅限欧盟。该水印是词选择上的统计模式而非可见标签,OpenAI 称其不识别用户、账号或提示词,检测器不公开,仅获批研究人员可申请使用。
Awaiting translation
OpenAI 表示欧盟符合条件的 ChatGPT 和 Codex 文本将携带名为 textGrain 的隐藏水印,未来几周内面向所有套餐推出,目前仅限欧盟。该水印是词选择上的统计模式而非可见标签,OpenAI 称其不识别用户、账号或提示词,检测器不公开,仅获批研究人员可申请使用。
Awaiting translation
OpenAI 在 ChatGPT 和 Codex 中为生成文本加入不可见水印,以符合欧盟 AI 法案对机器可读标识的要求,目前仅面向欧盟用户,未来几周自动生效。
Awaiting translation
GitHub 和 Microsoft 于 2026 年 10 月 5 日开放研究预览版 ReviewBench,用于评测 AI 代码评审智能体。该基准包含来自 187 个公开仓库、19 种语言的 219 个 pull request,语言和仓库规模分布基于对 GitHub 上 1.039 亿个 pull request 的分析,并刻意提高了实质性改动的占比。
Awaiting translation
SemiAnalysis 实测 Anthropic、OpenAI 等九家 AI 订阅套餐后得出,同样 200 美元,Claude 订阅折算的 Token 用量约为 OpenAI 的 5 倍。
Awaiting translation
Anthropic Subscriptions Offer 5x+ More Value Than OpenAI Limit testing every AI subscription plan from Anthropic, OpenAI, Meta, SpaceXAI, MiniMax, Moonshot, Zdotai, Cursor, and Cognition https://newsletter.semianalysis.com/p/anthropic-subscriptions-offer-5x
OpenAI 在 9 月 22 日至 10 月 1 日间连续发布 Codex CLI 0.156.0 到 0.160.0,官方称之为一次大更新,带来全屏终端界面、/agents 视图和并行任务管理。
Awaiting translation
OpenAI 在 28 天更新的第 1 天宣布,通过 ChatGPT 订阅使用 GPT-6 Astra 和 GPT-6.1 Sol 时默认速度提升约 50%,用户无需改设置,两小时内生效。
Awaiting translation
Day 1/ We have optimized the default speed to be ~50% faster across GPT-6 Astra and GPT-6.1 Sol through the subscription across all our products and partners using Sign in With ChatGPT (including OpenCode, Pi, Amp, Devin, ...). No changes needed on your end and this should be felt within the next two hours.
OpenAI 宣布未来几周内对欧盟用户在 ChatGPT 和 Codex 生成的文字加入不可见水印,以满足欧盟《人工智能法案》要求,API 用户可在全球范围自行开启,默认关闭。
Awaiting translation
We're expanding our approach to content provenance to include text in response to EU regulatory requirements, while recognizing the significant limitations of current text watermarking technology. Our tools already help verify whether an image or audio file was created with our models. This work builds on those efforts to help people better understand when content may have been generated or edited with an OpenAI model. In the EU, we’ll start watermarking eligible text from ChatGPT and Codex over the coming weeks to comply with the EU AI Act. Customers using our API can turn on text watermarking for select models worldwide today.
作者用 npm 发布时间戳统计 2026 年 Q3 的发布次数:Claude Code 发布 79 个版本、Codex CLI 38 个、Gemini CLI 15 个,中位发布间隔分别为 0.96 天、1.64 天和 6.09 天。
Awaiting translation
GPT-5.5 于 2026 年 4 月 23 日发布,是 OpenAI 自 GPT-4.5 以来首个完全重训练的基座模型,在 Terminal-Bench 2.0 上得分 82.7%,相同 Codex 任务下比 GPT-5.4 少用 40% token,但输入输出价格翻倍至每百万 token 5 美元和 30 美元。
Awaiting translation
9 月 15 日发布的 System One 决策模型 Jev 一周内 GitHub 相关项目超两千个、star 破四万,一批论文随之涌现。
Awaiting translation
作者上手体验了 OpenAI DevDay 发布的 Dots、Decisions API、开源 Codex Harness 更新和 ChatGPT Space 四项能力。
Awaiting translation
Awaiting translation
OpenAI 在 DevDay 2026 上向每位参会者赠送了一台透明外壳的 ModRetro Chromatic: DevDay Edition 限定掌机,该机型由 ModRetro 与 OpenAI 联合开发,限量 8000 台。
Awaiting translation
OpenAI 在 9 月 29 日 DevDay 上发布 GPT-6.1 Sol,距离 9 月 22 日发布 GPT-6 Sol 仅一周,官方测试中接近 Astra 水准,API 标准输入输出单价为 Astra 的 1/5。
Awaiting translation
OpenAI has released GPT-6.1 Sol for coding, document processing, and task automation. The company says it comes close to GPT-6 Astra on some tests. On DeepSWE 1.1, a benchmark of real-world codebase tasks, the model matches Astra while costing about one-fifth as much to run. On OSWorld 2.0, which tests app control, it beats GPT-6 Sol by 7 percentage points at the highest reasoning tier.
Why it matters: GPT-6.1 Sol matches Astra on DeepSWE 1.1 at roughly one-fifth the cost, which gives you a sense of how the price-performance tradeoff for coding tasks has shifted.
Simon Willison live-blogged the OpenAI DevDay 2026 keynote from Fort Mason in San Francisco, where OpenAI announced the personal agent Dots, ChatGPT Space, GPT-6.1 Sol, Ultrafast, and more.
Why it matters: A running, item-by-item record of what OpenAI announced at DevDay, for a quick look at what Dots, GPT-6.1 Sol, Ultrafast, and Codex Security actually look like.
OpenAI 在 9 月 29 日旧金山 DevDay 上发布 GPT-6.1 Sol,API 名为 gpt-6.1-sol,定价为每百万输入 token 2 美元、输出 10 美元,缓存输入 0.10 美元,标准价格是 GPT-6 Astra 的五分之一。
Awaiting translation
Why it matters: OpenAI DevDay 发布 GPT-6.1 Sol,价格降至 Astra 的五分之一,并同步更新 Codex、Agents API 与插件体系,可据此判断成本与工具链变化。
OpenAI 发布 GPT-6 Sol 和 Luna,Sol 定价 $2/$10 MTok、约为 GPT-5.6 的一半,DeepSWE 68.8%,Luna 定价 $0.10/$0.50 MTok。
Awaiting translation
OpenAI 将重新向新订阅者开放 200 美元的 Pro 订阅,但调整了额度计算方式。据 Тибо Соттио 说法,按 API 消耗折算,新方案提供的用量约为旧版 Pro 200 的一半。
Awaiting translation
Tibo 发帖称 Codex 不会恢复 5 小时用量限制,用户仍可自由安排每周额度。Pro 200 美元订阅明天重新开放新用户订阅,但按 API 等价价值计算,新套餐大约只有旧版的一半 API 用量价值;明天还会给 Pro 订阅加入一些不消耗 usage 的新功能,具体未公布。
Awaiting translation
OpenAI 发布 Codex CLI 0.157.0,主要变化是工具会为支持的交互式会话自动启动后台服务器,让 Codex 的使用不再绑定在单个终端窗口。新增的 f 键可以从 CLI 分叉其他应用中打开的对话,保留草稿和排队中的提示词,/import 命令也能在远程和本地后台会话中使用。原文指出这并不代表关闭终端后任务仍会继续执行,OpenAI 没有这样的声明,后台服务器只是为共享会话打基础。
Awaiting translation
未发布的 ChatGPT 订阅配置中出现新档位 Pro Max,据 TestingCatalog 称月费 500 美元,是现价 200 美元 Pro 的 2.5 倍。
Awaiting translation
OpenAI 于 9 月 22 日在 API 发布 gpt-6-sol 和 gpt-6-luna,两者接受文本和图像输入、只生成文本。标准处理下,GPT-6 Sol 每百万输入 token 收费 2 美元、缓存 token 0.20 美元、输出 10 美元;GPT-6 Luna 分别为 0.10、0.01 和 0.50 美元。
Awaiting translation
开发者 pdfu 在 iOS 27 和 macOS Golden Gate 的私有框架中发现,Apple 的新 Siri 架构设计了与第三方 AI 模型对接的机制。
Awaiting translation
On September 3, 2026, OpenAI released GPT-6 Astra and Astra Pro, initially limited to enterprises in the Daybreak cybersecurity program, with paid ChatGPT, the API, and AWS opening up over the following days.
Why it matters: We break down the benchmark comparison between GPT-6 Astra and Fable 5.1, pointing out that the tested versions and harnesses differ across teams, so readers can judge which scores are actually comparable.
斯坦福大学研究人员主导、Terminal-Bench 团队联合全球科研机构专家打造的 Terminal-Bench-Science 0.1 发布,首批含生命、物理、地球、数学和工程科学领域的 70 项任务。
Awaiting translation
OpenAI has launched Daybreak, combining ChatGPT, Codex Security, and the open-source Codex Security CLI into a security defense workflow that covers pre-merge PR reviews, repository and vulnerability backlog scans, and regular CI checks.
Why it matters: The official documentation walks through the full Codex Security workflow—from PR reviews and repository scans to CLI-based batch scanning—so you can decide how to plug it into your existing security processes.
OpenAI has open-sourced the harness that drives the Codex app, CLI, and IDE extensions, and through the Codex app-server client protocol it exposes capabilities like creating threads, starting turns, receiving events, and handling approval requests.
Why it matters: With the Codex harness and app-server protocol now public, developers can see how to embed the agent in their own products and where the boundaries are.
Anthropic 宣布从 8 月 14 日起,Pro、Max 和 Team 套餐的新会话默认运行 auto mode,并停止对分类器额外 token 开销收费;Enterprise、Claude API、AWS、Bedrock、Google Cloud 和 Microsoft Foundry 暂时保持可选,计划下个月改为默认。
Awaiting translation
Why it matters: Anthropic 公布 auto mode 的安全评测数据与内部拦截案例,可据此判断默认权限模式对现有工作流的影响。
Vercel ships v0 API, giving programmatic, headless access to the v0 app-generation agent: send a prompt, v0 generates an app, spins up a dev server in the Vercel Sandbox, and returns a preview URL you can embed in your own UI. The API is now generally available.
Why it matters: v0 opens up its app-generation capability as an API, so readers can judge how to wire it into their own product or agent workflow.
The Terminal-Bench team releases Terminal-Bench 3.0, whose first version spans 7 domains and 74 tasks, with the strongest model passing about 34%. Building on Terminal-Bench 2.1, this release broadens task diversity and adds CI/CD, semantic versioning, and result migration to keep improving the benchmark.
Why it matters: Terminal-Bench 3.0 rebuilds the benchmark with 74 tasks and CI/CD-based versioning, so readers can see how the new benchmark separates models.
OpenAI has added custom repository rules to Codex Code Review: you can put review guidelines in AGENTS.md, and Codex applies them during review and cites where each one came from in its findings. In OpenAI's own evaluation, the rule-guided version caught 98% of the required custom issues, versus 58.3% for the baseline. The guidance is to start with non-obvious invariants like compatibility requirements and data boundaries, put repo-level rules in the root directory and service-level rules in the corresponding directory, and leave formatting and mechanical checks to CI.
Why it matters: OpenAI lays out the capabilities, the syntax, and the evaluation data for Codex Code Review custom rules, so you can judge how to bake your team's review experience into AGENTS.md.
Lovable 上线 agent integrations,任何公开已发布的 Lovable 应用都能添加 MCP server,从而被 ChatGPT、Claude 等 AI 工具直接调用。
Awaiting translation
Why it matters: Lovable 把已发布应用接入 MCP,读者可据此判断自家应用如何被 ChatGPT、Claude 直接调用。
Codex 团队在 GPT-5.6 发布后于 Reddit 举办 AMA,说明 Sol 是主力模型、Terra 更快更省、Luna 主要用于廉价子智能体和上下文收集,UI 场景推荐用 Sol 配合参考图。
Awaiting translation
Lovable 在早期访问中测试 GPT-5.5,其内部基准显示最难任务通过率从 GPT-5.4 的 36.9% 升至 41.6%,每次请求工具调用减少 23.1%,用户卡住的消息占比下降 9.9%。GPT-5.5 每请求输出 token 减少 33%,日常任务成本效率提升约 15%,将很快向 Lovable 构建者开放。
Awaiting translation
OpenAI 宣布 Codex 可通过 Figma MCP Server 生成 Figma 设计文件,并支持设计稿与代码双向流转。
Awaiting translation
Why it matters: 官方给出 Codex 与 Figma MCP 双向打通的具体操作步骤,读者可据此判断设计到代码的往返流程能否落地。
AI 代码审查平台 Graphite 完成 5200 万美元 B 轮融资,由 Accel 领投,Anthropic 的 Anthology Fund、Shopify Ventures、Figma Ventures、Andreessen Horowitz 等参投。
Awaiting translation