Skip to content

#Product updates

0 items today
10/6Tue
  1. 宝玉67

    OpenAI 在 28 天更新的第 1 天宣布,通过 ChatGPT 订阅使用 GPT-6 Astra 和 GPT-6.1 Sol 时默认速度提升约 50%,用户无需改设置,两小时内生效。

    Awaiting translation

    QuotedTibo@thsottiaux

    Day 1/ We have optimized the default speed to be ~50% faster across GPT-6 Astra and GPT-6.1 Sol through the subscription across all our products and partners using Sign in With ChatGPT (including OpenCode, Pi, Amp, Devin, ...). No changes needed on your end and this should be felt within the next two hours.

10/5Mon
  1. AI Hero · Skills Updates62

    AI Hero Skills v1.3 released: new /implement-spec, /pr, and /retro skills, and CONTEXT.md renamed to GLOSSARY.md

    AI Hero Skills v1.3 is out, adding three skills—/implement-spec, /pr, and /retro—and extending the main flow from /grill-with-docs → /to-spec → /to-tickets into implementation, PRs, and retros.

    Why it matters: The author rounds out the skill set into a full path from writing specs to PRs and retros, and lays out the trade-offs at each step along with the rough edges they already know about.

10/3Sat
10/1Thu
9/28Mon
9/26Sat
  1. Boris Cherny60

    Claude Tag 在 Slack 中现已支持个人连接器,可直接访问个人有权限的 Drive 文档、Salesforce 账号或数仓表,今天在 Teams 上线、下周面向 Enterprise 开放。

    Awaiting translation

    QuotedNoah Zweben@noahzweben

    Claude Tag in Slack can now use your personal connectors! You can now securely access that Drive doc, Salesforce account, or Warehouse table that you have personal access to right where the work happens. Avail. on Teams today and Enterprise next week https://claude.com/blog/claude-tag-now-supports-personal-connectors-in-channels

9/24Thu
  1. Lovable · Blog71

    Lovable ships Chats, and opens up the trajectory and inbox architecture behind its multi-agent collaboration

    Lovable launches Chats, an agent that runs at the workspace level, can hold conversations across projects and trigger builds; once changes are confirmed, it hands the task off to the project's builder agent and brings progress back into the conversation.

    Why it matters: Lovable shares the three-layer architecture behind Chats—trajectory, inbox, and activation—which you can adapt for your own multi-agent orchestration.

  2. Lovable · Blog36

    Lovable 聊天功能免费开放,可对话探索应用创意与改进

    Lovable 聊天功能现已免费开放,Free、Pro 和 Business 工作区每天都有免费聊天额度。用户可以让它评估该做哪个创意、把客户反馈整理成计划,或分析已上线应用的代码和 Lovable Cloud 数据库用量。聊天在生成图片、视频或转交 Plan、Build 时仍照常消耗 credits,当前聊天定价(含每日免费额度)适用至 2026 年 10 月 31 日。

    Awaiting translation

9/23Wed
  1. Lovable · Blog60

    Lovable Ships Opus 5.5: Faster Builds, Quality on Par with Opus 5

    Lovable has shipped Opus 5.5, which the company says matches Opus 5 in results while cutting the number of steps by one-third to one-half. On Lovable's internal benchmarks, Opus 5.5 ties Opus 5 on 0-to-1 builds and iterative code changes, and comes out 4% to 6% ahead on validation discipline; across all reasoning effort levels, steps per task drop by 26% to 57% and input tokens fall by 21% to 59%, with the differences significant at the 95% confidence level.

    Why it matters: Lovable shares official comparison data between Opus 5.5 and Opus 5 on step counts and tokens, so readers can judge the real change in build efficiency.

9/22Tue
9/15Tue
  1. Lovable · Blog64

    Lovable open-sources OJ, a Rust preview engine that beats Vite on cold start and memory

    Lovable has released OJ, a preview engine written from scratch in Rust. It reads your existing vite.config.ts and runs real Vite plugins through a compatibility layer, all in a single binary, with no toolchain installed into the project.

    Why it matters: Lovable rewrote its preview engine OJ in Rust, sharing cold start and memory comparisons against Vite, plus canary data from production.

  2. Cline · Blog62

    Cline releases the open-source desktop app Cline Desktop, aimed at open-weight models.

    Cline has released an early version of its open-source desktop app, Cline Desktop, moving the agent runtime that previously lived in the VS Code extension and CLI into a standalone workspace. It supports parallel sessions, scheduled tasks, and a Marketplace for extending tools and integrations.

    Why it matters: The official release lays out the desktop app's capabilities and open entry points, so readers can judge whether it fits their multi-agent parallel workloads.

9/10Thu
  1. Cursor · Changelog76

    Cursor launches “Projects,” a feature that uses a coordinating agent to take on long-running development work

    Cursor introduces “Projects,” a feature built for long-running work like a single feature, a migration, or an entire application. It keeps context over months and delegates tasks to thousands of sub-agents. Projects are powered by cloud agents: the coordinating agent doesn’t write code, it only plans, assigns work, and hands back results, spinning up local agents when on-device testing is needed. Each project keeps a set of files synced between the cloud and local machines, steadily accumulating research findings, artifacts, and knowledge of the codebase.

    Why it matters: The official docs lay out the context-sharing and auto-triggering mechanisms for project-based multi-agent collaboration, which you can use to judge how long-running tasks get taken over.

9/9Wed
9/7Mon
9/5Sat
  1. GitHub Blog · Copilot71

    GitHub Copilot launches Project HydraFusion, using multi-model runtime orchestration to improve coding quality

    GitHub has launched Project HydraFusion as a research preview in the Copilot CLI. It uses runtime orchestration to pick an execution plan across models from multiple providers. Users select it just like any other model, and billing follows each model's standard rates.

    Why it matters: GitHub lays out three orchestration modes for HydraFusion and compares cost versus quality across three benchmarks, so you can judge the trade-offs of multi-model orchestration on real coding tasks.

9/2Wed
  1. Cursor · Changelog66

    Cursor launches self-hosted machines, keeping tool execution within your own network

    Cursor supports self-hosted machines: code repositories, build artifacts, and secrets all stay on internal machines within your own infrastructure, and the agent handles tool calls locally. My Machines connects a single laptop or VM to a personal workflow, while Team Pools are named worker queues for teams or enterprises—scaling capacity up with requests and down when workers disconnect. Pools aren't tied to code repositories, and idle machines can sleep and then resume within a reconnection window.

    Why it matters: The official docs lay out pooled scheduling and sandbox integration for self-hosted machines, so readers can judge whether tool execution can stay within their own network.

9/1Tue
  1. Lovable · Blog38

    Lovable 接入 Fable 5.1:迭代修复最高提升 17%,成本降低 31%

    Lovable 现已接入 Fable 5.1,早期测试显示其在修复和改进现有应用上比 Fable 5 最高提升 17%,单任务成本最多降低 31%。该模型在中等和高推理强度下的 UI 与视觉设计质量最高提升 3.5%,并会在完成任务前打开浏览器运行应用进行自我验证。Lovable 正将卡住的会话以及更长、更复杂的任务路由到 Fable 5.1。

    Awaiting translation

8/28Fri
8/27Thu
  1. Cursor · Changelog42

    Cursor 云端智能体无需代码仓库即可从零开始,支持保存到 Origin

    Cursor 云端智能体不再需要连接 GitHub 或其他第三方 SCM 提供商,用户可直接输入提示开始工作,Cursor 会在后台创建 Origin 代码仓库。满意后可点击“创建代码仓库”将工作保存到 Origin,并设置私有或内部可见性。Cursor 还能通过端口转发在浏览器中实时预览云端智能体环境,连接 Vercel 账户后点击“发布”即可生成可访问 URL。

    Awaiting translation

8/25Tue
8/21Fri
  1. OpenAI Developer Blog · Codex65

    OpenAI Releases Daybreak and Codex Security, a Security Workflow

    OpenAI has launched Daybreak, combining ChatGPT, Codex Security, and the open-source Codex Security CLI into a security defense workflow that covers pre-merge PR reviews, repository and vulnerability backlog scans, and regular CI checks.

    Why it matters: The official documentation walks through the full Codex Security workflow—from PR reviews and repository scans to CLI-based batch scanning—so you can decide how to plug it into your existing security processes.

8/19Wed
  1. Cursor · Changelog66

    Cursor Updates Cloud Agents and Harness: Event Subscriptions, Independent Sub-Agent Runs, and /goal

    Cursor has updated its cloud agents and Cursor harness so cloud agents can subscribe to event sources, resume when there's new activity in a PR, Slack thread, or scheduled task, and keep going until the work is done—fixing CI failures and handling bot comments.

    Why it matters: Cloud agents are moving from one-shot runs to subscribing to events and following up continuously on PRs and Slack threads, which gives readers a way to judge how the boundaries of automation are shifting.

  2. OpenAI Developer Blog · Codex71

    OpenAI open-sources the Codex harness and the app-server client protocol

    OpenAI has open-sourced the harness that drives the Codex app, CLI, and IDE extensions, and through the Codex app-server client protocol it exposes capabilities like creating threads, starting turns, receiving events, and handling approval requests.

    Why it matters: With the Codex harness and app-server protocol now public, developers can see how to embed the agent in their own products and where the boundaries are.

8/13Thu
  1. Augment Code · Blog62

    Augment Code Expands Cosmos: Turning Code Review into an Agentic PR-to-Merge Loop

    Augment Code has extended its Cosmos review system from code review to a full PR-to-merge loop, adding four capabilities: Verifier, PR Fixer, Review Dashboard, and cosmos approve. Dedicated Experts handle risk analysis, line-by-line correctness review, design review, runtime verification, and fixes.

    Why it matters: Augment has expanded code review into a PR-to-merge loop covering fixes, verification, and approval, giving readers a way to judge how multi-agent division of labor plays out in practice.

8/12Wed
8/5Wed
  1. Vercel · v0 Blog62

    Vercel ships v0 API for programmatic access to its app-generation agent

    Vercel ships v0 API, giving programmatic, headless access to the v0 app-generation agent: send a prompt, v0 generates an app, spins up a dev server in the Vercel Sandbox, and returns a preview URL you can embed in your own UI. The API is now generally available.

    Why it matters: v0 opens up its app-generation capability as an API, so readers can judge how to wire it into their own product or agent workflow.

7/30Thu
  1. Terminal-Bench · News62

    Terminal-Bench ships new Harbor features, turning the benchmark into a versioned asset that keeps getting updated

    The Terminal-Bench team has shipped a batch of new Harbor features that let datasets be released by version and let leaderboards migrate to new versions by reusing, re-evaluating, or rerunning trials. Tasks use semantic versioning: patch-level changes reuse old results as-is, validator changes only require re-evaluating saved artifacts, and only major changes that significantly alter the agent environment require a rerun. Dataset versions follow the highest version number among the tasks, and leaderboards use diffs to rerun only the tasks with major changes.

    Why it matters: The Terminal-Bench team maintains the benchmark like software, laying out concrete mechanisms for task semantic versioning and leaderboard upgrades that you can carry over to your own evaluation pipeline.

7/29Wed
7/22Wed
  1. Lovable · Blog28

    Lovable 成为首个获得 AIUC-1 认证的 AI 编程智能体平台

    Lovable 成为首个获得 AIUC-1 认证的 AI 编程智能体平台,该标准是业界首个专为 AI 智能体打造的安全、保障与可靠性标准。AIUC-1 由 Stanford、MIT、MITRE 和云安全联盟参与制定,包含六大原则下的 51 项要求,覆盖密钥管理、安全代码生成默认设置、沙箱执行、人工监督和企业治理,每项要求均需提供证据并由第三方独立验证,而非自我声明。

    Awaiting translation

7/20Mon
  1. OpenAI Developer Blog · Codex71

    Codex Code Review now supports custom review rules in AGENTS.md

    OpenAI has added custom repository rules to Codex Code Review: you can put review guidelines in AGENTS.md, and Codex applies them during review and cites where each one came from in its findings. In OpenAI's own evaluation, the rule-guided version caught 98% of the required custom issues, versus 58.3% for the baseline. The guidance is to start with non-obvious invariants like compatibility requirements and data boundaries, put repo-level rules in the root directory and service-level rules in the corresponding directory, and leave formatting and mechanical checks to CI.

    Why it matters: OpenAI lays out the capabilities, the syntax, and the evaluation data for Codex Code Review custom rules, so you can judge how to bake your team's review experience into AGENTS.md.

7/15Wed
6/24Wed
  1. Andrej Karpathy60

    Anthropic 发布 Claude Tag,团队可在 Slack 中把 Claude 加为团队成员,让它访问指定频道和工具,通过 @ 它来委派任务。Karpathy 认为这是 LLM 交互界面的第三次重大改版:第一代是访问网站,第二代是下载到电脑的应用,第三代则是自带工具和组织级上下文、与人类团队并行工作的持久异步实体。

    Awaiting translation

    QuotedClaude@claudeai

    Introducing Claude Tag, a new way for teams to work with Claude. In Slack, Claude joins as a team member with access to the channels and tools you choose. Tag Claude in and delegate tasks to it while you focus on other work.

6/18Thu
  1. AI Hero · Skills Updates69

    AI Hero Skills v1 发布:token 消耗降低 63%,新增 /ask-matt 与 /writing-great-skills

    AI Hero 的 skills 目录发布 v1,通过在各 Skill 上启用 disable-model-invocation: true,让 Skill 描述不再进入模型选择 Skill 时查看的上下文窗口,Skill 描述的 token 成本降低 63%。

    Awaiting translation

    Why it matters: v1 用 disable-model-invocation 把 Skill 描述移出上下文窗口,并区分用户调用与模型调用,读者可据此判断自己的 Skill 组织方式。

  2. Augment Code · Blog71

    Augment launches Project Builder on the Cosmos platform, taking large projects from design doc to merge

    Augment launches Project Builder on the Cosmos platform. This Cosmos expert turns a one-line feature description into a design doc grounded in the real codebase; after human review, it orchestrates worker agents to implement the work and drive it to merge.

    Why it matters: Augment has shared how Project Builder handles design review and orchestration, plus the code volume and launch timelines of three production projects—enough to judge whether design-first plus agent orchestration is workable.

5/19Tue
5/18Mon
  1. Lovable · Blog62

    Lovable 上线 Skills,把重复指令变成可复用技能

    Lovable 上线 Skills 功能,把重复交代的工作方式写成可复用的 markdown 技能文件,在相关任务出现时按需加载。技能以文件夹形式组织,主文件 SKILL.md 含 name、description 和 instructions,description 是决定是否触发的唯一依据,支持文件只在主文件引用且确实需要时才加载。

    Awaiting translation

    Why it matters: 官方详解 Lovable Skills 的文件结构、触发机制与写法,并给出可对照的正反示例。