用 GPT-Live-1 和 Codex 把 LED 显示屏变成语音助手 Jack
开发者用 Raspberry Pi、128 × 64 HUB75 LED 面板、Codex 和 GPT-Live-1 全双工语音模型,把一块航班追踪屏改造成语音助手 Jack。
Awaiting translation
开发者用 Raspberry Pi、128 × 64 HUB75 LED 面板、Codex 和 GPT-Live-1 全双工语音模型,把一块航班追踪屏改造成语音助手 Jack。
Awaiting translation
Awaiting translation
MiMo-V2.6-Pro debuts as the top open weights model on the Artificial Analysis Intelligence Index (46). At $0.13 per Intelligence Index task, it lands on the Intelligence vs. Cost per Task Pareto frontier @Xiaomi has just released MiMo-V2.6-Pro, an open weights model with major advances in intelligence over its predecessor, MiMo-V2.5-Pro (Intelligence Index: 26). Despite the improvement, it retains the same attractive pricing at $0.435 per 1M input tokens (with a 99% cache-hit discount) and $0.87 per 1M output tokens. This makes MiMo-V2.6-Pro one of the most cost-efficient models to deploy. MiMo-V2.6-Pro is an MoE model with 1.02T total parameters and 42B active parameters. Stay tuned for additional analysis of the model. Check out MiMo-V2.6-Pro full benchmarking breakdown here: https://artificialanalysis.ai
Lovable 与 AWS、CrowdStrike、Databricks、Docker、Google Cloud、Okta、Proofpoint、Salesforce、ServiceNow、Wiz、Zscaler 共同成为 Blueprint Alliance 创始成员,该联盟推动面向企业级 AI 智能体安全与治理的开放参考架构。
Awaiting translation
Awaiting translation
Awaiting translation
Thorsten Ball 在 Register Spill 中称 Jev 不是 LLM,而是“为软件直接使用的快速结构化决策”而建的模型,TypeSafe 将其描述为“非结构化状态输入、带类型的概率决策输出”的前沿智能函数调用。
Awaiting translation
Awaiting translation
Lovable 宣布收购 Sutro,Sutro 团队过去五年构建了语言、编译器和后端平台,让应用行为与规则显式化。Sutro 创始人 Tomas Halgas 及三名工程师 Max Gfeller、Tony Zhan、Hirad Arshadi 加入 Lovable,Tomas 将负责 Lovable 的技术布道。
Awaiting translation
Mitchell Hashimoto 提出“白板答辩”标准:任何面向客户的系统,开发者都应能随时被叫住,清楚解释其工作原理并为自己的决策辩护,这是他对负责任使用 AI 的衡量标准。
Awaiting translation
Lovable 接入 Salesforce 的 Headless 360,用户可用自然语言构建应用与智能体,直接读写 Accounts、Contacts、Leads、Cases 和 Opportunities,并沿用 Salesforce 的身份与权限。
Awaiting translation
Awaiting translation
Lovable has released OJ, a preview engine written from scratch in Rust. It reads your existing vite.config.ts and runs real Vite plugins through a compatibility layer, all in a single binary, with no toolchain installed into the project.
Why it matters: Lovable rewrote its preview engine OJ in Rust, sharing cold start and memory comparisons against Vite, plus canary data from production.
Awaiting translation
Cline has released an early version of its open-source desktop app, Cline Desktop, moving the agent runtime that previously lived in the VS Code extension and CLI into a standalone workspace. It supports parallel sessions, scheduled tasks, and a Marketplace for extending tools and integrations.
Why it matters: The official release lays out the desktop app's capabilities and open entry points, so readers can judge whether it fits their multi-agent parallel workloads.
Awaiting translation
Thorsten Ball 在 Register Spill 中称 Fable 5.1 和 GPT-6 Astra 带来质变,基准分数从 65% 升到 79% 也无法体现这种变化。
Awaiting translation
GitHub 日本和韩国市场的营销负责人把活动运营流程自动化:用 Issue 表单收集活动信息,用 event-setup 标签触发 GitHub Actions,自动复制落地页、生成 UTM 链接、提交邀请邮件并同步项目看板,原本要花近一天的工作缩短到几分钟。
Awaiting translation
GitHub Copilot 应用内置 diff、终端和浏览器三个面板,让开发者无需在编辑器、终端和浏览器之间切换即可完成 AI 编码闭环。diff 面板用绿色和红色高亮显示代码的新增、删除和修改,支持接受更改、留下评论或让 Copilot 继续修改;终端面板可直接运行项目命令并支持多窗口切换;浏览器面板则能预览界面并用 Pick & Polish 工具选中元素让智能体调整。
Awaiting translation
斯德哥尔摩地面工程公司老板 Kedde 与一位懂技术的朋友用 Lovable 搭建物流平台 Kbag,三周、约 500 美元成本做出首个版本,以实时地址定价取代传统分区定价,并通过 API 接入 Volvo Connect Fleet Management 获取车辆路线、油耗、载重与司机剩余驾驶时间数据。
Awaiting translation
In an official blog post, OpenAI lays out recommendations for adjusting Skills, AGENTS.md, and task prompts under GPT-6 Astra: Skill descriptions should be as short as possible and state clearly when they apply, and multi-flow Skills should use a root document for minimal routing instead of turning the Skill into an overly specific step-by-step checklist.
Why it matters: OpenAI has published guidance on cleaning up Skills, AGENTS.md, and prompts under GPT-6 Astra, and it carries over to existing repository setups.
Cursor introduces “Projects,” a feature built for long-running work like a single feature, a migration, or an entire application. It keeps context over months and delegates tasks to thousands of sub-agents. Projects are powered by cloud agents: the coordinating agent doesn’t write code, it only plans, assigns work, and hands back results, spinning up local agents when on-device testing is needed. Each project keeps a set of files synced between the cloud and local machines, steadily accumulating research findings, artifacts, and knowledge of the codebase.
Why it matters: The official docs lay out the context-sharing and auto-triggering mechanisms for project-based multi-agent collaboration, which you can use to judge how long-running tasks get taken over.
Lovable 正式启动 Partner Program,分为面向独立开发者和小型机构的 Expert 与面向咨询公司、系统集成商的 Solution Partner 两条路径,并提供 Partner Directory 供企业查找和雇佣伙伴。
Awaiting translation
Lovable 上线 drafts 功能,为项目创建带独立对话和预览的副本,团队可并行探索同一项目的多个版本,改动只有接受并发布后才会应用到线上应用。该功能适用于任何项目,首个版本仅覆盖前端改动,涉及数据库结构或登录设置的修改仍需在项目对话中完成。
Awaiting translation
Thorsten Ball 在 Register Spill 的 Joy & Curiosity #98 中借用《反脆弱》里的扁桃体切除研究,提出工程师对 Sol、Fable、Astra 等模型输出的“代码差、注释蠢”评价,可能源于“天真干预主义”偏见——AI 已在 20 分钟内端到端完成前后端改动、内外部文档和测试,并在无头浏览器中跑完全流程、附上录屏为证。
Awaiting translation
GitHub has launched Project HydraFusion as a research preview in the Copilot CLI. It uses runtime orchestration to pick an execution plan across models from multiple providers. Users select it just like any other model, and billing follows each model's standard rates.
Why it matters: GitHub lays out three orchestration modes for HydraFusion and compares cost versus quality across three benchmarks, so you can judge the trade-offs of multi-model orchestration on real coding tasks.
The author built the space exploration game Void Explorer in Codex with Astra, featuring 2,048 star systems and over 10,000 procedurally generated planets, and shared the full workflow from prompts to architecture, testing, and performance measurement.
Why it matters: Using Astra in Codex, the author built an entire game and showed a transferable collaborative workflow that spans prompts, testing, and performance measurement.
用户向 Codex 中的 Astra 描述一栋极简住宅的需求,Astra 通过 Blender Python API(bpy)生成可编辑 3D 场景,完成建筑、家具、材质、灯光与相机,并自行检查预览渲染、修正细节。项目从带家具的起居亭扩展为围绕庭院布局的 U 型单层住宅,含三间卧室、办公室、浴室和更衣室,随后导出到 Unreal Engine 5 探索实时漫游。
Awaiting translation
GitHub Copilot 应用支持同时运行多个 AI 智能体会话,每个会话运行在独立的 Git worktree 上,互不干扰,可并行推进同一项目的不同任务。每个会话保留各自的上下文,切换时无需重新说明,用户可在会话视图中查看进度并审阅结果。
Awaiting translation
Cline 把 VS Code 扩展从约 76,000 行单体核心迁移到 Cline SDK,并自建灰度发布机制:一个安装包内打包 loader、legacy 和 next 两套扩展,由 PostHog 功能开关按百分比决定激活哪套,崩溃时自动回退到 legacy,开关可随时降到 0% 作为 kill switch。
Awaiting translation
Why it matters: Cline 官方复盘如何把 1100 万用户的 VS Code 扩展迁到新 harness,含灰度机制与前后指标对比。
Mitchell Hashimoto 公布 Superlogical 服务器内存测试结果:在繁忙终端和客户端规模下,其内存占用显著低于 tmux 及其他复用器,而 tmux 空载状态表现优秀。
Awaiting translation
GitHub Copilot 团队复盘了四项降低 AI 编码成本的改动,核心原则是按完整任务而非单次工具调用衡量效率。
Awaiting translation
Why it matters: GitHub Copilot 团队复盘四项降本改动,并给出可迁移的评估方法:按完整任务而非单次工具调用衡量成本。
Cursor supports self-hosted machines: code repositories, build artifacts, and secrets all stay on internal machines within your own infrastructure, and the agent handles tool calls locally. My Machines connects a single laptop or VM to a personal workflow, while Team Pools are named worker queues for teams or enterprises—scaling capacity up with requests and down when workers disconnect. Pools aren't tied to code repositories, and idle machines can sleep and then resume within a reconnection window.
Why it matters: The official docs lay out pooled scheduling and sandbox integration for self-hosted machines, so readers can judge whether tool execution can stay within their own network.
Lovable 现已接入 Fable 5.1,早期测试显示其在修复和改进现有应用上比 Fable 5 最高提升 17%,单任务成本最多降低 31%。该模型在中等和高推理强度下的 UI 与视觉设计质量最高提升 3.5%,并会在完成任务前打开浏览器运行应用进行自我验证。Lovable 正将卡住的会话以及更长、更复杂的任务路由到 Fable 5.1。
Awaiting translation
Thorsten Ball 在 Joy & Curiosity #97 中提出,AI 智能体让 bug 的发现和修复更快、可并行、可异步,开发者应重新校准代码库能容忍的 bug 数量。
Awaiting translation
Terminal-Bench 发布 4.0 版本,校准任务的时间、CPU 和内存资源,修复 19 个任务并移除 8 个饱和或存在质量问题的任务,所有任务统一设为 8 小时 agent 超时。
Awaiting translation
Cursor 云端智能体不再需要连接 GitHub 或其他第三方 SCM 提供商,用户可直接输入提示开始工作,Cursor 会在后台创建 Origin 代码仓库。满意后可点击“创建代码仓库”将工作保存到 Origin,并设置私有或内部可见性。Cursor 还能通过端口转发在浏览器中实时预览云端智能体环境,连接 Vercel 账户后点击“发布”即可生成可访问 URL。
Awaiting translation
斯坦福大学研究人员主导、Terminal-Bench 团队联合全球科研机构专家打造的 Terminal-Bench-Science 0.1 发布,首批含生命、物理、地球、数学和工程科学领域的 70 项任务。
Awaiting translation
Cline 让八个模型在自家 harness 里做 IMO 2026 六道题,证明由 GPT-5.5 和 Claude Opus 5 双盲按 0–7 分制评分、Gemini 3.1 Pro 仲裁,金牌线为 29 分。
Awaiting translation
Why it matters: Cline 用同一套 harness 盲评八个模型做 IMO 2026,给出分数与单次成本对照,可看开源权重模型的实际性价比。
GitHub Copilot 应用支持用自动化分诊 Dependabot pull request:用自然语言描述任务,按风险分组、识别安全的补丁与次版本更新、核验 CI 状态并给出摘要。自动化可选手动、每小时、每天、每周或 issue 创建时触发,也可选择在云端或本地运行,每次运行记录都会保存。
Awaiting translation
The Cline team built a code review agent with the Cline SDK, splitting review into two agent loops—review and judge—then using a driver script to batch-submit the surviving issues as a single COMMENT event to the GitHub PR.
Why it matters: A full breakdown of the plugin, Hooks, and two-stage loop behind a code review agent, transferable to other automated review scenarios.