Show HN:Craigslist 式 AI 智能体 Skill 市场,由人工策展
一个名为 Craigslist for agent skills 的 Skill 市场在 Hacker News 的 Show HN 板块发布,所有 Skill 由人工策展。
Awaiting translation
一个名为 Craigslist for agent skills 的 Skill 市场在 Hacker News 的 Show HN 板块发布,所有 Skill 由人工策展。
Awaiting translation
GitHub 日本和韩国市场的营销负责人把活动运营流程自动化:用 Issue 表单收集活动信息,用 event-setup 标签触发 GitHub Actions,自动复制落地页、生成 UTM 链接、提交邀请邮件并同步项目看板,原本要花近一天的工作缩短到几分钟。
Awaiting translation
有开发者反馈 Claude Code 项目开发数月后开始频繁触发压缩会话(Compacting conversation),压缩完记忆就错乱,询问如何迁移必要记忆。
Awaiting translation
In an official blog post, OpenAI lays out recommendations for adjusting Skills, AGENTS.md, and task prompts under GPT-6 Astra: Skill descriptions should be as short as possible and state clearly when they apply, and multi-flow Skills should use a root document for minimal routing instead of turning the Skill into an overly specific step-by-step checklist.
Why it matters: OpenAI has published guidance on cleaning up Skills, AGENTS.md, and prompts under GPT-6 Astra, and it carries over to existing repository setups.
文章从 Claude Code、Codex 和 Gemini CLI 的官方文档中整理出四种提示词模式:先计划再编辑、给智能体一个可运行的检查、让智能体反过来访谈你、把反复重打的提示词存成文件,并给出各家对应的命令、参数和文件格式。
Awaiting translation
Why it matters: 横向对照 Claude Code、Codex、Gemini CLI 三家文档,给出计划模式、可运行检查、访谈式提问和保存提示词四种模式的命令与文件格式。
Simon Willison 发布 .blend URL Viewer,粘贴 CORS 可访问的 .blend 文件或 GitHub 仓库链接即可在浏览器中查看 Blender 5.x 文件,支持网格几何体、曲线、文本、材质、光照和已保存的相机位置,并提供轨道控制、线框模式和模型适配。
Awaiting translation
i-have-adhd 是一个给编码助手用的 Skill,目标是让回答先给行动、步骤编号、结尾只留一个具体下一步,不再写“希望有帮助”之类的客套话。它由 10 条规则组成,全文在 SKILL.md 中,并给出前后对比:原本绕一大圈讲 auth 流程的回答,变成先执行 npm install jsonwebtoken@latest、再改 src/auth.ts:42 的具体步骤。
Awaiting translation
Cursor engineer Lauren Tan shares how she gets AI agents to submit and merge PRs on their own: the key is validation—letting the agent run code, capture CPU traces, and open an iOS simulator to check its own work.
Why it matters: Cursor engineers break trust in AI agents down into reusable validation skills, feature maps, and evals, so readers can build their own automated validation workflows.
Hacker News 上有人提问该去哪里学习 AI 编程相关的 Skill,回答建议关注 GitHub trending 仓库,那里常有快速升温的 Skill 项目。
Awaiting translation
Author Ryan Lopopolo argues that an agent is a parameterized program built on top of a set of capabilities: models and configurations, reasoning and tool-call loops, computers, disks, context, Skills, tools, connectors, runtimes, network policies, identity, IAM, guardrails, I/O channels, and system prompts.
Why it matters: Drawing on his experience building multiple agents, the author proposes a platform architecture that decouples capability interfaces from their implementations — a useful reference for teams building Agent platforms.
一位 Claude Code 工程师写的《找到你的未知》长文四天获得三百多万浏览,核心是把提示词看作地图、代码库看作地盘,两者之间的差距就是未知,模型越强,卡住它的越可能是使用者自己讲不清未知。
Awaiting translation
作者用自己在用的 Remotion 视频工程演示 CLAUDE.md 的分层写法:根目录 CLAUDE.md 只放路由,41 行,指向 L0 官方、共享、出片三层资源;产出目录的 CLAUDE.md 同样只做路由,61 行。
Awaiting translation
Anthropic 工程师在《Claude 5 时代模型的上下文工程新规则》中称,针对 Opus 5、Fable 5 这一代模型把 Claude Code 系统提示词删掉 80% 以上,跑官方评测没有可测出的损失。
Awaiting translation
宝玉在演讲中提出 AI 原生思维,主张做 AI 产品要盯着模型能力边界线找需求,并按能力、成本、价值三条边界判断值不值得做。他以自己做的字幕翻译 App BaoCut 为例,说明从模拟字幕组的 V1 转向以终为始的 V2 后,用词级时间戳对齐、术语表注入和 Agent 自验证替代人工校对,一次成本优化把调用从 33 次降到 12 次、单集处理从 31 分钟降到 18 分钟。
Awaiting translation
Addy Osmani 在博客中提出,智能体能跳过写代码、调试、读别人代码这些原本积累经验的环节,因此刻意练习变得必要,他建议在提示前先形成假设、多问为什么、读 diff、预测失败点,并偶尔手写小问题。
Awaiting translation
WikiSkill 是一个让 Agent Skill 与持久知识库(wiki)协同演化的框架,它把原始执行经验、累积知识和可执行 Skill 分离,并持续将经验沉淀进 wiki 供后续 Skill 更新使用。
Awaiting translation
Warp 内部尝试用 AI 做 Code Review 时发现 Agent 不了解项目、团队规范和历史经验,手动改系统提示词和 AGENTS.md 效果都不理想。
Awaiting translation
Addy Osmani 建议每隔几周运行一次 Claude Code 的 /doctor,单独用 /memory 检查记忆,并让每条指令重新证明自己的价值,因为模型、harness 和代码库都在变,旧配置会留下。
Awaiting translation
Why it matters: 作者结合自身配置审计经验与近期研究,说明 Agent 配置文件为何会腐化,以及如何按节奏清理。
The Cline team built a code review agent with the Cline SDK, splitting review into two agent loops—review and judge—then using a driver script to batch-submit the surviving issues as a single COMMENT event to the GitHub PR.
Why it matters: A full breakdown of the plugin, Hooks, and two-stage loop behind a code review agent, transferable to other automated review scenarios.
V2EX 网友讨论 Vibe Coding 如何清理 AI 生成的冗余代码。有回复建议用 LSP、eslint 等工具接上 MCP,一句指令即可完成且节省 token;也有人指出 AI 倾向用大量 patch 处理逻辑上不会发生的 case,而非解决 root cause,需要定期重构精简。
Awaiting translation
GitSkills 数据集对 187 万个去重 Skill 做语言识别后发现,14.3% 的 Skill 正文不是英语,其中中文占 6.2%,高于 GitHub 仓库文档 3.3% 的中文比例。非英语 Skill 占比从 2026 年 Q1 的 13.0% 升至 Q2 的 16.3%,三个月涨 3 个百分点,而 GitHub 整体非英语文档从 3.7% 到 13.0% 用了十年。
Awaiting translation
宝玉以给字幕转录翻译 App BaoCut 加远程转录功能为例,复盘了自己的 AI 原生开发流程:可行性分析、设计文档、高精度原型、编码实现、测试验证五步一个没少,但执行主体从人变成 Agent,人只在关键路径做确认和决策。
Awaiting translation
Paper Compute 团队审计了 6 月 16 日至 8 月 14 日两个月的 766 次真实 Agent 会话,其中只有 29 次被沉淀为可复用 Skill,其余 737 次留在会话库里。
Awaiting translation
Paper Compute 将 tapes 以 Apache/MIT 双许可开源,并推出 cassettes 集成范式,用于连接、加工和构建 AI trace 数据。
Awaiting translation
Drew Breunig 提出用「情境化智能体」来定义 harness:Harrison Chase 所说的系统提示词、规划工具、文件系统和子智能体构成开发者控制的核心循环,harness 则管理循环之外的世界,包括会话、环境、仓库、记忆、Skills、团队、组织和模型这些由内向外、使用人数递增而变动递减的层次。
Awaiting translation
Anthropic 官方博客拆解 Claude Code 会话的 token 成本:请求分 prefill 与 decode 两阶段,输出 token 单价约为输入的 5 倍,缓存读取为输入价格的 0.1x、写入最高 2x。
Awaiting translation
Baruch Sadogursky 与 Patrick Debois 在 InfoQ 演讲中用 Claude Code 现场演示:把全部项目文档塞进 CLAUDE.md 后,给接口加错误处理会因约定冲突返回 500,改用按描述懒加载的 Skill 后同一提示词通过测试。
Awaiting translation
Why it matters: 两位作者用现场演示拆解上下文工程的四类反模式,并给出 Skill、检索通道、外部记忆与评测的对应做法。
AI Hero Skills v1.2 发布,整套 Skill 打包为 Claude Code 插件,并为每个 SKILL.md 添加 Codex 元数据,使同一套 Skill 在 Claude Code 和 Codex 中通用。
Awaiting translation
Vercel ships v0 API, giving programmatic, headless access to the v0 app-generation agent: send a prompt, v0 generates an app, spins up a dev server in the Vercel Sandbox, and returns a preview URL you can embed in your own UI. The API is now generally available.
Why it matters: v0 opens up its app-generation capability as an API, so readers can judge how to wire it into their own product or agent workflow.
文章提出把每日站会当作检测层,用来发现分散在个人 AI 辅助会话中的重复工作:当多名工程师用智能体各自重新解决同一问题时,重复不会出现在工单看板上,却会先以抱怨或玩笑的形式在站会里冒出来。
Awaiting translation
The OpenAI Codex Cookbook lays out a set of repository conventions for wiring Codex into your development process: use AGENTS.md for persistent repository instructions, PLANS.md as the source of phase plans, and split the work into phased build files under harness/build/, with each phase spelling out its goals, acceptance criteria, boundaries, and approval gates.
Why it matters: OpenAI lays out a complete directory convention for constraining Codex with AGENTS.md, PLANS.md, and phased build files—one you can adapt to your own repository.
Hacker News 上一位用户发帖追问:Claude Code、Codex 等框架为何要引入 Skill 概念,而不是用组织良好的 Markdown 文档加 AGENTS.md 指向文件。
Awaiting translation
SimpleEnglish 是一个让大语言模型按 ASD-STE100 简化技术英语写作的 Agent Skill,兼容 Claude Code、Cursor、VS Code Copilot、OpenAI Codex、Gemini CLI 等遵循 Agent Skills 标准的工具,MIT 许可、无依赖。
Awaiting translation
作者发布 Bullshit Detector,一套可移植的 Agent Skills,把 YouTube、TikTok、文章、推文或 PDF 拆成逐条声明,联网核查后给出确认、存疑、误导、错误、无法核实五类判定和 0–10 的 BS 分数,并附来源链接。
Awaiting translation
面对 GitHub 等基础设施被 AI 智能体流量压垮的现状,作者主张从 tokenmaxxing 转向 valuemaxxing,用任务完成数、节省时间和避免返工来衡量价值,而非 token 消耗量。他指出 Claude Code 会话默认 30 天后删除,导致已付费的上下文白白流失,并认为 Skill 是比 markdown 文件更好的上下文路由方式,但大规模管理 Skill 仍无解。
Awaiting translation
宝玉针对 @cellier_ 提出的三个 Agent 判断给出补充:通用 Agent 赛道将是少数赢家通吃,小通用 Agent 同样没有生存空间。他认为开放与封闭的差别不在开源或能否切换模型,而在插件生态,Skill 加 MCP 让 Agent 能完成各类任务。他还判断 Agent 产品体验护城河不深,模型能力和成本才是拉开差距的关键,小团队更适合基于 Agent 做插件。
Awaiting translation
Thariq Shihipar, a member of Anthropic's engineering team, wrote up the new context engineering rules for Claude 5, saying the team has cut over 80% of the system prompt from Claude Code for models like Claude Opus 5 and Claude Fable 5, with no measurable loss on coding evals.
Why it matters: Anthropic lays out the new context engineering rules for Claude 5 and explains how to trim the system prompt, CLAUDE.md, and Skills.
Drew Breunig 发布 drskill,用于评估全局或项目环境中的 Skill 与 MCP,可通过 uv tool install drskill 或 pip install drskill 安装后运行 drskill scan。
Awaiting translation
Ingot 是面向个人开发者的本地优先库和 MCP server,为 Agent 的 Skill 指令提供证据门禁的变更控制。
Awaiting translation
JetBrains 用 Claude Code 跑 SkillsBench 的 86 个真实编程任务,对比安装与不安装 Caveman 的效果,发现输出 Token 只从约 59.2 万降到 54.2 万,节省 8.5%,远低于项目宣传的 65%。
Awaiting translation