Skip to content

Writing Agent Skills, finding useful Skills, and troubleshooting Skills that do not take effect.

Latest curated items

Items 1–20 · 33 total
Today10/6Tue
  1. DEV Community · MCP82

    Skills Are Not Tools: Why I Gave My Coding Agent 11 MCP Servers and It Fell Apart

    In February I hooked up 11 MCP Servers to my coding agent. Just the tool list in an empty session ate up 34,000 tokens, with Datadog alone contributing over two hundred tools. The agent got slower and kept picking the wrong tools.

    Why it matters: I'm writing this up because 11 MCP Servers of my own blew up my context, and it showed me that tools and knowledge belong in different containers.

10/5Mon
  1. AI Hero · Skills Updates62

    AI Hero Skills v1.3 released: new /implement-spec, /pr, and /retro skills, and CONTEXT.md renamed to GLOSSARY.md

    AI Hero Skills v1.3 is out, adding three skills—/implement-spec, /pr, and /retro—and extending the main flow from /grill-with-docs → /to-spec → /to-tickets into implementation, PRs, and retros.

    Why it matters: The author rounds out the skill set into a full path from writing specs to PRs and retros, and lays out the trade-offs at each step along with the rough edges they already know about.

10/4Sun
  1. 宝玉82

    Drawing on a podcast episode, Baoyu walks through how Lauren Tan, who works on Grok Bot at SpaceXAI, merged 2500 PRs in a single month: at night she lets the AI check and merge on its own, then spot-checks the next morning instead of reviewing each one.

    Quotedlauren@poteto

    i had a lot of fun chatting with @mattpocockuk today about how i was able to land 2,500 PRs last month! Matt is a wonderful interviewer so i think the interview turned out really interesting both of our skill plugins work great together, so i recommend giving both a try and picking the best skills that suit your workflow https://www.youtube.com/watch?v=MN9dGgmLyso

    Why it matters: Using Lauren Tan's practice of merging 2500 PRs in a month, Baoyu explains that skipping individual reviews rests on validation Skills and rule constraints, and lays out the conditions under which he'd apply the same approach.

9/11Fri
  1. OpenAI Developer Blog · Codex66

    OpenAI on How to Rewrite Skills and Prompts for GPT-6 Astra

    In an official blog post, OpenAI lays out recommendations for adjusting Skills, AGENTS.md, and task prompts under GPT-6 Astra: Skill descriptions should be as short as possible and state clearly when they apply, and multi-flow Skills should use a root document for minimal routing instead of turning the Skill into an overly specific step-by-step checklist.

    Why it matters: OpenAI has published guidance on cleaning up Skills, AGENTS.md, and prompts under GPT-6 Astra, and it carries over to existing repository setups.

9/10Thu
  1. Vibe Code Textbook · Articles78

    编程智能体的四种提示词模式:plan mode、skills 与保存的提示词

    文章从 Claude Code、Codex 和 Gemini CLI 的官方文档中整理出四种提示词模式:先计划再编辑、给智能体一个可运行的检查、让智能体反过来访谈你、把反复重打的提示词存成文件,并给出各家对应的命令、参数和文件格式。

    Awaiting translation

    Why it matters: 横向对照 Claude Code、Codex、Gemini CLI 三家文档,给出计划模式、可运行检查、访谈式提问和保存提示词四种模式的命令与文件格式。

9/8Tue
  1. Habr · Cursor82

    How Cursor engineers merge 800+ PRs a month: using validation skills and evals to build trust in AI agents

    Cursor engineer Lauren Tan shares how she gets AI agents to submit and merge PRs on their own: the key is validation—letting the agent run code, capture CPU traces, and open an iOS simulator to check its own work.

    Why it matters: Cursor engineers break trust in AI agents down into reusable validation skills, feature maps, and evals, so readers can build their own automated validation workflows.

9/5Sat
  1. Ryan Lopopolo66

    An agent platform built for inventing agents: decoupling capability interfaces from their implementations

    Author Ryan Lopopolo argues that an agent is a parameterized program built on top of a set of capabilities: models and configurations, reasoning and tool-call loops, computers, disks, context, Skills, tools, connectors, runtimes, network policies, identity, IAM, guardrails, I/O channels, and system prompts.

    Why it matters: Drawing on his experience building multiple agents, the author proposes a platform architecture that decouples capability interfaces from their implementations — a useful reference for teams building Agent platforms.

8/27Thu
  1. Addy Osmani · Blog71

    如何审计你的 Agent 配置文件:CLAUDE.md、Skills 与 Hooks 的定期清理

    Addy Osmani 建议每隔几周运行一次 Claude Code 的 /doctor,单独用 /memory 检查记忆,并让每条指令重新证明自己的价值,因为模型、harness 和代码库都在变,旧配置会留下。

    Awaiting translation

    Why it matters: 作者结合自身配置审计经验与近期研究,说明 Agent 配置文件为何会腐化,以及如何按节奏清理。

8/26Wed
  1. Cline · Blog71

    Building a Code Review Agent on the Cline Loop with the Cline SDK

    The Cline team built a code review agent with the Cline SDK, splitting review into two agent loops—review and judge—then using a driver script to batch-submit the surviving issues as a single COMMENT event to the GitHub PR.

    Why it matters: A full breakdown of the plugin, Hooks, and two-stage loop behind a code review agent, transferable to other automated review scenarios.

8/14Fri
  1. InfoQ · AI Coding Presentations80

    InfoQ 演讲:300 个精准 token 胜过 10 万个噪声 token,上下文工程的架构

    Baruch Sadogursky 与 Patrick Debois 在 InfoQ 演讲中用 Claude Code 现场演示:把全部项目文档塞进 CLAUDE.md 后,给接口加错误处理会因约定冲突返回 500,改用按描述懒加载的 Skill 后同一提示词通过测试。

    Awaiting translation

    Why it matters: 两位作者用现场演示拆解上下文工程的四类反模式,并给出 Skill、检索通道、外部记忆与评测的对应做法。

8/5Wed
  1. Vercel · v0 Blog62

    Vercel ships v0 API for programmatic access to its app-generation agent

    Vercel ships v0 API, giving programmatic, headless access to the v0 app-generation agent: send a prompt, v0 generates an app, spins up a dev server in the Vercel Sandbox, and returns a preview URL you can embed in your own UI. The API is now generally available.

    Why it matters: v0 opens up its app-generation capability as an API, so readers can judge how to wire it into their own product or agent workflow.

8/3Mon
  1. OpenAI · Codex Cookbook71

    Iterating on a Development Workflow with Codex: From AGENTS.md to Phased Build Files

    The OpenAI Codex Cookbook lays out a set of repository conventions for wiring Codex into your development process: use AGENTS.md for persistent repository instructions, PLANS.md as the source of phase plans, and split the work into phased build files under harness/build/, with each phase spelling out its goals, acceptance criteria, boundaries, and approval gates.

    Why it matters: OpenAI lays out a complete directory convention for constraining Codex with AGENTS.md, PLANS.md, and phased build files—one you can adapt to your own repository.

7/26Sun
  1. Hacker News · Context Engineering 讨论82

    Anthropic publishes new context engineering rules for Claude 5

    Thariq Shihipar, a member of Anthropic's engineering team, wrote up the new context engineering rules for Claude 5, saying the team has cut over 80% of the system prompt from Claude Code for models like Claude Opus 5 and Claude Fable 5, with no measurable loss on coding evals.

    Why it matters: Anthropic lays out the new context engineering rules for Claude 5 and explains how to trim the system prompt, CLAUDE.md, and Skills.

7/8Wed
  1. AI Hero · Skills Updates65

    AI Hero skills repo ships v1.1: adds /wayfinder, renames /to-spec and /to-tickets

    AI Hero's skills repo ships v1.1, renaming /to-prd to /to-spec, merging /to-plan and /to-issues into /to-tickets, and adding new Skills like /wayfinder, /research, and /prototype.

    Why it matters: The author walks through the full Skill flow from grilling to deployment and gives the migration commands for the renames, the merge, and the new /wayfinder—useful for anyone building an AI development workflow.

7/6Mon
  1. 宝玉80

    The Claude Code team explains the four types of AI agent loops and how to use them

    The Claude Code team defines an AI agent loop as repeatedly running a work cycle until a predefined stopping condition is met, and splits it into four types—turn-based, goal-oriented, time-based, and proactive—along four dimensions: what triggers it, how it stops, the underlying instructions it uses, and the tasks it fits.

    Why it matters: The Claude Code team breaks loops into four types—turn-based, goal-oriented, time-based, and proactive—and lays out how each is triggered, how it stops, and how tokens are controlled.

6/18Thu
  1. AI Hero · Skills Updates69

    AI Hero Skills v1 发布:token 消耗降低 63%,新增 /ask-matt 与 /writing-great-skills

    AI Hero 的 skills 目录发布 v1,通过在各 Skill 上启用 disable-model-invocation: true,让 Skill 描述不再进入模型选择 Skill 时查看的上下文窗口,Skill 描述的 token 成本降低 63%。

    Awaiting translation

    Why it matters: v1 用 disable-model-invocation 把 Skill 描述移出上下文窗口,并区分用户调用与模型调用,读者可据此判断自己的 Skill 组织方式。

6/3Wed
  1. claude.dev · Anthropic Developer Blog86

    Anthropic shares what it learned from using Skills inside Claude Code: nine categories and tips for writing them

    Inside Claude Code, Anthropic has already built up hundreds of Skills in active use. The team sorts them into nine categories—library and API references, product validation, data fetching and analysis, business process automation, code scaffolding, code quality and review, CI/CD and deployment, runbooks, and infrastructure operations—and notes that the best Skills should fall cleanly into one of them.

    Why it matters: Anthropic’s internal framework for categorizing hundreds of Skills, along with its experience writing them, can carry over to a team building its own Skill library.

5/18Mon
  1. Lovable · Blog62

    Lovable 上线 Skills,把重复指令变成可复用技能

    Lovable 上线 Skills 功能,把重复交代的工作方式写成可复用的 markdown 技能文件,在相关任务出现时按需加载。技能以文件夹形式组织,主文件 SKILL.md 含 name、description 和 instructions,description 是决定是否触发的唯一依据,支持文件只在主文件引用且确实需要时才加载。

    Awaiting translation

    Why it matters: 官方详解 Lovable Skills 的文件结构、触发机制与写法,并给出可对照的正反示例。

4/24Fri
  1. Jesse Vincent66

    Prime Radiant 发布 Greenfield 与 Iterative Development 研究预览

    Prime Radiant 发布两款新技术的研究预览:Greenfield 把现有软件(代码库、文档、API 客户端等)转成行为规格语料,Iterative Development 则是一套基于 Superpowers 的智能体方法论,把大规格拆成需求、打包成开发 epic 后交给编码智能体实现。

    Awaiting translation

    Why it matters: Prime Radiant 公开两套工具的研究预览,读者可了解从旧代码库提取行为规格再驱动智能体重建产品的思路。

4/10Fri
  1. claude.dev · Anthropic Developer Blog74

    Anthropic 工程师谈 Claude Code 的工具设计:如何像智能体一样思考

    Anthropic 的 Thariq Shihipar 复盘了 Claude Code 工具设计中的取舍,核心主张是工具要贴合模型自身能力,而判断能力边界只能靠观察输出和反复实验。

    Awaiting translation

    Why it matters: Anthropic 工程师复盘 Claude Code 工具设计的取舍,给出可迁移到自建智能体的判断方法。