A study of 8,135 agent trials finds skills work mainly as procedural anchors. Use this evidence-based checklist to design better SKILL.md files.
产品的 AI 可用性要看 Agent 能否完成真实任务,并通过外部状态、恢复能力、成本与回归稳定性验证。
Test whether AI agents can use your product with dynamic task evals that measure correct state changes, recovery, cost, and regression stability.
用内容、发现、调用与运行时四层模型,判断 SKILL.md 在 Codex、Claude Code、Gemini CLI 和 GitHub Copilot 之间到底能移植什么。
A four-layer compatibility model for moving SKILL.md workflows across Codex, Claude Code, Gemini CLI, and GitHub Copilot without confusing shared synt