Agent Skills: Turning Senior-Engineer Workflows into Rules for AI Coding Agents
addyosmani/agent-skills is a set of 25 engineering skills and 9 slash commands for AI coding agents. It covers six lifecycle stages: define, plan, build, verify, review and ship. The repository turns senior-engineer habits, such as spec before code, small commits and tests as proof, into rules that agents like Claude Code, Cursor and Codex can load by context. Its /build auto mode needs one plan approval yet keeps per-task tests and commits. One npx command installs it.
Background and Problem Definition
AI coding agents can now produce working code, yet working code is not the same as shippable code. The same model may write tests first today and skip them tomorrow. It may review a change in one session and ignore review in the next. The weak point is not raw capability. It is consistency. Nobody has turned the habits of a senior engineer into rules that an agent follows every time.
addyosmani/agent-skills targets exactly this gap. The repository describes itself as production-grade engineering skills for AI coding agents. It packages the workflows, quality gates and best practices that senior engineers use, so that agents can follow them across every phase of development. At the time of writing the repository has about 101,276 stars. Its topics include agent-skills, claude-code, codex, cursor and antigravity, which signals that it aims at many agent tools rather than one vendor.
Architectural Core and Technical Principles
The backbone is a six-stage lifecycle: Define, Plan, Build, Verify, Review and Ship. Each stage maps to a slash command: /spec, /plan, /build, /test, /review and /ship. The README lists nine commands in total. The other three are /constraints for setting the quality bar, /webperf for auditing web performance, and /code-simplify for simplifying code. Each command carries a short key principle. Spec before code. Small, atomic tasks. One slice at a time. Tests are proof. Decide the quality bar once and enforce it everywhere. Improve code health in review. Measure before you optimize. Clarity over cleverness. Faster is safer. These lines read as plain advice, but each one counters a known agent failure: oversized changes, skipped tests and premature optimization.
The second layer is automatic activation. According to the README, skills activate based on what you are doing. Designing an API triggers api-and-interface-design. Building UI triggers frontend-ui-engineering. The commands are explicit entry points, and the skills are knowledge packs attached to the situation. The repository holds 25 skills. Named examples include code-review-and-quality, a five-axis review before merge, interview-me, which interrogates requirements one question at a time, and test-driven-development, which enforces red, green, refactor. The third layer is /build auto. When a spec already exists, it generates the plan and implements every task in one approved pass. The human approves the plan once. The README is careful to say that this removes the human stepping between tasks, not the verification. Every task stays test-driven and is committed on its own, and the run pauses on failures or risky steps. This design separates the level of automation from the strength of verification, and it is the most notable trade-off in the project.
Practical Evaluation and Applications
Installation is short. The general route is the open skills CLI. The command npx skills add addyosmani/agent-skills installs all 25 skills. Adding --list lets you browse first, and --skill with a name installs one skill only. The README says the CLI reaches more than 70 agents, including Claude Code, Cursor, Codex, Copilot and Cline. Claude Code users can also use the plugin marketplace with /plugin marketplace add addyosmani/agent-skills, then /plugin install agent-skills@addy-agent-skills.
The README admits one real gotcha. A per-skill install copies only the skill folder, not the repo-level references directory. The skill still works, but paths to shared checklists break. The suggested workarounds are a whole-repo integration, a clone of the repository, or copying the needed checklist into a references directory inside the installed skill. The gap is tracked in issue 361. Installing a single skill is therefore not free.
One limit must be stated plainly. This article rests on the README summary. We did not read all 25 skills in full, and we ran no controlled experiment. We cannot report how much a team's defect rate falls after adoption. The visible value today is the completeness of the process design and the ease of distribution. Real impact depends on whether your agent actually obeys the skills, and each team should measure that in its own codebase.
Industry Impact and Outlook
The importance of this project lies less in algorithms and more in packaging engineering discipline as a distributable artifact. Team standards used to live in a wiki and relied on goodwill. Here they live inside the agent context and are enforced by slash commands and situational triggers. A star count above one hundred thousand shows how strong the demand is for agents that follow the rules.
It also shows a trend: the skill format is converging across tools. One skill set can be read by Claude Code, Cursor, Codex and others, so a team can keep its process assets in a repository instead of a vendor setting. Three points deserve watching. First, when the distribution gap for shared references will close. Second, whether the pause logic of autonomous modes such as /build auto reacts early enough in real projects. Third, whether automatic triggering stays accurate as the skill count grows. A team that wants to adopt it should first run the full /spec to /ship chain on one small project, then decide on wider rollout.
Sources
FAQ
How many skills and commands does agent-skills provide, and which stages do they cover?
The README states 25 skills and 9 slash commands across six stages: Define, Plan, Build, Verify, Review and Ship. The commands are /spec, /plan, /build, /test, /constraints, /review, /webperf, /code-simplify and /ship.
Does /build auto skip verification?
No. The README says it removes only the human stepping between tasks. The plan is approved once, but every task stays test-driven and is committed individually, and the run pauses on failures or risky steps.
What is the limit of installing a single skill?
A per-skill npx install copies only the folder under skills, not the repo-level references directory, so paths to shared checklists break. Use a whole-repo integration, clone the repository, or copy the checklist into the skill's own references folder. The gap is tracked in issue 361.