← プロジェクト一覧に戻る

Addy Osmani's Agent Skills

仕様・計画・実装・テスト・レビュー・公開という順序を、コーディングエージェントに守らせる25個のスキル

エージェントスキルMIT
スター
90.5k
フォーク
9.7k
オープンIssue
121
最終コミット
2026年8月28日

Addy Osmani's Agent Skillsとは

放っておくと、コーディングエージェントは指示を言い終えた瞬間からコードを書き始めます。これは、そこに本来の開発の順序を戻すスキル集です。9つのコマンドが、何を作るかを決める段階から公開までを覆い、背後のスキルは作業内容に応じて自動で有効になります。インターフェースの設計を始めればその担当スキルが、画面側に触れればそちら向けのスキルが働きます。注目すべきは自動モードの扱いで、省かれるのは各作業の合間に人が入ることだけであり、確認は残ります。各作業は引き続きテスト駆動で行われ、個別にコミットされ、失敗すれば止まります。実体はMarkdownなので、共通のskillsコマンドで70種類ほどのエージェントに導入できます。

Addy Osmani's Agent Skillsで何ができますか?

  • まず仕様を書かせる — 1つのコマンドが要件を1問ずつ問いただします。コードを1行も書く前に、何を作るのかが決まります。
  • コミットできる大きさに作業を割る — 計画のスキルは1つの大きな変更ではなく、小さく独立した作業に分けます。後からレビューできる成果物になるのはこのためです。
  • 人が離れても確認は残す — 自動モードは1度の承認で計画全体を実行しますが、各作業はテスト駆動のまま個別にコミットされ、失敗すれば停止します。
  • 頼む前にレビューを済ませる — レビュー用のスキルが5つの観点で変更を通します。同僚に渡る前に、指摘しやすい問題を先に片付けられます。
  • 必要なスキルが自分で立ち上がる — API設計、画面側、性能といった作業内容に応じてスキルが有効になるため、どれを呼ぶかを覚えておく必要がありません。
  • 使っているエージェントに導入する — 共通のskillsコマンドで配布される素のMarkdownなので、70種類ほどのエージェントに、まとめてでも1つずつでも入れられます。

Addy Osmani's Agent Skillsを選ぶ前に

  • スキルを1つだけ導入すると共通のチェックリストが欠けるとREADMEは注意していますが、追跡していたIssueはその後閉じています。どちらとも決めつけず手元で確認してください。
  • ここに書かれているのは特定の開発観です。仕様が先、コミットは小さく、テストが根拠、という前提です。自分たちの進め方が本当に違う場合、スキルは合わせるのではなく抵抗します。

よくある質問

Addy Osmani's Agent Skillsは商用利用できますか?

Addy Osmani's Agent SkillsはMITライセンスで公開されています。OSI承認のオープンソースライセンスで、商用利用が認められています。

Addy Osmani's Agent Skillsはどの形で使えますか?

Addy Osmani's Agent Skillsはローカル実行の形で利用できます。

ドキュメント

addyosmani/agent-skills のREADMEより転載(MIT)。 原文を読む ↗

Agent Skills

Production-grade engineering skills for AI coding agents.

Skills encode the workflows, quality gates, and best practices that senior engineers use when building software. These ones are packaged so AI agents follow them consistently across every phase of development.

Addy's Agent Skills

  DEFINE          PLAN           BUILD          VERIFY         REVIEW          SHIP
 ┌──────┐      ┌──────┐      ┌──────┐      ┌──────┐      ┌──────┐      ┌──────┐
 │ Idea │ ───▶ │ Spec │ ───▶ │ Code │ ───▶ │ Test │ ───▶ │  QA  │ ───▶ │  Go  │
 │Refine│      │  PRD │      │ Impl │      │Debug │      │ Gate │      │ Live │
 └──────┘      └──────┘      └──────┘      └──────┘      └──────┘      └──────┘
  /spec          /plan          /build        /test         /review       /ship

Commands

9 slash commands that map to the development lifecycle. Each one activates the right skills automatically.

What you’re doingCommandKey principle
Define what to build/specSpec before code
Plan how to build it/planSmall, atomic tasks
Build incrementally/buildOne slice at a time
Prove it works/testTests are proof
Set the quality bar/constraintsDecide it once, enforce it everywhere
Review before merge/reviewImprove code health
Audit web performance/webperfMeasure before you optimize
Simplify the code/code-simplifyClarity over cleverness
Ship to production/shipFaster is safer

Want fewer manual steps once the spec exists? /build auto generates the plan and implements every task in a single approved pass — you approve the plan once, then it runs autonomously. It removes the human stepping between tasks, not the verification: every task is still test-driven and committed individually, and it pauses on failures or risky steps.

Skills also activate automatically based on what you’re doing — designing an API triggers api-and-interface-design, building UI triggers frontend-ui-engineering, and so on.


Quick Start

Fastest path — any agent, one command. The open skills CLI installs into 70+ agents (Claude Code, Cursor, Codex, Copilot, Cline, and more):

npx skills add addyosmani/agent-skills            # install all 25 skills
npx skills add addyosmani/agent-skills --list     # browse before installing

Or grab individual skills:

npx skills add addyosmani/agent-skills --skill code-review-and-quality   # five-axis review before merge
npx skills add addyosmani/agent-skills --skill interview-me              # requirements interrogation, one question at a time
npx skills add addyosmani/agent-skills --skill test-driven-development   # red-green-refactor, enforced

Installing one skill? A per-skill npx install copies only skills/<name>/, not the repo-level references/ directory. The skill still works, but paths to supplementary shared checklists are unavailable. Use a whole-repo integration, clone the repository, or copy the needed checklist into a references/ directory inside the installed skill. This portability gap is tracked in #361.

Prefer a native integration? Pick your tool below.

Marketplace install:

/plugin marketplace add addyosmani/agent-skills
/plugin install agent-skills@addy-agent-skills

SSH errors? The marketplace clones repos via SSH. If you don’t have SSH keys set up on GitHub, either add your SSH key or use the full HTTPS URL to force HTTPS cloning during the marketplace-add step:

/plugin marketplace add https://github.com/addyosmani/agent-skills.git
/plugin install agent-skills@addy-agent-skills

If /plugin install still fails with git@github.com: Permission denied (publickey) on Windows or macOS, the recommended workaround is to configure Git once to rewrite GitHub SSH URLs to HTTPS for subprocess clones:

git config --global url."https://github.com/".insteadOf git@github.com:

Local / development:

git clone https://github.com/addyosmani/agent-skills.git
claude --plugin-dir /path/to/agent-skills

Put workflow skills under .cursor/skills/ (sync from agent-skills/skills/) and short policies in .cursor/rules/*.mdc — do not paste full skills into rules. See docs/cursor-setup.md.

Install as a native plugin for skills, subagents, and slash commands. See docs/antigravity-setup.md.

Install from the repo:

agy plugin install https://github.com/addyosmani/agent-skills.git

Install from a local clone:

git clone https://github.com/addyosmani/agent-skills.git
agy plugin install ./agent-skills

Install as native skills for auto-discovery, or add to GEMINI.md for persistent context. See docs/gemini-cli-setup.md.

Install from the repo:

gemini skills install https://github.com/addyosmani/agent-skills.git --path skills

Install from a local clone:

gemini skills install ./agent-skills/skills/

Add skill contents to your Windsurf rules configuration. See docs/windsurf-setup.md.

Copy skills to .opencode/skills/ (or ~/.config/opencode/skills/), add a project-local AGENTS.md, and use the built-in skill tool for agent-driven execution. Optional slash commands can be added under .opencode/commands/.

See docs/opencode-setup.md.

Use agent definitions from agents/ as Copilot personas and skill content in .github/copilot-instructions.md. See docs/copilot-setup.md.

Install as a native Codex plugin (Codex CLI v0.122+):

codex plugin marketplace add addyosmani/agent-skills
codex plugin add agent-skills@agent-skills

The first command registers the marketplace; the second installs the plugin. Codex reads the root skills/ directory directly through .codex-plugin/plugin.json. Once installed, invoke skills in chat using @ (e.g., @spec-driven-development). See docs/codex-setup.md for local installation and troubleshooting.

Install natively with the built-in cmd skills command. Command Code clones the repo, discovers every SKILL.md, and installs into .commandcode/skills/:

cmd skills add addyosmani/agent-skills            # pick skills to install (project)
cmd skills add addyosmani/agent-skills --global   # install for all projects (~/.commandcode/skills/)
cmd skills add addyosmani/agent-skills -s spec-driven-development  # install a specific skill

Installed skills show up in the TUI slash menu, e.g. /spec-driven-development. See docs/commandcode-setup.md.

Skills are plain Markdown - they work with any agent that accepts system prompts or instruction files. See docs/getting-started.md.


Adoption

Already installed? How you roll the pack out depends on your codebase. The Adoption Guide covers two paths: the full lifecycle from day one for a greenfield project, or an incremental, verification-first rollout for an established codebase.


All 24 Skills

The commands above are entry points. The pack includes 25 skills total — 24 lifecycle skills plus the using-agent-skills meta-skill. Each skill is a structured workflow with steps, verification gates, and anti-rationalization tables. You can also reference any skill directly.

Meta - Discover which skill applies

SkillWhat It DoesUse When
using-agent-skillsMaps incoming work to the right skill workflow and defines shared operating rulesStarting a session or deciding which skill applies

Define - Clarify what to build

SkillWhat It DoesUse When
interview-meOne-question-at-a-time interview that extracts what the user actually wants instead of what they think they should want, until ~95% confidenceThe ask is underspecified, or the user invokes “interview me” / “grill me”
idea-refineStructured divergent/convergent thinking to turn vague ideas into concrete proposalsYou have a rough concept that needs exploration
spec-driven-developmentWrite a PRD covering objectives, commands, structure, code style, testing, and boundaries before any codeStarting a new project, feature, or significant change
constraint-driven-developmentInterviews you for a quality bar with sane default thresholds, writes CONSTRAINTS.md, places each check by cost, and catches agents silencing checks or skipping tests to get greenNo standards are written down, or an agent is producing more than anyone reads

Plan - Break it down

SkillWhat It DoesUse When
planning-and-task-breakdownDecompose specs into small, verifiable tasks with acceptance criteria and dependency orderingYou have a spec and need implementable units

Build - Write the code

SkillWhat It DoesUse When
incremental-implementationThin vertical slices - implement, test, verify, commit. Feature flags, safe defaults, rollback-friendly changesAny change touching more than one file
test-driven-developmentRed-Green-Refactor, test pyramid (80/15/5), test sizes, DAMP over DRY, Beyonce Rule, browser testingImplementing logic, fixing bugs, or changing behavior
context-engineeringFeed agents the right information at the right time - rules files, context packing, MCP integrationsStarting a session, switching tasks, or when output quality drops
source-driven-developmentGround every framework decision in official documentation - verify, cite sources, flag what’s unverifiedYou want authoritative, source-cited code for any framework or library
doubt-driven-developmentAdversarial fresh-context review of every non-trivial decision in-flight - CLAIM → EXTRACT → DOUBT → RECONCILE → STOP, with optional user-authorized cross-model escalationStakes are high (production, security, irreversible), working in unfamiliar code, or a confident output is cheaper to verify now than to debug later
frontend-ui-engineeringComponent architecture, design systems, state management, responsive design, WCAG 2.1 AA accessibilityBuilding or modifying user-facing interfaces
api-and-interface-designContract-first design, Hyrum’s Law, One-Version Rule, error semantics, boundary validationDesigning APIs, module boundaries, or public interfaces

Verify - Prove it works

SkillWhat It DoesUse When
browser-testing-with-devtoolsChrome DevTools MCP for live runtime data - DOM inspection, console logs, network traces, performance profilingBuilding or debugging anything that runs in a browser
debugging-and-error-recoveryFive-step triage: reproduce, localize, reduce, fix, guard. Stop-the-line rule, safe fallbacksTests fail, builds break, or behavior is unexpected

Review - Quality gates before merge

SkillWhat It DoesUse When
code-review-and-qualityFive-axis review, change sizing (~100 lines), severity labels (Nit/Optional/FYI), review speed norms, splitting strategiesBefore merging any change
code-simplificationChesterton’s Fence, Rule of 500, reduce complexity while preserving exact behaviorCode works but is harder to read or maintain than it should be
security-and-hardeningOWASP Top 10 prevention, auth patterns, secrets management, dependency auditing, three-tier boundary systemHandling user input, auth, data storage, or external integrations
performance-optimizationMeasure-first approach - Core Web Vitals targets, profiling workflows, bundle analysis, anti-pattern detectionPerformance requirements exist or you suspect regressions

Ship - Deploy with confidence

SkillWhat It DoesUse When
git-workflow-and-versioningTrunk-based development, atomic commits, change sizing (~100 lines), the commit-as-save-point patternMaking any code change (always)
ci-cd-and-automationShift Left, Faster is Safer, feature flags, quality gate pipelines, failure feedback loopsSetting up or modifying build and deploy pipelines
deprecation-and-migrationCode-as-liability mindset, compulsory vs advisory deprecation, migration patterns, zombie code removalRemoving old systems, migrating users, or sunsetting features
documentation-and-adrsArchitecture Decision Records, API docs, inline documentation standards - document the whyMaking architectural decisions, changing APIs, or shipping features
observability-and-instrumentationStructured logging, RED metrics, OpenTelemetry tracing, symptom-based alerting - instrument as you buildAdding telemetry, or shipping anything that runs in production
shipping-and-launchPre-launch checklists, feature flag lifecycle, staged rollouts, rollback procedures, monitoring setupPreparing to deploy to production

Agent Personas

Pre-configured specialist personas for targeted reviews:

AgentRolePerspective
code-reviewerSenior Staff EngineerFive-axis code review with “would a staff engineer approve this?” standard
test-engineerQA SpecialistTest strategy, coverage analysis, and the Prove-It pattern
security-auditorSecurity EngineerVulnerability detection, threat modeling, OWASP assessment
web-performance-auditorWeb Performance EngineerCore Web Vitals audit with Quick/Deep modes and a metric-honesty rule; run it via /webperf

See docs/agents.md for the decision matrix, orchestration rules, and how personas compose with skills and slash commands.


Reference Checklists

Quick-reference material that skills pull in when needed:

ReferenceCovers
definition-of-done.mdProject-wide standing bar every change clears, contrasted with per-task acceptance criteria
testing-patterns.mdTest structure, naming, mocking, React/API/E2E examples, anti-patterns (JavaScript/TypeScript)
security-checklist.mdPre-commit checks, auth, input validation, headers, CORS, OWASP Top 10
performance-checklist.mdCore Web Vitals targets, frontend/backend checklists, measurement commands
accessibility-checklist.mdKeyboard nav, screen readers, visual design, ARIA, testing tools
observability-checklist.mdOn-call questions, structured logging, RED/USE metrics, tracing, symptom-based alerting, pre-launch gate
orchestration-patterns.mdEndorsed multi-persona orchestration patterns, anti-patterns, and the “personas don’t invoke personas” rule

How Skills Work

Every skill follows a consistent anatomy:

┌─────────────────────────────────────────────────┐
│  SKILL.md                                       │
│                                                 │
│  ┌─ Frontmatter ─────────────────────────────┐  │
│  │ name: lowercase-hyphen-name               │  │
│  │ description: Guides agents through [task].│  │
│  │              Use when…                    │  │
│  └───────────────────────────────────────────┘  │                                                                                                
│  Overview         → What this skill does        │
│  When to Use      → Triggering conditions       │
│  Process          → Step-by-step workflow       │
│  Rationalizations → Excuses + rebuttals         │
│  Red Flags        → Signs something's wrong     │
│  Verification     → Evidence requirements       │
└─────────────────────────────────────────────────┘

Key design choices:

  • Process, not prose. Skills are workflows agents follow, not reference docs they read. Each has steps, checkpoints, and exit criteria.
  • Anti-rationalization. Every skill includes a table of common excuses agents use to skip steps (e.g., “I’ll add tests later”) with documented counter-arguments.
  • Verification is non-negotiable. Every skill ends with evidence requirements - tests passing, build output, runtime data. “Seems right” is never sufficient.
  • Progressive disclosure. The SKILL.md is the entry point. Supporting references load only when needed, keeping token usage minimal.

Project Structure

agent-skills/
├── skills/                            # 25 skills (24 lifecycle + 1 meta)
│   ├── interview-me/                  #   Define
│   ├── idea-refine/                   #   Define
│   ├── spec-driven-development/       #   Define
│   ├── constraint-driven-development/ #   Define
│   ├── planning-and-task-breakdown/   #   Plan
│   ├── incremental-implementation/    #   Build
│   ├── context-engineering/           #   Build
│   ├── source-driven-development/     #   Build
│   ├── doubt-driven-development/      #   Build
│   ├── frontend-ui-engineering/       #   Build
│   ├── test-driven-development/       #   Build
│   ├── api-and-interface-design/      #   Build
│   ├── browser-testing-with-devtools/ #   Verify
│   ├── debugging-and-error-recovery/  #   Verify
│   ├── code-review-and-quality/       #   Review
│   ├── code-simplification/           #   Review
│   ├── security-and-hardening/        #   Review
│   ├── performance-optimization/      #   Review
│   ├── git-workflow-and-versioning/   #   Ship
│   ├── ci-cd-and-automation/          #   Ship
│   ├── deprecation-and-migration/     #   Ship
│   ├── documentation-and-adrs/        #   Ship
│   ├── observability-and-instrumentation/ # Ship
│   ├── shipping-and-launch/           #   Ship
│   └── using-agent-skills/            #   Meta: how to use this pack
├── agents/                            # 4 specialist personas
├── references/                        # 7 supplementary checklists
├── hooks/                             # Session lifecycle hooks
├── .claude/commands/                  # 8 slash commands (Claude Code)
├── .gemini/commands/                  # 8 slash commands (Gemini CLI)
├── commands/                          # 8 slash commands (Antigravity CLI)
├── plugin.json                        # Antigravity plugin manifest
└── docs/                              # Setup guides per tool

このREADMEは一部を省略しています。全文はGitHubにあります。 原文を読む ↗

Addy Osmani's Agent Skills
AIに聞く
GitHub