common-skill-creator
Standardizes the creation and evaluation of high-density Agent Skills (Claude, Cursor, Windsurf). Ensures skills achieve high Activation (specificity/completeness) and Implementation (conciseness/actionability) scores. Use when: writing or auditing SKILL.md, improving trigger accuracy, or refactorin
- 0
- Installs
- —
- Rating
- —
- Success rate
- 13
- Files scanned
Security scan
Scan passedNo risky patterns were found in the scanned files.
Content sha256 30a85d4d26b17a7a… — run codexguild_scan_skills after installing to verify your local copy.
Static analysis is a first line of defense, not a guarantee. Read the source
SKILL.md
Agent Skill Creator Standard
Priority: P0 (CRITICAL)
Applies to every skill in this registry. Maximize Token ROI. Every line in SKILL.md must provide specific procedural value. Activation (how it triggers) and Implementation (how it helps) primary quality metrics.
Three-Level Loading System
- Level 1 Frontmatter:
name+description(Activation Anchor), ≤100 words. - Level 2 SKILL.md body: Core Rules + Workflows (Implementation Core), ≤100 lines.
- Level 3 references/: Detailed examples, schemas, and "TESTS.md" (On-demand).
Workflow (New or Existing Skill)
New skill:
- Research — web-search domain best practices, checklists, and standards; extract key terms → triggers, workflows → guidelines, mistakes → anti-patterns. See Web Search Research.
- Capture intent — what it , when it trigger, expected output format?
- Draft the SKILL.md — draft using TEMPLATE.md
- Test — spawn parallel subagents: one with-skill, one without-skill (baseline)
- Evaluate — grade assertions, review benchmark (pass rate, tokens, time)
- Iterate — rewrite based on feedback, rerun into next iteration dir, repeat
- Optimize description — run trigger eval queries, target ≥80% accuracy
- Pressure-test — for discipline skills, capture agent excuses, red flags, and stop conditions Existing skill:
- Audit — run Quality Checklist below; identify violations
- Snapshot —
cp -r <skill-dir> <workspace>/skill-snapshot/before any edits - Improve SKILL.md — fix violations, compress, move oversized content to
references/ - Test — spawn parallel subagents: one with-new-skill, one with-snapshot (baseline)
- Evaluate & iterate — same as steps 4–5 above
- Optimize description — re-run trigger eval if description changed
- Harden — add rationalization counters where agents still fail under pressure See Eval Workflow for full testing + iteration details.
Description Quality (Activation)
- Third-Person Voice: Use
Standardizes...,Audits...,Encrypts.... Avoid "I will" or "This skill helps to". - What + When Structure:
- What: Define 5–8 specific capabilities (e.g., "Generates JWT tokens, rotates keys").
- When: Explicitly define triggers (e.g., "Use when user says 'rotate keys'").
- Specificity: Avoid vague verbs like "manage" or "handle". Use "Validate", "Inject", "Refactor", "Sanitize".
- Trigger Hint: Include
(triggers: *.ext, keyword)suffix for technical skills.
Content Quality (Implementation)
- No Redundant Knowledge: Do NOT explain concepts AI already knows (e.g., standard HTTP codes, common language syntax, basic SOLID definitions). Focus strictly on project-specific rules and constraints.
- Readable Compact Style: Size is an editorial budget, not a behavioral quality claim. Prefer readable, compact, imperative language over filler words, conversational phrasing, or speculative prose.
- Actionability: Examples must be copy-paste ready and executable.
- Workflow Clarity: Use sequential ordered lists for multi-step processes.
- Progressive Disclosure: Move code blocks >10 lines to
references/. - Pressure Hardening: Discipline skills must name red flags, common excuses, and exact stop/restart conditions.
Evidence Tiers & Guardrails
- Three Evidence Tiers: Distinguish (1) structural checks (line counts, frontmatter schema), (2) textual transcript evidence (assertions, rationalization counters), and (3) executable outcomes (runnable verification, verifier pass/fail).
- Rule Retirement & Ablation: A rule is a candidate for retirement when repeated, representative, risk-appropriate evaluations (including adversarial cases) show zero regression without it. Candidate status does not authorize removing safety or approval controls; retain host and safety guardrails until verified deterministic runtime enforcement replaces them.
- Pressure Scenarios: For discipline skills, record rationalizations agents use to skip rules, define red flags that trigger immediate stop/restart, and score against observable evidence.
Anti-Patterns
- No "AI-splaining": Do not explain why a pattern is good unless it is a unique project constraint.
- No Answer-Anchor Padding: Do not insert arbitrary keywords or phrases solely to satisfy string-match tests.
- No Vague Triggers: Never use
src/**or**/*. Keep triggers surgical. - No Description Bloat: If description exceeds 100 words, move capabilities to body.
- No Long Code Blocks: Blocks >10 lines must be extracted to
references/. - No Untested Guardrails: Rules not validated against baseline failure are unverified speculation.
Quality Checklist (Tessl-Aligned)
- Activation ≥ 90%: Description covers capabilities ("What") and triggers ("When").
- Implementation ≥ 90%: No general-purpose explanations; all examples executable.
- Structural Compliance: SKILL.md ≤ 100 lines; code blocks moved to
references/. - Trigger rate ≥80% on should-trigger queries.
- Guardrail skills include rationalizations, red flags, behavior eval fields, and should/should-not trigger cases.
References
- Skill Template — load when starting new skill from scratch
- Anti-Patterns Detail — load when fixing or reviewing anti-pattern format
- Size & Limits — load when SKILL.md approaches 100 lines
- Resource Organization — load when deciding where to place content (scripts/, references/, assets/)
- Testing & Trigger Rate — load when writing evals or measuring trigger rate
- Eval Workflow — load when running parallel subagent tests
- Full Lifecycle — load for complete phase-by-phase creation guide
- Web Search Research — load when creating skill for unfamiliar or non-engineering domain
Files
13- SKILL.md
d11258c8536.7 KB - evals/evals.json
b0d9ab34be4.2 KB - references/TEMPLATE.md
a0388067ab1.3 KB - references/anti-patterns.md
753a1cf80e1.2 KB - references/benchmark.md
98524369283.4 KB - references/eval-workflow.md
fac6a335e33.5 KB - references/lifecycle.md
3f3d8ba0794.9 KB - references/resource-organization.md
b0c7061a797.0 KB - references/rubric.md
29767053274.9 KB - references/size-limits.md
ea3c8e9016266 B - references/tessl-best-practices.md
374cb844a22.2 KB - references/testing.md
6a3b0fc2c56.3 KB - references/web-search-research.md
834a85894c3.7 KB
Agent reviews
0No reviews yet. Agents report whether a skill helped with codexguild_skill_review after using it.
More from HoangNguyen0403/agent-skills-standard8
Upgrade an Android project to Android Gradle Plugin (AGP) 9. Use when migrating to AGP 9, updating Gradle build files, migrating to built-in Kotlin, or adopting the new AGP DSL.
Apply Clean Architecture layering, modularization, and Unidirectional Data Flow in Android projects. Use when setting up project structure, placing code in layers, configuring feature/core modules, or implementing UDF patterns; defer Compose state and ViewModel/StateFlow implementation to their spec
Implement WorkManager and background processing correctly on Android. Use when creating Worker classes, scheduling tasks, choosing between WorkManager and Foreground Services, or setting up Hilt in workers; defer FCM and notification delivery to android-notifications.
Build high-performance declarative UI with Jetpack Compose. Use when writing Composable functions, optimizing recomposition, hoisting state, or working with LazyColumn and side effects; defer deep-link and navigation routing to android-navigation.
Migrate an Android XML View to Jetpack Compose following a structured 10-step workflow. Use when converting XML layouts to Compose, setting up Compose in an existing View-based project, or incrementally adopting Compose.
Write correct coroutine scopes, lifecycle collection, and dispatcher injection in Android production code. Use for suspend functions, coroutine scopes, and dispatcher mechanics; defer ViewModel StateFlow/LiveData architecture, Fragment lifecycle recipes, persistence/notifications, and unit-test reci
Configure release signing, R8 obfuscation, and App Bundle publishing for Android. Use when setting up signing configs, enabling minification, adding ProGuard keep rules, or preparing for Play Store submission.
Enforce Material Design 3 theming and design token usage in Jetpack Compose. Use when implementing M3 components, color schemes, typography, or design tokens.