
What Are Agent Skills? How They Differ from Prompts
Understand when a reusable Agent Skill is worth building, what belongs in SKILL.md, and when a normal prompt is enough.

Evaluate trigger precision, output quality, failure behavior, context cost, and consistency with a compact test set.
A Skill is not reliable because one demo looked good. Test the conditions that are most likely to expose ambiguity.
Use one request that should trigger the Skill, one nearby request that should not, and one ambiguous request that requires clarification.
Check required sections, file type, naming, citations, and validation results. Do not grade only tone or visual polish.
Remove one required input. A reliable Skill should request the missing value or follow a documented safe fallback—not invent it.
Simulate unavailable tools, unreachable sources, and invalid files. The Skill should report the limitation and preserve partial work.
Record which references are loaded by default and how often the Skill needs manual correction. A smaller predictable Skill often outperforms a broad one.

Understand when a reusable Agent Skill is worth building, what belongs in SKILL.md, and when a normal prompt is enough.

A practical structure for writing small, predictable Agent Skills with clear triggers, resources, and evaluation cases.

Start with a task that has clear rules, stable inputs and a result you can inspect. A first automation could organize a test enquiry into a spreadsheet without automatically replying to the customer.
One useful workflow at a time
Get new Agent Skills, prompt patterns, and honest implementation notes. No fabricated metrics, no daily noise.