Skip to content

Navigation Menu

Sign in
Sign up

Are Agent Skills actually portable? — five-client evidence map #10

Discussion options

Five major AI coding clients now document some form of Agent Skills support. That does not prove that one skill behaves the same across them.

The new Doubt evidence map separates the portable core from the client-specific behavior:

  • 5 claims
  • 6 evidence items
  • 6 official sources
  • 1 contradiction
  • 1 explicit unknown

The practical verdict: author one conservative core, then install and test it per host. SKILL.md, references, assets, and scripts may be portable; discovery, activation, consent, tool use, and output behavior may not be.

Explore the evidence:

The map also exposed the next missing edge: no same-fixture behavioral benchmark across Claude Code, Codex, GitHub Copilot, Cursor, and Gemini CLI.

If you can run even one client, contribute a reproducible result to benchmark issue #9. Positive, neutral, and negative evidence are all useful.

You must be logged in to vote

Replies: 1 comment

Comment options

alsoleg89
Jul 31, 2026
Maintainer Author

The reproducible benchmark kit is now published in
Doubt v0.5.0.

It freezes one synthetic fixture, the unmodified skill payload, and direct,
implicit, and negative prompts. One-client contributions are accepted with
exact version/configuration metadata and sanitized raw output; honest failures
and blocked runs count as evidence.

Start with the
benchmark README
or claim a client in issue #9.

Current baseline: 0 submitted clients.

You must be logged in to vote
0 replies
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment
Labels
None yet
1 participant

AltStyle によって変換されたページ (->オリジナル) /