Topic hub

Agent skills

Portable skill files, behavioral contracts, tool instructions, and evals that make agent capability reusable.

Retrieval answer

Portable skill files, behavioral contracts, tool instructions, and evals that make agent capability reusable. Agent skills package repeatable behavior outside the current chat. The important boundary is whether the skill changes task outcomes, not whether the instruction prose looks good. Skills become infrastructure when they are versioned, evaluated, and shared across agents.

Pattern memory

What patterns are emerging?

1 patterns
  1. medium

    Skills become a portable capability layer

    Agent skills are emerging as a portable capability layer, but their value depends on progressive disclosure, provenance, security review, and behavioral evaluation.

Field notes

What should readers understand next?

6 notes
  1. Dr. Skill Audits What An Agent Loads Before It Works

    Dr. Skill scans skills and MCP servers for collisions, duplication, secrets, drift, missing metadata, and unused loadout, with local and CI-friendly commands.

  2. Genkit Adds Progressive Disclosure For Agent Skills

    Genkit now loads Agent Skills through middleware that discovers SKILL.md metadata first and activates full instructions, references, and scripts only when needed.

  3. Mistral Treats Prompts And Skills As Production Records

    Mistral Studio adds immutable versions, ownership, promotion labels, lineage, rollback, and audit logs for prompts and skills used in production AI systems.

  4. Stripe Builds A Shared Agent Platform For Knowledge Work

    Stripe's Kai combines surface-agnostic APIs, domain-owned AgentStudio assets, per-session sandboxes, long-horizon state, and more than 1,000 internal tools and skills.

  5. Claude Cowork Turns Screen Recordings Into Reusable Skills

    Claude Cowork's recorded-skill flow makes workflow capture a first-party path from human demonstration to reusable agent capability.

  6. Hermes Agent Makes the Learning Loop Part of the Runtime

    Nous Research's Hermes Agent packages memory, skills, messaging, scheduling, tool use, and sandboxing as a runtime rather than a single chat surface.

Raw signals

What changed recently?

83 signals
  1. Agent skills need behavioral evals, not prose review

    Benchmarks show that expert-authored skills can help while self-generated skills can underperform a no-skill baseline.

  2. Improve routes architecture and execution to different models

    The Improve skill audits a repository with a stronger planning model and delegates bounded fixes to cheaper executors.

  3. Large-task planning moves from chat into a ticket graph

    Matt Pocock's workflow turns discovery, specification, tickets, implementation, and review into durable linked artifacts.

  4. Animation Vocabulary Turns Motion References into an Agent Skill

    A compact design skill maps visual intent to named motion patterns, giving interface agents a reusable language for implementing animation decisions.

  5. A GitHub Skill Can Resolve Review Comments as a Batch

    A reusable Codex workflow gathers pull-request feedback, plans changes, applies fixes, and reports resolution state as one auditable review task.

  6. Pydantic AI loads capabilities only when needed

    A capability can bundle instructions, tools, model settings, and hooks while exposing only a compact description until activation.

  7. Creative Ideation Becomes a Routable Agent Skill

    A packaged workflow decomposes creative exploration into reference gathering, divergence, critique, and selection instead of asking a model to be creative in one step.

  8. Anthropic: Goal-Scoped Agent Loops

    The archive captures X source as a dated public record from X source. It documents long-running work gaining explicit goals, state, stopping rules, and recovery and is retained as pressure-testing evidence for the goal-scoped agent loops trend.

  9. Agentic Resource Discovery Indexes Skills, MCP, and Agents

    Hugging Face treats agent capabilities as searchable resources, creating a discovery layer for tools that can be loaded at runtime.

  10. Hermes Turns Solved Tasks into New Skills

    An agent runtime can preserve successful procedures as reusable capability modules, converting execution traces into a compounding operational library.

  11. CLAUDE.md Shrinks as Skills, Hooks, and Subagents Mature

    Agent behavior moves out of one giant context file into scoped mechanisms that load only for the relevant task and lifecycle stage.

  12. Codex Turns a Recorded Mac Workflow into a Reusable Skill

    Record and Replay captures a human procedure, converts it into agent instructions, and preserves the successful interaction as reusable behavior.

  13. Taste Lab Uses a Design System as a Reasoning Brief

    Taste Lab packages aesthetic decisions and implementation rules as agent-readable context rather than relying on vague style prompts.

  14. Last30Days Packages Social Research as an Agent Skill

    A reusable research workflow gathers recent public discussion, filters repetition, and produces a time-bounded evidence set for another agent or analyst.

  15. Claude Code: Verification Bandwidth

    The archive captures Claude Code as a dated public record from Claude Code. It documents evaluation, review, and observability becoming the bottleneck after generation accelerates and is retained as pressure-testing evidence for the verification bandwidth trend.

  16. GitHub / cursor/plugins: Verification Bandwidth

    The archive captures GitHub / cursor/plugins as a dated public record from GitHub / cursor/plugins. It documents evaluation, review, and observability becoming the bottleneck after generation accelerates and is retained as pressure-testing evidence for the verification bandwidth trend.

  17. SkillSpector Adds a Security Gate for Agent Skills

    NVIDIA's scanner treats installable agent instructions as executable supply-chain artifacts that require inspection before use.

  18. Walkinglabs / Learn Harness Engineering: Verification Bandwidth

    The archive captures Walkinglabs / Learn Harness Engineering as a dated public record from Walkinglabs / Learn Harness Engineering. It documents evaluation, review, and observability becoming the bottleneck after generation accelerates and is retained as pressure-testing evidence for the verification bandwidth trend.

  19. Claude for Legal Packages a Legal Team as Plugins

    Anthropic's legal repository expresses domain knowledge, procedures, and tools as installable modules for repeatable professional work.

  20. Claude: Goal-Scoped Agent Loops

    The archive captures YouTube source as a dated public record from YouTube source. It documents long-running work gaining explicit goals, state, stopping rules, and recovery and is retained as supporting evidence for the goal-scoped agent loops trend.

  21. SKILL.md Acts as a Behavior Loader, Not a Better Prompt

    Skills package procedures, tools, and progressive context so an agent can load behavior only when the task requires it.

  22. Claude: Goal-Scoped Agent Loops

    The archive captures X source as a dated public record from X source. It documents long-running work gaining explicit goals, state, stopping rules, and recovery and is retained as pressure-testing evidence for the goal-scoped agent loops trend.

  23. Cursor Ships Its Development Procedures as Agent Skills

    Cursor Team Kit packages verification, CI repair, review preparation, and code cleanup as reusable agent behavior.

  24. GitHub / NousResearch/hermes-agent: Generative Media Infrastructure

    The archive captures GitHub / NousResearch/hermes-agent as a dated public record from GitHub / NousResearch/hermes-agent. It documents image, video, audio, and multimodal generation becoming application infrastructure and is retained as branch-opening evidence for the generative media infrastructure trend.

Discovery graph / next reads

Continue through New Runtime

Open the graph
  1. 01related materialSkills become a portable capability layerContinue through the Agent skills topic.
  2. 02related materialDr. Skill Audits What An Agent Loads Before It WorksContinue through the Agent skills topic.
  3. 03related materialGenkit Adds Progressive Disclosure For Agent SkillsContinue through the Agent skills topic.
  4. 04related materialMistral Treats Prompts And Skills As Production RecordsContinue through the Agent skills topic.
  5. 05related materialStripe Builds A Shared Agent Platform For Knowledge WorkContinue through the Agent skills topic.

These links are also published in this page’s JSON twin and as typed edges in DiscoveryGraph v1.

Who read this page?Machine requests, hidden until opened

Loading the privacy-safe route aggregate…

Open the JSON contract