One-entry orchestrator for coding agents: jailed vault, FTS knowledge graph, 18-axis review, in-harness Chrome QA. Curated packs — not a 90k-skill dump
SuperSkill has 23 tools with schemas and descriptions present, but quality is uneven. Most tools have basic descriptions (50-150 chars) and reasonable input schemas with type definitions. However, several critical gaps exist: (1) missing output schema documentation across all tools, (2) sparse parameter descriptions (many params lack detail about constraints, formats, or ranges), (3) several tools lack clarity on when to use them vs. similar tools (task, learn, brainstorm have overlapping responsibilities), (4) no documented error recovery guidance, (5) sensitive operations (write, delete) lack confirmation/dry-run patterns. Naming is generally verb-based (read, write, decide, task) but some names are overloaded (task action='list|create|update|close' should be split). Tool definitions are explicitly visible in src/mcp-server.ts and registered with schemas, so no scoring caps apply for inference. Average per-tool score across 23 tools is ~52, placing this in the 'Fair/Poor' range, significant gaps in descriptions and error handling, but base structure is present.
Log brainstorm entries with topic organization
Capture session output and artifacts
Reference credentials/secrets without exposing values
Log an architecture decision record
Mark notes as deprecated with migration guidance
Log environment facts and system configuration
Extract structured data from unstructured notes using AI
Output schemas not documented. No visible response shape definitions for any of the 23 tools. LLMs cannot plan downstream calls or extract fields without knowing what structure to expect.
Sparse parameter descriptions. Many parameters have only a brief noun phrase ('File path to read', 'Task title') without constraints, format, range, or dependency hints. Example: 'depth' in read tool lacks guidance on valid range or behavior for deeply nested dirs.
| Scored | Grade | Overall | Spec posture | Rubric |
|---|---|---|---|---|
| 2026-09-22 | D | 52 | 2026-07-28+ | v2 |
Initialize superskill vault for a project or repo
Manage learning notes (discoveries and insights)
Create knowledge graph links between notes
Get comprehensive project context including decisions, learnings, and recent work
Clean vault by removing stale notes based on retention policy
Read a vault note
Get resume context for interrupted sessions with formatting support
Restore repository state from snapshot
Full-text search vault knowledge base with FTS5
Manage agent sessions for multi-turn coordination
Install skills from sources (registry, GitHub, local)
Remove installed skills
Capture current repository state snapshot for rollback
Vault statistics: counts, ages, trending topics
Manage project tasks/todos
Write/create a vault note
Overloaded 'action' parameter in task, learn, session, and prune. Single parameter accepting enum of 4+ actions (list|create|update|close) signals multiple responsibilities. Should split into separate tools or clearly document which params apply per action to avoid LLM confusion.
No error recovery guidance. Tools provide no hints for what LLM should do on failure. Example: write tool with no guidance if path is invalid or permissions denied. Pattern: recovery-guide.
Destructive operations lack confirmation/dry-run pattern. Tools like prune, rollback, skill_remove perform irreversible state changes with no explicit confirmation step or dry-run mode. Agents can accidentally delete significant vault data.
Sensitive operation (cred_refs) documentation is vague. Tool claims to 'reference credentials/secrets without exposing values' but no guidance on how the secrets are actually injected, retrieved, or validated at runtime. Security posture unclear.
Rate limiting implemented (30 writes/min) but no guidance in tool descriptions about backoff or retry behavior. If an LLM hits the limit, it has no signal to slow down or batch differently.
Overlapping tool purposes. task, learn, and brainstorm all log/store structured notes with different metadata. No clear guidance on when to use task vs learn vs brainstorm. LLM may pick wrong tool or duplicate entries.
Pagination not documented for search and stats tools. If search can return large result sets, no limit, page, or offset params visible. Risk of context window exhaustion.
Parameter naming inconsistency. Some tools use 'project' (string slug), others infer it. No consistent pattern for optional vs required project context across the 23 tools.