{"id":24663,"plugin_id":"plugins~Plugin_a2d7fcc77268819187a5d61e6a1452eb","kind":"skill","collection_source":null,"comparison_source":null,"observed_at":"2026-09-30T23:18:34.333Z","digest":"0cb4916db9daca36ce655fe37f87bfafa3d517c93bb537e6ef61266529f573ab","against":null,"payload":{"name":"metric-pack-designer","description":"Design custom metric packs for plugin-eval so teams can add local evaluation rubrics that emit schema-compatible checks and metrics. Use when the user wants their own evaluation criteria or visualizations.","included_files":[{"relative_path":"agents/openai.yaml","size_in_bytes":116}],"skill_md_contents":"---\nname: metric-pack-designer\ndescription: Design custom metric packs for plugin-eval so teams can add local evaluation rubrics that emit schema-compatible checks and metrics. Use when the user wants their own evaluation criteria or visualizations.\n---\n\n# Metric Pack Designer\n\nUse this skill when the user wants to extend `plugin-eval` with a local rubric.\n\n## Workflow\n\n1. Clarify the custom rubric categories and target kinds.\n2. Define the smallest useful `checks[]` and `metrics[]` payload.\n3. Create a metric-pack manifest plus a script that prints JSON to stdout.\n4. Run the pack through `plugin-eval analyze <path> --metric-pack <manifest.json>`.\n\n## Design Rules\n\n- Keep IDs stable across runs so comparisons stay meaningful.\n- Emit only `checks[]`, `metrics[]`, and optional `artifacts[]`.\n- Do not try to overwrite the core score or summary.\n- Prefer deterministic local signals over subjective text generation.\n\n## Reference\n\n- `../../references/metric-pack-manifest.md`\n"},"changes":[],"summary":"First saved snapshot. No earlier version is available for comparison.","summary_kind":"deterministic","summary_metadata":{}}