agency-evaluation-criteria

Evaluate AI agency deliverables against quality criteria and generate evaluation-report.md.

Updated Mar 25, 2026
One-click install
npx skills add https://github.com/taewook486/anki_rag --skill agency-evaluation-criteria
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: agency-evaluation-criteria
Source: https://github.com/taewook486/anki_rag/tree/main/.claude/skills/agency-evaluation-criteria
Command: npx skills add https://github.com/taewook486/anki_rag --skill agency-evaluation-criteria

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Provides a standardized, skeptical evaluation framework and automated testing to ensure AI agency deliverables meet quality and compliance expectations.

Core Features & Use Cases

  • Weighted scoring across Design Quality, Originality, Completeness, and Functionality with explicit pass/fail criteria.
  • Automated testing via Playwright to verify deliverables against BRIEF, copy, and design-spec requirements.
  • Output generation of an evaluation-report.md containing overall scores, per-dimension evidence, defect lists with references, screenshots, and actionable improvements.
  • Applicable to agency projects including copywriting, UX/UI design, and implementation artifacts to enforce quality gates.

Quick Start

Provide your BRIEF, Original copy, and Original design-spec to generate a detailed evaluation report.

Frequently Asked Questions about agency-evaluation-criteria

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I evaluate AI agency project deliverables against quality criteria?▼

Evaluate AI agency deliverables by scoring them against weighted criteria like Design Quality, Originality, Completeness, and Functionality. This process uses input contracts such as BRIEF, Original copy, and design-spec files to generate a detailed evaluation report.

Can I use Playwright to automate quality testing for design and copy artifacts?▼

Yes, Playwright automates testing to verify deliverables against BRIEF, copy, and design-spec requirements. It captures screenshots and validates functionality to ensure AI agency outputs meet explicit pass/fail criteria.

What is included in an automated evaluation report for agency outputs?▼

An evaluation report includes an overall score, per-dimension scores with evidence, a defect list with file references and screenshots, and actionable improvement recommendations based on the input contracts.

How do I audit AI agency outputs for originality and completeness?▼

Audit AI agency outputs by applying a standardized evaluation framework that scores originality and completeness against the original copy and design specifications, producing a defect list with file references for any missing elements.

What inputs do I need to generate a quality evaluation report for an agency project?▼

You need to provide the BRIEF, Original copy.md, Original design-spec.md, and the built application artifacts. These inputs are used to evaluate the deliverables and generate the final evaluation-report.md.

Does the evaluation framework support different types of agency deliverables?▼

Yes, the framework is applicable to agency projects including copywriting, UX/UI design, and implementation artifacts, enforcing quality gates across various deliverable types using a weighted scoring system.