AI for QA / Scenario

Judge Two Prompt Variants

Evaluate two prompts against an unattended-pipeline rubric.

Difficulty
Hard
Format
Scenario
Points
200
Estimate
15 min

// MISSION BRIEF

Your Mission

Two prompt variants are up for adoption in a test-generation pipeline. Score them against a rubric that values correctness, coverage, and, crucially, repeatability. The prettier output is not automatically the better prompt.

// FIRST CONTACT

Battle teaser

First artifact

The rubric (pipeline use, runs unattended nightly)

Against the rubric, the better prompt for this pipeline is:

  1. AVariant B: creativity and volume matter most
  2. BA tie; both produce test cases
  3. CNeither; prompts cannot be compared
  4. DVariant A: it parses every run, holds format, scopes to the spec, and surfaces ambiguity instead of inventing
Answers, scoring, hints, and the full battle stay sealed.

// SKILL TAGS

prompt-engineeringevaluationscenarioai-for-qa