AI for QA / Scenario
Judge Two Prompt Variants
Evaluate two prompts against an unattended-pipeline rubric.
- Difficulty
- Hard
- Format
- Scenario
- Points
- 200
- Estimate
- 15 min
// MISSION BRIEF
Your Mission
Two prompt variants are up for adoption in a test-generation pipeline. Score them against a rubric that values correctness, coverage, and, crucially, repeatability. The prettier output is not automatically the better prompt.
// FIRST CONTACT
Battle teaser
First artifact
The rubric (pipeline use, runs unattended nightly)
Against the rubric, the better prompt for this pipeline is:
- AVariant B: creativity and volume matter most
- BA tie; both produce test cases
- CNeither; prompts cannot be compared
- DVariant A: it parses every run, holds format, scopes to the spec, and surfaces ambiguity instead of inventing
Answers, scoring, hints, and the full battle stay sealed.
// SKILL TAGS
prompt-engineeringevaluationscenarioai-for-qa