Question 1

A team notices vague, inconsistent LLM outputs for the same story for two different prompts. Which technique BEST helps choose the stronger wording among two prompt versions using predefined metrics?
  • Question 2

    An LLM prioritizes tests using likelihood X impact but ranks a trivial tooltip change above a payment failure.
    What defect does this MOST LIKELY show?
  • Question 3

    Which option BEST differentiates the three prompting techniques?
  • Question 4

    You are tasked with applying structured prompting to perform impact analysis on recent code changes. Which of the following improvements would BEST align the prompt with structured prompt engineering best practices for comprehensive impact analysis?
  • Question 5

    The model flags anomalies in logs and also proposes partitions for input validation tests. Which metrics BEST evaluate these two outcomes together?