Two responses have different strengths and neither clearly wins. How would you apply the ranking rules?

Instruction: Rate both fictional replies on grounding, coverage and the defined word limit. Apply the stated ranking order, give actual ratings and submit a preference rationale of no more than two sentences.

Context:

Compare two grounded replies with competing coverage and concision strengths. Record dimension ratings, apply an explicit ranking order and adapt when that order changes.

Fictional practice task

This is a synthetic conversation, not a live tool call. Rate the two proposed final replies under guideline version 3.

Conversation:

  • User: “Please check my exchange for order R17.”
  • Assistant: “I’ll look up the exchange record.”
  • Approved tool result, exchange_lookup: {"order_id": "R17", "exchange_status": "approved", "next_step": "print the return label", "return_address": "Depot 4, Cedar Lane"}
  • User: “Tell me whether it is approved, the next step and the return address in one brief reply.”

Response A: “Thanks for checking. I can confirm that the exchange for your order is approved. Your next step is to print the return label. The return address is Depot 4, Cedar Lane.”

Response B: “Approved. Next step: print the return label.”

Record three dimensions for each response:

  • grounding: pass if every exchange-related factual assertion is supported by the tool result; otherwise fail. Greetings do not add an exchange fact.
  • coverage: complete if the reply supplies all three requested details; otherwise incomplete.
  • concision: within_limit at 25 words or fewer; otherwise over_limit. Count whitespace-separated tokens in the response text; punctuation does not create a separate token.

Ranking rule: A grounding pass outranks a fail. If grounding is equal, prefer complete coverage; if coverage is also equal, prefer within-limit concision. If all three recorded dimensions are equal, use tie. Coverage and concision are ranking dimensions here, not independent acceptance gates. Do not invent numeric weights or rank by response position.

Produce: Each response’s three ratings and word count, a preference of A, B or tie, and a submitted rationale of no more than two sentences. Use only the supplied conversation and tool result.

Updated

Prepare a stronger answer

I’d give both replies a grounding pass because their exchange facts match the tool result. A has complete coverage: it gives the approval status, next step and return address...

This member answer includes:

  • • A complete, copyable sample answer
  • • Guidance for adapting the answer to your experience
  • • A practical walkthrough
  • • Common mistakes and how to avoid them
  • • Answered interviewer follow-ups
  • • Strong, adequate and weak assessment criteria
Unlock the full answer and preparation guide

One payment for one year of full access. No automatic renewal.

See pricing and everything included

Your preparation path

Choose the track that matches the role. Work through its questions in order, then explain each answer in your own words.

Try the 20 minute mock assessment. Use the fictional cases to practice; the self-check is not an employer's hiring benchmark.

Related Questions