フリー問題

NVIDIA-Certified Associate: Generative AI LLMs のフリー問題 14 / 20 問目

問題文

Two prompt templates are being compared on the same 200 questions. The team runs each template once with sampling enabled and different random seeds, then declares a winner from the single pair of scores. What is the flaw?

選択肢

  1. With sampling enabled, one run per template mixes the effect of the template with the randomness of decoding, so the difference may not be reproducible.
  2. The scores cannot be compared, and the two templates produce answers of different lengths, and length alone decides the score.
  3. Two hundred questions is too few for any comparison, and no design can produce a usable answer until the evaluation set holds several thousand items.
  4. Comparing templates requires that both be evaluated on different question sets so that neither template can benefit from questions it has already been tuned against.

解答・解説を確認するには

正解と解説の確認、回答の記録には無料登録が必要です。登録すると演習モードでフリー問題に回答し、正誤と解説をその場で確認できます。