フリー問題

NVIDIA-Certified Associate: Generative AI LLMs のフリー問題 18 / 20 問目

問題文

A team optimizes an assistant purely for the automatic helpfulness score. Six months later, answer length has doubled and users complain about wordiness. What general lesson does this illustrate?

選択肢

  1. Answer length should have been capped from the start, which is the only measure that prevents this specific outcome.
  2. Automatic helpfulness scores are invalid, and the only defensible approach is to optimize against human judgment collected on every release rather than against any automatic measure.
  3. The assistant has overfitted the evaluation set, and rotating the evaluation items each month would have prevented the drift in answer length from occurring at all.
  4. Optimizing a single measure lets everything it does not capture drift, so a set of guardrail metrics has to be tracked alongside the target.

解答・解説を確認するには

正解と解説の確認、回答の記録には無料登録が必要です。登録すると演習モードでフリー問題に回答し、正誤と解説をその場で確認できます。