フリー問題

NVIDIA-Certified Professional: Generative AI LLMs のフリー問題 7 / 20 問目

問題文

A model must fit into a smaller memory budget and also run faster on hardware that accelerates dense matrix multiplication. Which pruning approach matches both goals?

選択肢

  1. Structured pruning that removes whole units such as attention heads or feed-forward channels.
  2. Reducing the batch size at inference time.
  3. Unstructured magnitude pruning applied uniformly across all layers, which removes the least important individual weights and gives the best accuracy for a given reduction in the total number of nonzero parameters.
  4. Quantization of the weights to a lower-precision integer format, which shrinks the matrices the hardware multiplies.

解答・解説を確認するには

正解と解説の確認、回答の記録には無料登録が必要です。登録すると演習モードでフリー問題に回答し、正誤と解説をその場で確認できます。