フリー問題

NVIDIA-Certified Professional: Accelerated Data Science のフリー問題 7 / 20 問目

問題文

A nightly job reads 2 TB of raw event files, filters to one region, joins a small lookup table, aggregates by day, and writes the result. Which ordering of those steps moves the least data through the accelerated stack?

選択肢

  1. Join the lookup table first so that the region name is available as a readable label, then filter on that label, then aggregate, because filtering on a human-readable name is easier to review and the join itself is cheap when one side of it is small.
  2. Aggregate by day first to shrink the row count, then filter to the region, then join the lookup table, and the daily totals are far smaller than the raw events and every later step then works on that reduced table.
  3. Read everything, then let the query optimizer reorder the steps, and the optimizer sees the whole plan and will push the region filter down into the read for the team automatically.
  4. Filter to the region and select only the needed columns while reading, then join the lookup table, then aggregate.

解答・解説を確認するには

正解と解説の確認、回答の記録には無料登録が必要です。登録すると演習モードでフリー問題に回答し、正誤と解説をその場で確認できます。