フリー問題

Databricks Certified Associate Developer for Apache Spark のフリー問題 10 / 20 問目

問題文

A large fact DataFrame must be enriched with attributes from a small reference DataFrame, and every fact row must survive even when no reference row matches. Which join type meets that requirement in PySpark?

選択肢

  1. A left semi join, which keeps the fact rows and adds the reference attributes.
  2. An inner join, because the reference DataFrame is a complete dimension by construction and therefore every fact row is guaranteed to find a match, which makes the outer variant an unnecessary cost.
  3. A full outer join, which is the only way to keep unmatched rows on either side, so unmatched reference rows appear as well and have to be filtered out afterwards.
  4. A left outer join with the fact DataFrame on the left, which keeps unmatched fact rows with null attributes.

解答・解説を確認するには

正解と解説の確認、回答の記録には無料登録が必要です。登録すると演習モードでフリー問題に回答し、正誤と解説をその場で確認できます。