フリー問題

NVIDIA-Certified Professional: Agentic AI のフリー問題 12 / 20 問目

問題文

Before launch, a team wants to know how many concurrent users their agent can serve while keeping the ninety-fifth percentile latency under the objective. What kind of measurement answers this?

選択肢

  1. The token count of a typical request multiplied by the expected number of users.
  2. A single-request timing measurement, repeated one hundred times.
  3. A load test that drives increasing concurrency against the workflow and records latency and throughput at each level.
  4. A measurement of the model server's peak throughput in isolation, because the model call is the slowest part of the workflow and the workflow's own capacity is therefore determined entirely by the throughput that the server can sustain.

解答・解説を確認するには

正解と解説の確認、回答の記録には無料登録が必要です。登録すると演習モードでフリー問題に回答し、正誤と解説をその場で確認できます。