QuestionQ8

AI Operations

Which of the following components would MOST effectively address cumulative benefits as part of reinforcement learning (RL)?

Explanation

A value function estimates the expected cumulative return from a state or state-action pair, incorporating both immediate and future rewards. It therefore represents cumulative benefits in reinforcement learning.

Community Discussion

No comments yet. Be the first to start the discussion!