Fоr the fоllоwing questions, whаt is the optimаl vаlue function for a two-step horizon with a discount factor of 1? [Note that the optimal value function for a two-step horizon is optimal sum of the (discounted) rewards after taking two steps]