Skip to content
AD-802 (B) · Reinforcement Learning/Important Questions

Reinforcement Learning (AD-802 (B)) - Important Questions

  1. Unit 17 Marks High Priority

    Define reinforcement learning and give a real-life example.

    Predicted for DEC-2026

  2. Unit 17 Marks High Priority

    Explain various approaches to implement reinforcement learning.

    Predicted for DEC-2026

  3. Unit 17 Marks High Priority

    Differentiate between reinforcement learning and supervised learning.

    Predicted for DEC-2026

  4. Unit 27 Marks High Priority

    What is the key objective of bandit algorithms in reinforcement learning?

    Predicted for DEC-2026

  5. Unit 27 Marks High Priority

    After 12 iterations of the UCB1 algorithm applied on a 4-arm bandit problem, we have n1=3, n2=4, n3=3, n4=2 and Q12(1)=0.55, Q12(2)=0.63, Q12(3)=0.61, Q12(4)=0.40. Which arm should UCB1 play next? Justify with calculation.

    Predicted for DEC-2026

  6. Unit 27 Marks High Priority

    What is meant by passive and active reinforcement learning and how do we compare the two?

    Predicted for DEC-2026

  7. Unit 214 Marks High Priority

    Explain the following terms with respect to bandit algorithms: Median Elimination, Upper Confidence Bound (UCB) algorithm, and Probably Approximately Correct (PAC).

    Predicted for DEC-2026

  8. Unit 37 Marks High Priority

    Explain the Q-Function and Q-Learning Algorithm.

    Predicted for DEC-2026

  9. Unit 37 Marks High Priority

    Discuss Maximum Likelihood and Least Square Error Hypothesis.

    Predicted for DEC-2026

  10. Unit 47 Marks High Priority

    Explain Fitted-Q and Deep Q-Learning Problems.

    Predicted for DEC-2026

  11. Unit 47 Marks High Priority

    Explain any one advanced Q-learning algorithm.

    Predicted for DEC-2026

  12. Unit 47 Marks High Priority

    Explain learning policies by imitating optimal controllers.

    Predicted for DEC-2026

  13. Unit 57 Marks High Priority

    Explain inverse reinforcement learning.

    Predicted for DEC-2026

  14. Unit 414 Marks High Priority

    Explain the following: (i) Deep Q-Network (DQN) and Policy Gradient (ii) Applications and issues in deep reinforcement learning.

    Predicted for DEC-2026

Go to where you left off?

Quick Add to Notes

Save questions, your own notes and screenshots into notes filed by unit. It takes a free account.

Create free account

Have an account? Log in