Question 2
Computing the PPO objective requires solving a constrained optimisation subproblem at each step, making it more expensive than TRPO.
Question 3
In Proximal Policy Optimization (PPO), the clipped surrogate objective for a single state-action pair is:
Which of the following statements are correct?
16 more questions in this paper
Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.
More on the Mathematical Foundations of Generative AI End Term 13 Sept 2026 paper
The IIT Madras BS Mathematical Foundations of Generative AI (Mathematical Foundations of Generative AI) End Term paper sat on 13 Sept 2026, in the May 2026 term: 19 questions for 50 marks in 180 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.
| Feature | Mathematical Foundations of Generative AI End Term 13 Sept 2026 at a glance |
|---|---|
| Term | May 2026 term |
| Subject | Mathematical Foundations of Generative AI |
| Questions | 19 |
| Marks | 50 |
| Duration | 180 min |
| Numerical | 7 |
| MSQ | 8 |
| MCQ | 4 |
| Official paper | Mathematical Foundations Of Generative Ai 13 Sep 26 |
| Negative marking | No negative marking. |
| Updated |