Question 5
The model trains almost exclusively on the largest dataset (highest resource task).
The sampling distribution approaches a uniform distribution, where all tasks (large and small) are sampled with nearly equal probability.
The model trains almost exclusively on the smallest dataset (lowest resource task).
The sampling distribution remains proportional to the original dataset sizes.