Gradient descent is run on the quadratic loss
using the standard update
starting from with a fixed learning rate .
Compute , and determine the full range of for which this iteration converges to the minimiser from any starting point.
Select all that apply.