A one-parameter model is trained on the loss
using the gradient descent update rule
Training starts at with learning rate . Each step uses the slope measured at the position the algorithm is standing on at that moment.
Exactly two steps are taken. What is the value of afterwards?
Round your answer to 2 decimal places.