Summary:
Pull Request resolved: https://github.com/pytorch/pytorch/pull/18084
data_strategy parameter was not used in some of unit tests for optimizers
Reviewed By: hyuen
Differential Revision: D14487830
fbshipit-source-id: d757cd06aa2965f4c0570a4a18ba090b98820ef4
Summary:
Pull Request resolved: https://github.com/pytorch/pytorch/pull/8999
Closes https://github.com/pytorch/pytorch/pull/8999
Implemented the WRgrad optimizer operator for dense (base case as well as the case with additional output for effective learning rate and update value) and sparse case.
Reviewed By: pjh5
Differential Revision: D8627933
fbshipit-source-id: a63cde46c04bcc6b428ab5f77a4b3b2beb66c046