About gradient balancing #23

miranghimire · 2021-12-23T09:42:58Z

There are three backward calls inside gardient balancing between generator loss & OCR loss:

convolutional-handwriting-gan/models/ScrabbleGAN_baseModel.py

Line 354 in f7daa50

self.loss_T.backward(retain_graph=True)
convolutional-handwriting-gan/models/ScrabbleGAN_baseModel.py

Line 368 in f7daa50

self.loss_T.backward(retain_graph=True)
convolutional-handwriting-gan/models/ScrabbleGAN_baseModel.py

Line 374 in f7daa50

self.loss_T.backward()

Won't these calls accumulate the gradients during the call of optimizer.step(); I thought our objective here was to simply compute the gardient balancing terms and multiply those to the loss or could you please give overview of what's going on here inside gradient balancing incase I misunderstood something?

sharonFogel · 2021-12-26T11:13:22Z

I just looked at the code, I think you're right and there should be self.netG.zero_grad() between the first and second backprop. The third one is performed without gradient accumulation just so that the graph won't be retained.

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

About gradient balancing #23

About gradient balancing #23

miranghimire commented Dec 23, 2021 •

edited

Loading

sharonFogel commented Dec 26, 2021

About gradient balancing #23

About gradient balancing #23

Comments

miranghimire commented Dec 23, 2021 • edited Loading

sharonFogel commented Dec 26, 2021

miranghimire commented Dec 23, 2021 •

edited

Loading