Open
Description
There are three backward calls inside gardient balancing between generator loss & OCR loss:
Won't these calls accumulate the gradients during the call of optimizer.step(); I thought our objective here was to simply compute the gardient balancing terms and multiply those to the loss or could you please give overview of what's going on here inside gradient balancing incase I misunderstood something?
Metadata
Metadata
Assignees
Labels
No labels