upgrading InfogainLoss layer: (1) incorporating Softmax layer to make the gradeint computation robust, much like SoftmaxWithLoss layer (see: http://stackoverflow.com/a/34917052/1714410 for more information). (2) supporting loss along axis
S
shai committed
337b07589f4e44761bdb9ef4c242f83ca40c9da5
Parent: be163be