Paper recorded by Signals 4 on 2026-09-25 in cs.LG. Abstract reproduced from arXiv; link to the original below.
Published 2026-09-25 on arXiv · recorded by Signals 4 on 2026-09-28
Category: cs.LG · 机器学习 · first seen 2026-09-28
Direct feedback alignment (DFA) trains hidden layers through fixed random projections of output error. With tanh hidden units and independent sigmoid outputs, plain stochastic gradient descent can stall near the loss of a constant predictor of class frequencies. We trace this stall to the error's common mode, the component shared across inputs. An exact mean-covariance decomposition separates a ra