Neural networks rely on optimizers that treat each weight matrix as one object, but these matrices have two parts: magnitude and direction. Decoupling these parts could improve training. This approach is tested in a preprint, but its real-world impact is still unclear.
STATUS
ACTIVE
CATEGORY
Research
SOURCES
1 linked
ENTITIES
3 detected
OVERRIDE
Automated
MOMENTUM
2 hours ago