Search papers, labs, and topics across Lattice.
This study introduces Free-Energy-Gated Plasticity (FEGP), an extension of the Predictive-Coding-inspired Variational Recurrent Neural Network (PV-RNN), enabling real-time motor learning in human-robot interactions through adaptive synaptic weight changes. By dynamically regulating the learning rate based on variational free energy, the model successfully learned three cyclic motor patterns without the need for pretraining or task-boundary signals. Experimental results demonstrated that FEGP significantly enhanced repertoire coverage and retention of previously learned behaviors, highlighting the importance of temporal plasticity allocation in continuous learning scenarios.
Real-time motor learning in robots just got a major upgrade鈥攄ynamic plasticity allows for seamless adaptation without forgetting past behaviors.
Fully online embodied learning requires synaptic adaptation to acquire new behaviors while preserving previously learned dynamics during ongoing interaction. We extend the Predictive-Coding-inspired Variational Recurrent Neural Network (PV-RNN) to continuously adapt its synaptic weights and propose Free-Energy-Gated Plasticity (FEGP), which regulates the effective learning rate according to variational free energy. In real-time physical human-robot interaction, a randomly initialized network acquired three cyclic motor patterns without offline pretraining, replay, or task-boundary signals, with all three patterns emerging in autonomous rollouts. Controlled experiments over ten randomized teaching streams and five network initializations per stream showed that FEGP substantially improved repertoire coverage and retention of previously acquired patterns after they left the recent observation window. Neither a constant learning rate matched to the gate's time-averaged effective rate nor replay of the same gain values with disrupted temporal organization reproduced these improvements. These results indicate that the temporal allocation of plasticity relative to model-environment mismatch, rather than simply its average magnitude or distribution, is critical for maintaining previously acquired behaviors during continued online learning.