Self-editing meta-learning in Emergent Models
Prove whether, given sufficient memory and computation time and with reward supplied as an input, some Emergent Models can implement both a prediction subroutine and a reward-driven update subroutine that modifies a designated portion of the latent program, and characterize the maximum self-editing plasticity and its robustness implications.
References
In latent-universal settings, universality implies that the substrate can in principle represent any algorithm, while the latent nature of the program makes that algorithm part of the mutable state itself, suggesting the following conjecture: given sufficient memory and computation time, and providing the reward as an input, there exist some EMs that implement internally: (i) a prediction subroutine, which maps the current input to an output, and (ii) an update subroutine, which uses reward signal to modify a designated sub-portion of the program. In this view, the model could not only execute a prediction but also a feedback-driven self-improvement procedure acting on its own latent program, realizing a form of meta-learning \citep{vanschoren2018metalearningsurvey}. However, this remains speculative and requires further analysis, especially to determine the maximum degree of self-editing plasticity a substrate can support and the implications for robustness.