Nonparametric Neural Networks via Retrieval-Centric Deep Learning
October 6, 2026
This method replaces fixed-size weight matrices with a growing network that stores key-value pairs for every training data point. It utilizes functional gradient-based learning rules for kernelized attention layers to enable efficient retrieval and recombination during inference.
HOW THIS AFFECTS YOU
●
researcherThis provides a new framework for studying the duality between linear layers and attention mechanisms.