Research questionHow can self-organizing maps scale to millions of neurons within GPU memory and bandwidth limits?At MEDLINE scale, each training epoch repeatedly searches a large codebook for every sample, making memory traffic and codebook storage major costs. These constraints limit both training speed and the number of neurons that can fit on one GPU.