Research questionHow can machine-learning models predict rare, large events in scale-free processes beyond the training data’s observed range?Large events in power-law processes are sparsely represented in training data, so models must infer behavior beyond the scales they have seen. Self-similarity may provide useful structure, but coarse-graining and anomalous scaling make that structure difficult to represent.