Out-of-domain generalization from rephrasability and stability
Produced by Generalization Bounds for Autoregressive Processes and In-Context Learning
The task of predicting a conditional distribution for the next token given a preceding token context.
It is commonly trained with cross-entropy loss and evaluated through predictive risk or log loss.