Survey maps how foundation models can learn from their own outputs after training
A survey accepted to Findings of EMNLP 2026 catalogs 80 methods for adapting foundation models without human labels, preference data, stronger teachers or executable verifiers. It warns that internal learning signals can either improve a model or recursively amplify its errors.