Preprint introduces a benchmark and replay method for continual reasoning-model training
A new arXiv preprint studies whether reasoning models can learn tasks sequentially without losing ground, and proposes Continual Prompt Replay to close the gap with joint multitask training.