Paper proposes adaptive timing for correcting bias in LLM reasoning
A new preprint proposes triggering bias corrections only when evidence of stereotyping accumulates during an LLM’s reasoning, reporting fewer interventions than fixed-interval methods but important accuracy tradeoffs across model types.