
Daily Paper Cast · Jingwen Liang, Gengyu Wang
Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability
April 10, 2026·25 min·2 clips
A 2025 paper claimed SFT memorizes while RL generalizes, but new research questions if those experiments were flawed.
As heard by us
A careful walk-through that tests SFT claims against optimization, data, and model capability.
Daily Papercast treats the claim with a steady hand: does reasoning SFT really just memorize, or does the answer change with optimization, data, and model capability?
Why you'd press play
For a paper that tests whether repetition beats scale, press play here.
Listen to the show on