ReDraft, Don't Just Distill: Reference-Driven Revision for Continual VLLM Post-Training
arXiv:2609.16639v3 Announce Type: replace Abstract: Continual post-training of large multimodal models should add new capabilities while preserving those from pre-training, and the two goals pull in opposite directions. SFT gives explicit target supervision that learns a task from near-zero…