Back to feed
arXiv cs.LG·

Two to Tango: Coupled Task-Reference Selection for Safe LLM Fine-tuning

Signal
78
Hype
15
In three linesDualSelect, a fine-tuning method for LLMs, jointly selects safety references and compatible task samples to preserve safety alignment during adaptation. Tested on 1B-8B models, it improves Safety Avg. by at least 5.10 points over strongest baselines while maintaining task utility.
Read source
Your take?
Fine-tuningAI safetyAlignment

Summary generated by Claude — human-verified