Preprint2025
Selective Self-to-Supervised Fine-Tuning (S3FT)
Unknown
Selective Self-to-Supervised Fine-Tuning (S3FT) improves LLM fine-tuning by selectively using the model's own correct predictions or gold responses to reduce overfitting and boost generalization.
0Feb 1, 2025Fine Tuning