← back
arXivZilin Du, Bowen Yang, Boyang Albert LiThu, Oct 1, 2026, 10:23 AM PDT
score 16.5

New training method picks better data for language models

Original: Scalable, Transferable Meta-network for Data Selection Requires a Different Loss (and Why the Obvious Choice is Problematic)

Source: arxiv.org ↗

Writing ELI5 summary…