Tag: many

Many of Your DPOs are Secretly One: Attempting Unification Through Mutual Information

Many of Your DPOs are Secretly One: Attempting Unification Through Mutual Information arXiv:2501.01544v1 Announce Type: cross Abstract: Post-alignment of large language models (LLMs) is critical in improving their utility, safety, and alignment with human intentions. Direct preference optimisation (DPO) has become one of the most widely used algorithms for achieving this alignment, given its ability…

January 6, 2025

Many of Your DPOs are Secretly One: Attempting Unification Through Mutual Information