← All ICML reports
DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and Its Loss' Convexity is Dispensable)
Wenxuan Zhou, Shujian Zhang, brice magdalou, John Wheatley Lambert, Ehsan Amid, Richard Nock, Andrew Hard
SAI review preview
SAI reviewed this ICML 2026 paper and its available research artifacts. Sign in to read the complete referee report, inline feedback, and execution-based verification where available.
28Detailed review comments
CompletePaper and code review
Execution review not startedReplication status