Skip to content

BlankCode0/DPO_tldr_summarisation

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

1 Commit
 
 

Repository files navigation

YLMSR_SFT

2nd experiment of the paper "DPO: Your Language Model is Secretly a Reward Model"

About

2nd experiment of the paper "DPO: Your Language Model is Secretly a Reward Model"

Resources

Stars

Watchers

Forks

Releases

No releases published

Packages

 
 
 

Contributors