PeterLauLukCh's picture
Update README.md
dfa552a verified
metadata
license: mit
datasets:
  - trl-lib/ultrafeedback_binarized
base_model:
  - ComparisonPO/Mistral-Instruct-7B-DPO_clean