Hugging Face
Models
Datasets
Spaces
Community
Docs
Enterprise
Pricing
Log In
Sign Up
NoManDeRY
/
DPO-Shift-Llama-3-8B-Ultrafeedback-decrease_linear-1.0to0.95
like
0
Text Generation
Transformers
Safetensors
HuggingFaceH4/ultrafeedback_binarized
llama
alignment-handbook
Generated from Trainer
conversational
text-generation-inference
arxiv:
2502.07599
Model card
Files
Files and versions
xet
Community
1
Deploy
Use this model
db054ac
DPO-Shift-Llama-3-8B-Ultrafeedback-decrease_linear-1.0to0.95
1.54 kB
1 contributor
History:
1 commit
NoManDeRY
initial commit
db054ac
verified
11 months ago
.gitattributes
1.52 kB
initial commit
11 months ago
README.md
24 Bytes
initial commit
11 months ago