argilla/distilabel-intel-orca-dpo-pairs
Viewer
•
Updated
•
12.9k
•
7.02k
•
176
LLMs, NLP, Alignment, DPO, RLHF, data labeling, text-classification, text-generation, token-classification