Korean Datasets, Reward Models for RLHF
-
heegyu/KoSafeGuard-8b-0503
Text Generation • Updated • 111 • 5 -
heegyu/ko-reward-model-helpful-1.3b-v0.2
Text Classification • Updated • 15 -
heegyu/ko-reward-model-safety-1.3b-v0.2
Text Classification • Updated • 15 • 5 -
heegyu/ko-reward-model-helpful-roberta-large-v0.1
Text Classification • Updated • 11 • 1