LHK_DPO_v1
LHK_DPO_v1 is trained via Direct Preference Optimization(DPO) from TomGrc/FusionNet_7Bx2_MoE_14B.
Details
coming sooon.
Evaluation Results
coming soon.
Contamination Results
coming soon.
- Downloads last month
- 15
Inference Providers
NEW
This model isn't deployed by any Inference Provider.
๐
Ask for provider support