LHK_DPO_v1
LHK_DPO_v1 is trained via Direct Preference Optimization(DPO) from TomGrc/FusionNet_7Bx2_MoE_14B.
Details
coming sooon.
Evaluation Results
coming soon.
Contamination Results
coming soon.
- Downloads last month
- 9
Inference Providers
NEW
This model is not currently available via any of the supported Inference Providers.