metadata
language:
- ko
- en
license: mit
Model Card for free-solar-dpo-v0.1
Developed by : Freewheelin AI Technical Team
Hardware and Software
- Training Factors: We fine-tuned this model using the HuggingFace TRL Trainer
Method
- This model was trained using the learning method introduced in the SOLAR paper.