kp-mt5-large

This model is a fine-tuned version of jhpassion0621/kp-mt5-large on an unknown dataset. It achieves the following results on the evaluation set:

Model description

More information needed

More information needed

More information needed

The following hyperparameters were used during training:

Training Loss	Epoch	Step	Bleu	Gen Len	Validation Loss
1.0364	0.29	17000	32.5573	44.7582	0.8278
0.8819	0.58	34000	37.1161	45.0568	0.7062
0.7731	0.87	51000	40.329	45.7359	0.6188
0.7339	1.16	68000	41.7643	45.8618	0.5866
0.7093	1.45	85000	42.6878	45.5649	0.5657
0.6818	1.74	102000	43.2023	45.7701	0.5609
0.6739	2.00	117444	43.3983	45.6585	0.5586

Safetensors

Model size

1B params

Tensor type

F32

Inference Providers NEW

This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Unable to build the model tree, the base model loops to the model itself. Learn more.