A small lm. (Russian only) Created to emulate a really simple one way dialogue; WARNING!!! CAN SWEAR! It was trained on two T4s from scratch. Final training time: 1 hour 2 minutes. The model consists of 3 transformer blocks stacked forming 6 layers.
Inference Providers
NEW
This model is not currently available via any of the supported third-party Inference Providers, and
HF Inference API was unable to determine this model's library.