metadata
base_model: migtissera/Tess-XS-v1-3-yarn-128K
license: apache-2.0
metrics:
- accuracy
An instruct based fine tune of migtissera/Tess-XS-v1-3-yarn-128K.
It works well with long system prompts.
It isn't generic in a sense that it shouldn't be used for story telling, for example, but only for reasoning and text comprehension.
This model is trained on a private dataset. The high GSM8K score is NOT because of the MetaMath dataset.
Prompt Format:
SYSTEM: <ANY SYSTEM CONTEXT>
USER:
ASSISTANT: