mgoin's picture
Create README.md
e814ae3 verified
|
raw
history blame
379 Bytes
metadata
license: other
license_name: nvidia-open-model-license
license_link: >-
  https://developer.download.nvidia.com/licenses/nvidia-open-model-license-agreement-june-2024.pdf
tags:
  - fp8
  - vllm
base_model: nvidia/Minitron-4B-Base

Minitron-4B-Base-FP8

FP8 quantized checkpoint of nvidia/Minitron-4B-Base for use with vLLM.