Nemotron models that have been converted and/or quantized to work well in vLLM
-
mgoin/Nemotron-4-340B-Instruct-hf-FP8
Text Generation β’ 341B β’ Updated β’ 60 β’ 3 -
mgoin/Nemotron-4-340B-Base-hf-FP8
Text Generation β’ 341B β’ Updated β’ 94 β’ 2 -
mgoin/Nemotron-4-340B-Instruct-hf
Text Generation β’ 341B β’ Updated β’ 7 β’ 4 -
mgoin/Nemotron-4-340B-Base-hf
Text Generation β’ 341B β’ Updated β’ 3 β’ 1