Upload folder using huggingface_hub (#8)

- f46dbb480c00ee95ee37925fc2895bd6dc3be73ace2fd205ac7d6fecd142cbf2 (d4a4817aee7b0cd94a5c59a0a4bb6b3b24ccd0b3)
- 05b194b5078dc3c4217669c2656413d9dab6f45f2f2c6fec9875d9ac4aaa15bd (813e11ec09cd9166bc7156b035bca78987e3405a)

Files changed (3) hide show

README.md CHANGED Viewed

@@ -37,16 +37,17 @@ metrics:
 ![image info](./plots.png)
 **Important remarks:**
-- The quality of the model output might slightly vary compared to the base model. There might be minimal quality loss.
 - These results were obtained on NVIDIA A100-PCIE-40GB with configuration described in config.json and are obtained after a hardware warmup. Efficiency results may vary in other settings (e.g. other hardware, image size, batch size, ...).
 - You can request premium access to more compression methods and tech support for your specific use-cases [here](https://z0halsaff74.typeform.com/pruna-access?typeform-source=www.pruna.ai).
 ## Setup
 You can run the smashed model with these steps:
-0. Check cuda, torch, packaging requirements are installed. For cuda, check with `nvcc --version` and install with `conda install nvidia/label/cuda-12.1.0::cuda`. For packaging and torch, run `pip install packaging torch`.
-1. Install the `pruna-engine` available [here](https://pypi.org/project/pruna-engine/) on Pypi. It might take 15 minutes to install.
     ```bash
    pip install pruna-engine[gpu]==0.6.0 --extra-index-url https://pypi.nvidia.com --extra-index-url https://pypi.ngc.nvidia.com --extra-index-url https://prunaai.pythonanywhere.com/
     ```

 ![image info](./plots.png)
 **Important remarks:**
+- The quality of the model output might slightly vary compared to the base model.
 - These results were obtained on NVIDIA A100-PCIE-40GB with configuration described in config.json and are obtained after a hardware warmup. Efficiency results may vary in other settings (e.g. other hardware, image size, batch size, ...).
 - You can request premium access to more compression methods and tech support for your specific use-cases [here](https://z0halsaff74.typeform.com/pruna-access?typeform-source=www.pruna.ai).
+- Results mentioning "first" are obtained after the first run of the model. The first run might take more memory or be slower than the subsequent runs due cuda overheads.
 ## Setup
 You can run the smashed model with these steps:
+0. Check that you have linux, python 3.10, and cuda 12.1.0 requirements installed. For cuda, check with `nvcc --version` and install with `conda install nvidia/label/cuda-12.1.0::cuda`.
+1. Install the `pruna-engine` available [here](https://pypi.org/project/pruna-engine/) on Pypi. It might take up to 15 minutes to install.
     ```bash
    pip install pruna-engine[gpu]==0.6.0 --extra-index-url https://pypi.nvidia.com --extra-index-url https://pypi.ngc.nvidia.com --extra-index-url https://prunaai.pythonanywhere.com/
     ```

model/optimized_model.pkl CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:36a1b24e7bd91ed9824b8874ac12c2ff8c9bf13b49b368021cf3c9b1cfa2d3b0
-size 2743426249

 version https://git-lfs.github.com/spec/v1
+oid sha256:df20a0fcf3f3af8f539e8d25167d459a345fac72dea3e55260d473f6461cc740
+size 2743426819

plots.png CHANGED Viewed