|
Download README.md from huggingkot/Impish_LLAMA_3B-bnb-4bit: direct link, hf CLI and curl.
- Browser
- Download file 861 Bytes
-
https://huggingface.co/huggingkot/Impish_LLAMA_3B-bnb-4bit/resolve/main/README.md
- Command line
-
hf download hf://huggingkot/Impish_LLAMA_3B-bnb-4bit/README.md
-
curl -L -o README.md https://huggingface.co/huggingkot/Impish_LLAMA_3B-bnb-4bit/resolve/main/README.md
861 Bytes
metadata
base_model:
- SicariusSicariiStuff/Impish_LLAMA_3B
This is a converted weight from Impish_LLAMA_3B model in unsloth 4-bit dynamic quant using this collab notebook.
About this Conversion
This conversion uses Unsloth to load the model in 4-bit format and force-save it in the same 4-bit format.
How 4-bit Quantization Works
- The actual 4-bit quantization is handled by BitsAndBytes (bnb), which works under Torch via AutoGPTQ or BitsAndBytes.
- Unsloth acts as a wrapper, simplifying and optimizing the process for better efficiency.
This allows for reduced memory usage and faster inference while keeping the model compact.