Llama-3-Kimodo-GGML

Native F32 GGML/GGUF text-encoder components used by Kimodo. This is the reusable LLM2Vec encoder only; download a matching Kimodo diffusion model separately, for example Kimodo-SMPLX-RP-v1-GGML.

From a kimodo.cpp checkout with the Hugging Face CLI installed, install both with:

scripts/download_gguf_weights.sh --output "$PWD"

The components intentionally remain split into individual layers so kimodo.cpp can bound GPU memory use while evaluating the encoder.

Provenance and licence

The bundle is converted from Meta Llama-3-8B-Instruct and the MIT-licensed McGill LLM2Vec MNTP and supervised adapters. Built with Meta Llama 3.

LICENSE-META-LLAMA-3.txt and NOTICE accompany this distribution. Review the Meta Llama 3 Community License before use or redistribution. MANIFEST.json records the exact source commits and SHA-256 of each component.

Downloads last month
2,811
GGUF
Model size
0.5B params
Architecture
kimodo-llm2vec-layer
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support