llama embed nemotron 8b HNPU

Prebuilt HNPU artifacts for nvidia/llama-embed-nemotron-8b, a public embedding model.

For model behavior, license, intended use, and limitations, see the upstream model card.

Artifacts are architecture-pinned. Available artifact directories: v81/.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for runanywhere/llama_embed_nemotron_8b_HNPU

Finetuned
(6)
this model