| --- |
| license: llama3 |
| inference: false |
| base_model: meta-llama/Meta-Llama-3.1-8B-Instruct |
| base_model_relation: quantized |
| tags: |
| - green |
| - p8 |
| - llmware-chat |
| - ov |
| --- |
| |
| # llama-3.1-instruct-ov |
|
|
| **llama-3.1-instruct-ov** is an OpenVino int4 quantized version of Llama 3.1 Instruct, providing a very fast inference implementation, optimized for AI PCs using Intel GPU, CPU and NPU. |
|
|
| [**llama-3.1-instruct**](https://huggingface.co/meta-llama/Meta-Llama-3.1-8B-Instruct) is a leading open source general foundation model from Meta. |
|
|
|
|
| ### Model Description |
|
|
| - **Developed by:** meta-llama |
| - **Quantized by:** llmware |
| - **Model type:** llama-3.1 |
| - **Parameters:** 8 billion |
| - **Model Parent:** meta-llama/Meta-Llama-3.1-8B-Instruct |
| - **Language(s) (NLP):** English |
| - **License:** Llama 3.1 Community License |
| - **Uses:** General chat use cases |
| - **RAG Benchmark Accuracy Score:** NA |
| - **Quantization:** int4 |
| |
|
|
| ## Model Card Contact |
|
|
| [llmware on github](https://www.github.com/llmware-ai/llmware) |
|
|
| [llmware on hf](https://www.huggingface.co/llmware) |
|
|
| [llmware website](https://www.llmware.ai) |