| --- |
| license: apache-2.0 |
| inference: false |
| base_model: Qwen/Qwen2.5-Coder-7B-Instruct |
| base_model_relation: quantized |
| tags: [green, llmware-chat, p7, ov] |
| --- |
| |
| # qwen2.5-coder-7b-instruct-npu-ov |
|
|
| **qwen2.5-coder-7b-instruct-npu-ov** is an OpenVino int4 quantized version of [Qwen2.5-Coder-7B-Instruct](https://www.huggingface.co/Qwen/Qwen2.5-Coder-7B-Instruct), providing a fast inference implementation, optimized for AI PCs using Intel NPU. |
|
|
| This is from the latest release series from Qwen, and is an excellent chat/instruct coding assistant model. |
|
|
| ### Model Description |
|
|
| - **Developed by:** Qwen |
| - **Quantized by:** llmware |
| - **Model type:** qwen2.5 |
| - **Parameters:** 7 billion |
| - **Model Parent:** Qwen/Qwen2.5-Coder-7B-Instruct |
| - **Language(s) (NLP):** English |
| - **License:** Apache 2.0 |
| - **Uses:** Chat, general-purpose LLM |
| - **Quantization:** int4 |
| |
|
|
| ## Model Card Contact |
|
|
| [llmware on github](https://www.github.com/llmware-ai/llmware) |
|
|
| [llmware on hf](https://www.huggingface.co/llmware) |
|
|
| [llmware website](https://www.llmware.ai) |
|
|