Model HQ
DocumentationHardware optimized
AMD Supported Models
A complete catalog of 325 AI models optimized for enhanced performance across AMD CPUs, GPUs, and NPUs.
Hover the Encyclopedia tab on the right, or tap any icon to learn what a model type does.
325
Optimized models
ONNX
Runtime
CPU · GPU · NPU
Hardware targets
Why AMD
AMD optimization features
These models are optimized for enhanced performance on AMD hardware devices.
Performance benefits
- Optimized for AMD CPU, GPU, and NPU architectures
- Enhanced inference speed with ONNX Runtime and VITIS AI acceleration
- Efficient execution of AI models with optimized hardware utilization
- Streamlined AI inference for AMD Ryzen AI and Radeon platforms
Supported hardware
- AMD x86_64 CPU
- AMD Ryzen AI
- AMD GPUs and NPUs
- Windows 11 devices
Catalog
CPU Models
GGUF and tool models that run on the CPU alone — no GPU or NPU required.
Agentic
28 modelsslim-boolean-phi-3-gguf3.8Bslim-extract-phi-3-gguf3.8Bslim-extract-qwen-1.5b-gguf1.5Bslim-extract-qwen-nano-gguf0.5Bslim-sa-ner-phi-3-gguf3.8Bslim-summary-phi-3-gguf3.8Bslim-xsum-phi-3-gguf3.8Bslim-boolean-tool1.1Bslim-category-tool1.1Bslim-emotions-tool1.1Bslim-extract-tool1.1Bslim-intent-tool1.1Bslim-ner-tool1.1Bslim-nli-tool1.1Bslim-q-gen-phi-3-tool3.8Bslim-q-gen-tiny-tool1.1Bslim-qa-gen-phi-3-tool3.8Bslim-qa-gen-tiny-tool1.1Bslim-ratings-tool1.1Bslim-sa-ner-tool1.1Bslim-sentiment-tool1.1Bslim-sql-tool1.1Bslim-summary-tiny-tool1.1Bslim-summary-tool1.1Bslim-tags-3b-tool3Bslim-tags-tool1.1Bslim-topics-tool1.1Bslim-xsum-tool1.1BCoding
5 modelsqwen2.5-coder-0.5b-instruct-generic-cpu:4-foundry0.5Bqwen2.5-coder-1.5b-instruct-generic-cpu:4-foundry1.5Bqwen2.5-coder-14b-instruct-generic-cpu:4-foundry14Bqwen2.5-coder-7b-instruct-generic-cpu:4-foundry7Bqwen-2.5-7b-coder-gguf7BGeneral Chat
14 modelsgpt-oss-20b-generic-cpu:1-foundryNAdeepseek-r1-distill-qwen-14b-generic-cpu:4-foundry14Bdeepseek-r1-distill-qwen-7b-generic-cpu:4-foundry7Bqwen3-0.6b-generic-cpu:4-foundry0.6Bqwen3-1.7b-generic-cpu:2-foundry1.7Bqwen3-14b-generic-cpu:2-foundry14Bqwen3-4b-generic-cpu:3-foundry4Bqwen3-8b-generic-cpu:2-foundry8BPhi-4-mini-reasoning-generic-cpu:3-foundry3.8BPhi-4-generic-cpu:2-foundry14Bqwen3.5-0.8b-generic-cpu:2-foundry0.8Bqwen3.5-2b-generic-cpu:2-foundry2Bqwen3.5-4b-generic-cpu:2-foundry4Bqwen3.5-9b-generic-cpu:2-foundry9BGeneral Chat - GGUF
34 modelsdeepseek-qwen-14b-gguf14Bdeepseek-qwen-7b-gguf7Bdragon-llama-3.1-gguf8Bgemma-3-12b-gguf12Bgemma-3-4b-gguf4Bgemma-4-26b-gguf26Bgemma-4-2b-gguf2Bgemma-4-4b-gguf4Bgpt-oss-20b-gguf20Bgranite-4-micro-gguf1.1Bllama-2-7b-chat-gguf7Bliquidai-lfm2-2.6b-gguf2.6Bministral-3-14b-gguf14Bminicpm-2.6-gguf4Bmistral-small-3.2-24b-gguf24Bolmo-13b-gguf13Bopenhermes-2.5-mistral-7b-gguf7Bphi-3-gguf3.8Bphi-3.5-gguf3.8Bphi-4-gguf14Bphi-4-mini-gguf3.8Bphi-4-mini-reasoning-gguf3.8Bqwen-3.5-27b-gguf27Bqwen-3.5-35b-a3b-gguf35Bqwen-3.5-4b-gguf4Bqwen-3.5-9b-gguf9Bqwen2.5-32b-gguf32Bqwen2.5-72b-gguf72Bqwen3-1.7b-gguf1.7Bqwen3-14b-gguf14Bqwen3-8b-gguf8Bstarling-lm-7b-alpha-gguf7Btiny-llama-chat-gguf1.1Bzephyr-7b-beta-gguf7BInstruct
9 modelsqwen2.5-0.5b-instruct-generic-cpu:4-foundry0.5Bqwen2.5-1.5b-instruct-generic-cpu:4-foundry1.5Bqwen2.5-14b-instruct-generic-cpu:4-foundry14Bqwen2.5-7b-instruct-generic-cpu:4-foundry7BPhi-4-mini-instruct-generic-cpu:5-foundry3.8Bmistralai-Mistral-7B-Instruct-v0-2-generic-cpu:3-foundry7BPhi-3-mini-128k-instruct-generic-cpu:3-foundry3.8BPhi-3-mini-4k-instruct-generic-cpu:3-foundry3.8BPhi-3.5-mini-instruct-generic-cpu:2-foundry3.8BInstruct - GGUF
12 modelsgemma-2-27b-instruct-gguf27Bgemma-2-9b-instruct-gguf9Bllama-3-8b-instruct-gguf8Bllama-3.1-instruct-gguf8Bllama-3.2-1b-instruct-gguf1.1Bllama-3.2-3b-instruct-gguf3Bmistral-7b-instruct-v0.3-gguf7Bqwen-2.5-14b-instruct-gguf14Bqwen2-0.5b-instruct-gguf0.5Bqwen2-1.5b-instruct-gguf1.5Bqwen2-7B-instruct-gguf7Bqwen3-4b-instruct-gguf4BQuestion-answer
13 modelsbling-qwen-0.5b-gguf0.5Bbling-tiny-llama-onnx1.1Bbling-phi-3-onnx3.8Bbling-qwen-1.5b-ov1.5Bbling-tiny-llama-ov1.1Bbling-phi-3-ov3.8Bbling-qwen-mini-tool1.5Bdragon-qwen-7b-gguf7Bdragon-mistral-0.3-gguf7Bdragon-yi-9b-gguf9Bdragon-llama-answer-tool7Bdragon-mistral-answer-tool7Bdragon-yi-answer-tool6BVision
4 modelsqwen2.5-vl-3b-instruct-gguf3Bqwen3-vl-30b-gguf30Bqwen3-vl-4b-gguf4Bqwen3-vl-8b-gguf8BCatalog
GPU/CPU/NPU Models
ONNX and OpenVINO (OV) models that run on the available GPU, CPU, or NPU.
Agentic - ONNX
13 modelsslim-boolean-phi-3-onnx3.8Bslim-emotions-onnx1.1Bslim-extract-phi-3-onnx3.8Bslim-extract-tiny-onnx1.1Bslim-intent-onnx1.1Bslim-ner-onnx1.1Bslim-ratings-onnx1.1Bslim-sentiment-onnx1.1Bslim-sql-onnx1.1Bslim-summary-phi-3-onnx3.8Bslim-summary-tiny-onnx1.1Bslim-tags-onnx1.1Bslim-topics-onnx1.1BAgentic - OV
22 modelsslim-boolean-phi-3-ov3.8Bslim-category-ov1.1Bslim-emotions-ov1.1Bslim-extract-phi-3-ov3.8Bslim-extract-qwen-0.5b-ov0.5Bslim-extract-qwen-1.5b-ov1.5Bslim-extract-tiny-ov1.1Bslim-intent-ov1.1Bslim-ner-ov1.1Bslim-q-gen-tiny-ov1.1Bslim-qa-gen-tiny-ov1.1Bslim-ratings-ov1.1Bslim-sa-ner-phi-3-ov3.8Bslim-sentiment-ov1.1Bslim-sql-ov1.1Bslim-sql-phi-3-ov3.8Bslim-sql-qwen-base-ov1.5Bslim-summary-phi-3-ov3.8Bslim-summary-tiny-ov1.1Bslim-tags-ov1.1Bslim-topics-ov1.1Bslim-xsum-phi-3-ov3.8BCoding
2 modelscodegemma-7b-it-ov7Bqwen2.5-coder-7b-instruct-ov7BEmbedding - ONNX
9 modelsall-mini-lm-l6-v2-onnx0.02Bbge-base-en-v1.5-onnx0.11Bbge-large-en-v1.5-onnx0.33Bbge-small-en-v1.5-onnx0.03Bgte-base-onnx0.11Bgte-large-onnx0.33Bgte-small-onnx0.03Bindustry-bert-contracts-onnx0.11Bindustry-bert-insurance-onnx0.11BEmbedding - OV
14 modelsall-mini-lm-l6-v2-ov0.02Ball-mpnet-base-v2-ov0.11Bbge-base-en-v1.5-ov0.11Bbge-large-en-v1.5-ov0.33Bbge-small-en-v1.5-ov0.03Bgte-base-ov0.11Bgte-large-ov0.33Bgte-small-ov0.03Bindustry-bert-asset-management-ov0.11Bindustry-bert-contracts-ov0.11Bindustry-bert-insurance-ov0.11Bindustry-bert-loans-ov0.11Bindustry-bert-sec-ov0.11Bparaphrase-multilingual-MiniLM-L12-v2-ov0.12BGeneral Chat
14 modelsgpt-oss-20b-generic-gpu:1-foundry20Bdeepseek-r1-distill-qwen-14b-generic-gpu:4-foundry14Bqwen3-0.6b-generic-gpu:2-foundry0.6Bqwen3-1.7b-generic-gpu:2-foundry1.7Bqwen3-14b-generic-gpu:2-foundry14Bqwen3-4b-generic-gpu:2-foundry4Bqwen3-8b-generic-gpu:2-foundry8BPhi-4-mini-reasoning-generic-gpu:3-foundry3.8Bdeepseek-r1-distill-qwen-7b-generic-gpu:4-foundry7BPhi-4-generic-gpu:2-foundry14Bqwen3.5-0.8b-generic-gpu:2-foundry0.8Bqwen3.5-2b-generic-gpu:2-foundry2Bqwen3.5-4b-generic-gpu:2-foundry4Bqwen3.5-9b-generic-gpu:2-foundry9BGeneral Chat - OV
30 modelsqwen2-0.5b-chat-ov0.5Bqwen3-1.7b-ov1.7Bqwen3-14b-ov14Bqwen3-4b-ov4Bqwen3-8b-ov8Bdolphin-2.9.4-llama3.1-8b-ov8Bllama-2-13b-chat-ov13Bllama-2-chat-ov7Btiny-llama-chat-ov1.1Bphi-3-ov3.8Bphi-4-ov14Bphi-4-mini-ov3.8Bdolphin-2.9.3-mistral-7b-32k-ov7Bteknium-open-hermes-2.5-mistral-ov7Bzephyr-mistral-7b-chat-ov7Byi-1.5-34b-ov34Byi-6b-1.5v-chat-ov6Byi-9b-chat-ov9Bgemma-2-27b-ov27Bgemma-2b-it-ov2Bgemma-7b-it-ov7Bstablelm-2-12b-chat-ov12Bstablelm-2-zephyr-1_6b-ov1.6Bstablelm-zephyr-3b-ov3Bdreamgen-wizardlm-2-7b-ov7Bgranite-4-micro-ov1.1Bintel-neural-chat-7b-v3-2-ov7Bopenchat-3.6-8b-20240522-ov8Btiny-dolphin-2.8-1.1b-ov1.1Blcm-dreamshaper-ov1.1BInstruct
13 modelsqwen2.5-0.5b-instruct-generic-gpu:4-foundry0.5Bqwen2.5-1.5b-instruct-generic-gpu:4-foundry1.5Bqwen2.5-14b-instruct-generic-gpu:4-foundry14Bqwen2.5-7b-instruct-generic-gpu:4-foundry7Bqwen2.5-coder-0.5b-instruct-generic-gpu:4-foundry0.5Bqwen2.5-coder-1.5b-instruct-generic-gpu:4-foundry1.5Bqwen2.5-coder-14b-instruct-generic-gpu:4-foundry14Bqwen2.5-coder-7b-instruct-generic-gpu:4-foundry7BPhi-4-mini-instruct-generic-gpu:5-foundry3.8Bmistralai-Mistral-7B-Instruct-v0-2-generic-gpu:2-foundry7BPhi-3-mini-128k-instruct-generic-gpu:2-foundry3.8BPhi-3-mini-4k-instruct-generic-gpu:2-foundry3.8BPhi-3.5-mini-instruct-generic-gpu:2-foundry3.8BInstruct - ONNX
8 modelsllama-3.1-instruct-onnx8Bllama-3.2-1b-instruct-onnx1.1Bllama-3.2-3b-instruct-onnx3Bllama-2-chat-onnx7Btiny-llama-chat-onnx1.1Bphi-3-onnx3.8Bmistral-7b-instruct-v0.3-onnx7Bgemma-2b-it-onnx2BInstruct - OV
15 modelsqwen2-1.5b-instruct-ov1.5Bqwen2-7b-instruct-ov7Bqwen2.5-0.5b-instruct-ov0.5Bqwen2.5-1.5b-instruct-ov1.5Bqwen2.5-14b-instruct-ov14Bqwen2.5-32b-instruct-ov32Bqwen2.5-3b-instruct-ov3Bqwen2.5-72b-instruct-ov72Bllama-3.1-instruct-ov8Bllama-3.2-1b-instruct-ov1.1Bllama-3.2-3b-instruct-ov3Bmistral-7b-instruct-v0.2-ov7Bmistral-7b-instruct-v0.3-ov7Bmistral-nemo-instruct-2407-ov12Bmistral-small-instruct-2409-ov22BLanguage Detector
1 modelsxlm-roberta-language-detector-ov0.28BMath
1 modelsmathstral-7b-ov7BPrompt Safety
7 modelsprotectai-prompt-injection-onnx0.3Bunitary-toxic-roberta-onnx0.1Bvalurank-bias-onnx0.1Bmalicious-url-detector-ov0.1Bprotectai-prompt-injection-ov0.3Bunitary-toxic-roberta-ov0.1Bvalurank-bias-ov0.1BQuestion-answer
8 modelsdragon-mistral-0.3-onnx7Bdragon-qwen-7b-ov7Bdragon-llama2-ov7Bnvidia-llama3-chatqa-1.5-8b-ov8Bdragon-mistral-0.3-ov7Bdragon-mistral-ov7Bdragon-yi-6b-ov6Bdragon-yi-9b-ov9BRe-ranker
4 modelsjina-reranker-tiny-onnx0.1Bjina-reranker-turbo-onnx0.6Bjina-reranker-v1-tiny-en-ov0.1Bjina-reranker-v1-turbo-en-ov0.6BText-to-speech
1 modelsspeech-t5-tts-ov0.6BVision
5 modelsphi-3-vision-onnx4.2Bqwen2.5-vl-3b-ov3Bqwen2.5-vl-7b-ov7Bqwen2-vl-2b-instruct-ov2Bqwen2-vl-7b-instruct-ov7BCatalog
NPU Models
Models designed to run on Neural Processing Units for efficient, low-power inference.
Agentic
10 modelsslim-emotions-npu-ov1.1Bslim-extract-tiny-npu-ov1.1Bslim-intent-npu-ov1.1Bslim-ner-npu-ov1.1Bslim-ratings-npu-ov1.1Bslim-sentiment-npu-ov1.1Bslim-sql-npu-ov1.1Bslim-summary-tiny-npu-ov1.1Bslim-tags-npu-ov1.1Bslim-topics-npu-ov1.1BGeneral Chat
4 modelsDeepSeek-R1-Distill-Qwen-7B-vitis-npu:2-foundry7BPhi-4-mini-reasoning-vitis-npu:2-foundry3.8Bmistral-7b-v0.3-npu-ov7Byi-9b-npu-ov9BInstruct
10 modelsqwen2.5-0.5b-instruct-vitis-npu:3-foundry0.5Bqwen2.5-7b-instruct-vitis-npu:2-foundry7Bqwen2.5-coder-0.5b-instruct-vitis-npu:2-foundry0.5Bqwen2.5-coder-1.5b-instruct-vitis-npu:2-foundry1.5Bqwen2.5-coder-7b-instruct-vitis-npu:2-foundry7BPhi-4-mini-instruct-vitis-npu:2-foundry3.8BMistral-7B-Instruct-v0-2-vitis-npu:2-foundry7Bphi-3-mini-128k-instruct-vitis-npu:2-foundry3.8BPhi-3-mini-4k-instruct-vitis-npu:2-foundry3.8Bphi-4-mini-instruct-vitis-npu:2-foundry3.8BCatalog
Cloud Models
Models that run on remote servers over the internet, increasing speed and complex capabilities.
Cloud
15 modelsgpt-5.2-proNAgpt-5.2NAgpt-5-miniNAgpt-5-nanoNAgpt-4.1NAclaude-opus-4-5NAclaude-haiku-4-5NAclaude-sonnet-4-5NAclaude-sonnet-4-20250514NAclaude-opus-4-20250514NAgemini-3-pro-previewNAgemini-3-flash-previewNAgemini-2.5-proNAgemini-2.5-flashNAgemini-2.5-flash-liteNANext stepsCheck system requirements
Getting started with AMD models
- 01Ensure you have a device with compatible AMD hardware.
- 02Select models optimized for AMD from the Models section.
- 03The system automatically applies AMD optimizations when available.
- 04Monitor performance improvements in the system metrics.
Not sure what your hardware supports?
Check the system requirements to find the right models for your machine.
For AMD-specific optimization questions, contact our technical support team at support@aibloks.com.
Reference
