Model HQ

Hardware optimized

Qualcomm Supported Models

A complete catalog of 95 AI models optimized for Qualcomm Snapdragon processors using QNN runtime.

Hover the Encyclopedia tab on the right, or tap any icon to learn what a model type does.
95
Optimized models
QNN
Runtime
CPU · GPU · NPU
Hardware targets
Why Qualcomm

Qualcomm optimization features

Snapdragon-optimized models deliver faster inference by leveraging the Qualcomm Neural Network (QNN) SDK across Snapdragon hardware.

Performance benefits

  • Optimized for Qualcomm Snapdragon architectures
  • Enhanced inference speed with QNN runtime
  • Power-efficient execution on mobile and edge devices
  • Hardware-specific optimizations for Hexagon DSP and Adreno GPU

Supported hardware

  • Qualcomm Snapdragon processors
  • Hexagon Digital Signal Processors (DSP)
  • Adreno GPUs
  • Qualcomm AI Engine and NPUs
Catalog

CPU Models

GGUF and tool models that run on the CPU alone — no GPU or NPU required.

Agentic

23 models
slim-sentiment-tool1.1B
slim-extract-tool1.1B
slim-summary-tool1.1B
slim-boolean-tool1.1B
slim-tags-tool1.1B
slim-xsum-tool1.1B
slim-emotions-tool1.1B
slim-topics-tool1.1B
slim-sql-tool1.1B
slim-extract-qwen-1.5b1.5B
slim-ner-tool1.1B
slim-sa-ner-tool1.1B
slim-tags-3b-tool3B
slim-extract-tiny-tool1.1B
slim-ratings-tool1.1B
slim-intent-tool1.1B
slim-category-tool1.1B
slim-nli-tool1.1B
slim-q-gen-phi-3-tool3.8B
slim-summary-tiny-tool1.1B
slim-qa-gen-phi-3-tool3.8B
slim-qa-gen-tiny-tool1.1B
slim-summary-tiny1.1B

Coding

1 models
qwen2.5-7b-coder7B

General Chat

10 models
tiny-llama-chat1.1B
llama-2-7b-chat7B
qwen2.5-32b32B
bling-qwen-0.5b0.5B
bling-qwen-1.5b1.5B
openhermes-2.5-mistral7B
dragon-llama-3.18B
zephyr-7b-beta7B
starling-lm-7b-alpha7B
miniCPM-V-2_64B

Instruct

12 models
qwen-2.5-14b-instruct14B
qwen2-7B-instruct7B
qwen-2-0.5b-instruct0.5B
qwen2-1.5-instruct1.5B
llama-3.1-instruct8B
llama-3.2-ib-instruct3B
llama-3.2-1b-instruct1.1B
llama-3-8b-instruct8B
mistral-7b-instruct-v0.37B
gemma-2-9b-instruct9B
gemma-2-27b-instruct27B
gemma-2b-it2B

Question-answer

11 models
dragon-yi-answer-tool6B
dragon-llama-answer-tool7B
dragon-mistral-answer-tool7B
bling-answer-tool1.1B
bling-tiny-llama1.1B
bling-phi-33.8B
bling-phi-3.53.8B
dragon-mistral-0.37B
dragon-yi-9b9B
dragon-qwen-7b7B
bling-stablelm-3b3B
Catalog

GPU/CPU/NPU Models

OpenVINO (OV) and ONNX models that run on the available GPU, CPU, or NPU.

Agentic

15 models
slim-extract-phi-33.8B
slim-boolean-phi-33.8B
slim-summary-phi-33.8B
slim-emotions1.1B
slim-topics1.1B
slim-sql1.1B
slim-sentiment1.1B
slim-extract-tiny1.1B
slim-intent1.1B
slim-tags1.1B
slim-ratings1.1B
slim-ner1.1B
slim-extract-qwen-nano0.5B
slim-sa-ner-phi-33.8B
slim-xsum-phi-33.8B

General Chat

3 models
llama-2-chat7B
phi-33.8B
phi-3.53.8B

Instruct

10 models
qwen-2.5-14b-instruct14B
qwen2-7B-instruct7B
qwen-2-0.5b-instruct0.5B
qwen2-1.5-instruct1.5B
llama-3.2-3b-instruct-onnx3B
llama-3.1-instruct8B
llama-3.2-1b-instruct1.1B
llama-3-8b-instruct8B
mistral-7b-instruct-v0.37B
gemma-2b-it2B

Prompt Safety

3 models
protectai-prompt-injection0.3B
unitary-unbiased-toxic-roberta0.1B
valurank-distilroberta-bias0.1B

Question-answer

3 models
dragon-mistral-0.37B
dragon-yi-9b9B
dragon-qwen-7b7B

Re-ranker

2 models
jina-reranker-turbo0.6B
jina-reranker-tiny0.1B

Vision

1 models
phi-3-vision4.2B
Catalog

NPU Models

Models designed to run on Neural Processing Units for efficient, low-power inference.

Vision

1 models
llama-3.2-3b-onnx-qnn3B
Next steps

Getting started with Qualcomm models

  1. 01Ensure you have a Qualcomm Snapdragon processor with QNN support.
  2. 02Select models optimized for Qualcomm from the Models section.
  3. 03The system automatically applies Qualcomm optimizations when available.
  4. 04Monitor performance improvements in the system metrics.

Not sure what your hardware supports?

Check the system requirements to find the right models for your machine.

Check system requirements

For Qualcomm-specific optimization questions, contact our technical support team at support@aibloks.com.

Reference

Encyclopedia