Model HQ
DocumentationHardware optimized
NVIDIA Supported Models
A complete catalog of 122 AI models optimized optimized for NVIDIA GPUs and accelerators.
Hover the Encyclopedia tab on the right, or tap any icon to learn what a model type does.
122
Optimized models
CUDA
Runtime
NVIDIA GPU
Hardware target
Why NVIDIA
NVIDIA optimization features
These models are optimized for enhanced performance on NVIDIA GB10 devices (DGX Spark).
Performance benefits
- Optimized for NVIDIA GPUs such as GB10, GeForce, and RTX
- Enhanced AI inference performance with NVIDIA GPUs
- High-throughput execution for demanding AI workloads
- Hardware-specific optimizations for latest NVIDIA GPUs
Supported hardware
- NVIDIA DGX Spark
- NVIDIA GB10 GPUs
- NVIDIA GeForce GPUs
- NVIDIA RTX GPUs
Catalog
All Supported Models
Models optimized for NVIDIA GPUs and accelerators
Agentic
29 modelsslim-ner-tool1.1Bslim-sentiment-tool1.1Bslim-emotions-tool1.1Bslim-ratings-tool1.1Bslim-intent-tool1.1Bslim-nli-tool1.1Bslim-topics-tool1.1Bslim-tags-tool1.1Bslim-sql-tool1.1Bslim-category-tool1.1Bslim-xsum-tool1.1Bslim-extract-tool1.1Bslim-extract-phi-3-gguf3.8Bslim-extract-qwen-1.5b-gguf1.5Bslim-extract-qwen-nano-gguf0.5Bslim-extract-tiny-tool1.1Bslim-summary-tiny-tool1.1Bslim-summary-phi-3-gguf3.8Bslim-xsum-phi-3-gguf3.8Bslim-boolean-tool1.1Bslim-boolean-phi-3-gguf3.8Bslim-sa-ner-phi-3-gguf3.8Bslim-sa-ner-tool1.1Bslim-tags-3b-tool3Bslim-summary-tool1.1Bslim-q-gen-phi-3-tool3.8Bslim-q-gen-tiny-tool1.1Bslim-qa-gen-tiny-tool1.1Bslim-qa-gen-phi-3-tool3.8BCloud
15 modelsclaude-opus-4-5NAclaude-haiku-4-5NAclaude-sonnet-4-5NAclaude-sonnet-4-20250514NAclaude-opus-4-20250514NAgemini-3-pro-previewNAgemini-3-flash-previewNAgemini-2.5-proNAgemini-2.5-flashNAgemini-2.5-flash-liteNAgpt-5.2-proNAgpt-5.2NAgpt-5-miniNAgpt-5-nanoNAgpt-4.1NACoding
1 modelsqwen-2.5-7b-coder-gguf7BEmbedding
8 modelsall-mini-lm-L6-v20.02Ball-mpnet-base-v20.1Bindustry-bert-insurance0.1Bindustry-bert-contracts0.1Bindustry-bert-asset-management0.1Bindustry-bert-sec0.1Bindustry-bert-loans0.1Bnomic-ai/nomic-embed-text-v10.1BGeneral Chat
34 modelsllama-2-7b-chat-gguf7Bdragon-llama-3.1-gguf8Btiny-llama-chat-gguf1.1Bqwen3-1.7b-gguf1.7Bqwen3-8b-gguf8Bqwen3-14b-gguf14Bqwen-3.5-4b-gguf4Bqwen-3.5-9b-gguf9Bqwen-3.5-27b-gguf27Bqwen-3.5-35b-a3b-gguf35Bqwen2.5-32b-gguf32Bqwen2.5-72b-gguf72Bdeepseek-qwen-14b-gguf14Bdeepseek-qwen-7b-gguf7Bphi-3-gguf3.8Bphi-3.5-gguf3.8Bphi-4-gguf14Bphi-4-mini-gguf3.8Bphi-4-mini-reasoning-gguf3.8Bmistral-small-3.2-24b-gguf24Bministral-3-14b-gguf14Bopenhermes-2.5-mistral-7b-gguf7Bzephyr-7b-beta-gguf7Bstarling-lm-7b-alpha-gguf7Bgemma-3-4b-gguf4Bgemma-3-12b-gguf12Bgemma-4-4b-gguf4Bgemma-4-2b-gguf2Bgemma-4-26b-gguf26Bgpt-oss-20b-gguf20Bolmo-13b-gguf13Bgranite-4-micro-gguf1.1Bliquidai-lfm2-2.6b-gguf2.6Bminicpm-2.6-gguf4BInstruct
13 modelsgemma-2-9b-instruct-gguf9Bgemma-2-27b-instruct-gguf27Bllama-3.1-instruct-gguf8Bllama-3-8b-instruct-gguf8Bllama-3.2-1b-instruct-gguf1.1Bllama-3.2-3b-instruct-gguf3Bmistral-7b-instruct-v0.3-gguf7Bqwen2.5-vl-3b-instruct-gguf3Bqwen2-7B-instruct-gguf7Bqwen3-4b-instruct-gguf4Bqwen2-1.5b-instruct-gguf1.5Bqwen2-0.5b-instruct-gguf0.5Bqwen-2.5-14b-instruct-gguf14BRe-ranker
6 modelsjina-reranker-tiny-ppt0.1Bjina-reranker-turbo-ppt0.6Bjina-reranker-tiny-onnx0.1Bjina-reranker-turbo-onnx0.6Bjina-reranker-v1-turbo-en0.6Bjina-reranker-v1-tiny-en0.1BSpeech-to-text
1 modelswhisper-cpp-base-english0.07BQuestion-answer
11 modelsbling-qwen-mini-tool1.5Bdragon-qwen-7b-gguf7Bbling-phi-3-gguf3.8Bbling-phi-3.5-gguf3.8Bdragon-mistral-0.3-gguf7Bdragon-yi-9b-gguf9Bdragon-yi-answer-tool6Bbling-stablelm-3b-gguf3Bbling-answer-tool1.1Bdragon-llama-answer-tool7Bdragon-mistral-answer-tool7BVision
4 modelsqwen2.5-vl-3b-instruct-gguf3Bqwen3-vl-8b-gguf8Bqwen3-vl-4b-gguf4Bqwen3-vl-30b-gguf30BNext stepsCheck system requirements
Getting started with NVIDIA models
- 01Ensure you have a device with compatible NVIDIA GPU.
- 02Select models optimized for NVIDIA from the Models section.
- 03The system automatically applies NVIDIA optimizations when available.
- 04Monitor performance improvements in the system metrics.
Not sure what your hardware supports?
Check the system requirements to find the right models for your machine.
For NVIDIA-specific optimization questions, contact our technical support team at support@aibloks.com.
Reference
