mistral-nemo:12b-instruct-2407-q4_K_M: NVIDIA GPU requirements | LocalAIReady
← All Ollama models

Exact Ollama variant

Mistral Nemo

This page evaluates the immutable model blob behind mistral-nemo:12b-instruct-2407-q4_K_M. Results below use 8k context and 32 GB system RAM.

Q4_K_M quantization 6.96 GB exact model blob 1000k declared context

Exact model build verified

Technical details

SHA-256 digest: sha256:dd3af152229f92a3d61f3f115217c9c72f4b94d8be6778156dab23f894703c28

Run with Ollamaollama run mistral-nemo:12b-instruct-2407-q4_K_M
Check a custom hardware setup →

NVIDIA hardware comparison

Exact GPU VRAM Context Status Main limitation
GeForce RTX 3060 Laptop GPU 6GB 6.00 GB 8k Estimated Likely CPU/GPU offload VRAM
GeForce RTX 4060 8GB 8.00 GB 8k Estimated Likely CPU/GPU offload VRAM
GeForce RTX 3060 12GB 12.0 GB 8k Estimated Likely full GPU physical test data
GeForce RTX 5070 12GB 12.0 GB 8k Estimated Likely full GPU physical test data
GeForce RTX 4060 Ti 16GB 16.0 GB 8k Estimated Likely full GPU physical test data
GeForce RTX 5060 Ti 16GB 16.0 GB 8k Estimated Likely full GPU physical test data
GeForce RTX 3090 24GB 24.0 GB 8k Estimated Likely full GPU physical test data
GeForce RTX 4090 24GB 24.0 GB 8k Estimated Likely full GPU physical test data
GeForce RTX 5090 32GB 32.0 GB 8k Estimated Likely full GPU physical test data

Tested means the exact GPU, model digest, runtime and context were physically tested. Estimated never means tested.

Method

What is included

The comparison includes the exact model blob, an explicit 8k KV cache estimate, a conservative planning margin and possible system-RAM offload. It does not predict speed or claim a physical test where none exists.