TokenAssemble

Can a GMKtec EVO-X2 Ryzen AI Max+ 395 128GB run Llama-3.2-1B-Instruct?

Yes. Llama-3.2-1B-Instruct needs 2.01GB at Q8_0 with llama.cpp and 8K context; the GMKtec EVO-X2 Ryzen AI Max+ 395 128GB has 128GB of unified memory (~75% GPU-addressable) (95.5GB usable) — grade A, with decode speed in the interactive tier.

GMKtec EVO-X2 Ryzen AI Max+ 395 128GB · Llama-3.2-1B-Instruct · Q8_0

Fits — grade A

InteractivebetaROCm/Vulkan performance not yet calibrated
weights 1.32 GBKV cache 0.27 GBoverhead 0.42 GBpool 95.5 GBneeds 2.01 GB

assumes llama.cpp · batch 1 · 8K context · fp16 KV cache

Also runs with: Ollama (one-line install) · LM Studio (GUI)

$3,399.99 MSRP · as of 2026-07-11

Affiliate link — we may earn a commission. Verdicts are computed before any link is attached.

Every quant, graded

GMKtec EVO-X2 Ryzen AI Max+ 395 128GB × Llama-3.2-1B-Instruct, llama.cpp, 8K context.

QuantFile sizeFitSpeed tier
Q8_01.3 GBFits · Ainteractive · beta
Q6_K1.0 GBFits · Ainteractive · beta
Q5_K_M0.9 GBFits · Ainteractive · beta
Q4_K_M0.8 GBFits · Ainteractive · beta

Llama-3.2-1B-Instruct on similar hardware

More models on the GMKtec EVO-X2 Ryzen AI Max+ 395 128GB

Verdict computed from published specs and measured GGUF file sizes — see how the math works.