HP ZBook Ultra 14 G1a: which AI models can it run?

Estimated from the specifications below. Choose the memory size you have or plan to buy, then read the table for that size.

Specifications

Brand
HP
Chip
AMD Ryzen AI Max+ PRO 395
Processor
AMD Ryzen AI Max+ PRO 395 (16 cores)
Graphics
AMD Radeon 8060S (integrated)
Memory type
LPDDR5X-8533
Memory options
16 GB, 32 GB, 64 GB, 128 GB
Memory bandwidth
256 GB/s

Sources: https://h20195.www2.hp.com/v2/GetDocument.aspx?docname=c09119722, https://www.amd.com/en/products/processors/desktops/ryzen/ryzen-ai-halo/ryzen-ai-max-plus-395.html

Bandwidth is AMD's figure for the Ryzen AI Max+ 395 (on its Ryzen AI Halo page); HP lists the memory as LPDDR5X-8533, AMD lists a maximum of 8000, so real bandwidth may differ. Other chip options are Max PRO 390, 385 and 380.

With 16 GB of memory

About 12 GB of it can hold a model.

ModelFitNeeds (GB)Writing speed (tokens per second)SpeedQualityBasis
Llama 3.2 3B
llama3.2:3b
Fits 5 45 to 83 Fast not measured yet Estimated
Qwen3 4B
qwen3:4b · thinks first
Fits 6 36 to 67 Fast not measured yet Estimated
Gemma 3 4B
gemma3:4b
Fits 6.7 27 to 50 Fast not measured yet Estimated
Mistral 7B v0.3
mistral:7b
Fits 7.8 20 to 38 Fast not measured yet Estimated
Llama 3.1 8B
llama3.1:8b
Fits 8.3 18 to 34 Fast not measured yet Estimated
Qwen3 8B
qwen3:8b · thinks first
Fits 8.9 17 to 32 Fast not measured yet Estimated
Gemma 3 12B
gemma3:12b
Won't fit 15.9 not known n/a not measured yet Estimated
DeepSeek-R1 Distill Qwen 14B
deepseek-r1:14b
Won't fit 13.7 not known n/a not measured yet Estimated
Phi-4 14B
phi4:14b
Won't fit 13.9 not known n/a not measured yet Estimated
Qwen3 14B
qwen3:14b
Won't fit 13.4 not known n/a not measured yet Estimated
gpt-oss 20B
gpt-oss:20b
Won't fit 16.5 not known n/a not measured yet Estimated
Gemma 3 27B
gemma3:27b
Won't fit 27.2 not known n/a not measured yet Estimated
Qwen3 30B (A3B)
qwen3:30b
Won't fit 22.6 not known n/a 9.3/10 Estimated
Qwen3 32B
qwen3:32b
Won't fit 26.3 not known n/a not measured yet Estimated
Llama 3.1 70B
llama3.1:70b
Won't fit 51.5 not known n/a not measured yet Estimated
gpt-oss 120B
gpt-oss:120b
Won't fit 70.5 not known n/a not measured yet Estimated

Models marked "thinks first" can write out their reasoning before they answer, so a job takes longer than the speed alone suggests.

With 32 GB of memory

About 28 GB of it can hold a model.

ModelFitNeeds (GB)Writing speed (tokens per second)SpeedQualityBasis
Llama 3.2 3B
llama3.2:3b
Fits 5 45 to 83 Fast not measured yet Estimated
Qwen3 4B
qwen3:4b · thinks first
Fits 6 36 to 67 Fast not measured yet Estimated
Gemma 3 4B
gemma3:4b
Fits 6.7 27 to 50 Fast not measured yet Estimated
Mistral 7B v0.3
mistral:7b
Fits 7.8 20 to 38 Fast not measured yet Estimated
Llama 3.1 8B
llama3.1:8b
Fits 8.3 18 to 34 Fast not measured yet Estimated
Qwen3 8B
qwen3:8b · thinks first
Fits 8.9 17 to 32 Fast not measured yet Estimated
Gemma 3 12B
gemma3:12b
Fits 15.9 11 to 21 Fast not measured yet Estimated
DeepSeek-R1 Distill Qwen 14B
deepseek-r1:14b · thinks first
Fits 13.7 10 to 18 Unclear could be Moderate or Fast not measured yet Estimated
Phi-4 14B
phi4:14b
Fits 13.9 9.8 to 18 Unclear could be Moderate or Fast not measured yet Estimated
Qwen3 14B
qwen3:14b · thinks first
Fits 13.4 9.6 to 18 Unclear could be Moderate or Fast not measured yet Estimated
gpt-oss 20B
gpt-oss:20b · thinks first
Fits 16.5 22 to 41 Fast not measured yet Estimated
Gemma 3 27B
gemma3:27b
Tight 27.2 5.3 to 9.8 Moderate not measured yet Estimated
Qwen3 30B (A3B)
qwen3:30b · thinks first
Fits 22.6 26 to 49 Fast 9.3/10 Estimated
Qwen3 32B
qwen3:32b · thinks first
Tight 26.3 4.5 to 8.3 Unclear could be Slow or Moderate not measured yet Estimated
Llama 3.1 70B
llama3.1:70b
Won't fit 51.5 not known n/a not measured yet Estimated
gpt-oss 120B
gpt-oss:120b
Won't fit 70.5 not known n/a not measured yet Estimated

Models marked "thinks first" can write out their reasoning before they answer, so a job takes longer than the speed alone suggests.

With 64 GB of memory

About 60 GB of it can hold a model.

ModelFitNeeds (GB)Writing speed (tokens per second)SpeedQualityBasis
Llama 3.2 3B
llama3.2:3b
Fits 5 45 to 83 Fast not measured yet Estimated
Qwen3 4B
qwen3:4b · thinks first
Fits 6 36 to 67 Fast not measured yet Estimated
Gemma 3 4B
gemma3:4b
Fits 6.7 27 to 50 Fast not measured yet Estimated
Mistral 7B v0.3
mistral:7b
Fits 7.8 20 to 38 Fast not measured yet Estimated
Llama 3.1 8B
llama3.1:8b
Fits 8.3 18 to 34 Fast not measured yet Estimated
Qwen3 8B
qwen3:8b · thinks first
Fits 8.9 17 to 32 Fast not measured yet Estimated
Gemma 3 12B
gemma3:12b
Fits 15.9 11 to 21 Fast not measured yet Estimated
DeepSeek-R1 Distill Qwen 14B
deepseek-r1:14b · thinks first
Fits 13.7 10 to 18 Unclear could be Moderate or Fast not measured yet Estimated
Phi-4 14B
phi4:14b
Fits 13.9 9.8 to 18 Unclear could be Moderate or Fast not measured yet Estimated
Qwen3 14B
qwen3:14b · thinks first
Fits 13.4 9.6 to 18 Unclear could be Moderate or Fast not measured yet Estimated
gpt-oss 20B
gpt-oss:20b · thinks first
Fits 16.5 22 to 41 Fast not measured yet Estimated
Gemma 3 27B
gemma3:27b
Fits 27.2 5.3 to 9.8 Moderate not measured yet Estimated
Qwen3 30B (A3B)
qwen3:30b · thinks first
Fits 22.6 26 to 49 Fast 9.3/10 Estimated
Qwen3 32B
qwen3:32b · thinks first
Fits 26.3 4.5 to 8.3 Unclear could be Slow or Moderate not measured yet Estimated
Llama 3.1 70B
llama3.1:70b
Fits 51.5 2.1 to 3.9 Slow not measured yet Estimated
gpt-oss 120B
gpt-oss:120b
Won't fit 70.5 not known n/a not measured yet Estimated

Models marked "thinks first" can write out their reasoning before they answer, so a job takes longer than the speed alone suggests.

With 128 GB of memory

About 124 GB of it can hold a model.

ModelFitNeeds (GB)Writing speed (tokens per second)SpeedQualityBasis
Llama 3.2 3B
llama3.2:3b
Fits 5 45 to 83 Fast not measured yet Estimated
Qwen3 4B
qwen3:4b · thinks first
Fits 6 36 to 67 Fast not measured yet Estimated
Gemma 3 4B
gemma3:4b
Fits 6.7 27 to 50 Fast not measured yet Estimated
Mistral 7B v0.3
mistral:7b
Fits 7.8 20 to 38 Fast not measured yet Estimated
Llama 3.1 8B
llama3.1:8b
Fits 8.3 18 to 34 Fast not measured yet Estimated
Qwen3 8B
qwen3:8b · thinks first
Fits 8.9 17 to 32 Fast not measured yet Estimated
Gemma 3 12B
gemma3:12b
Fits 15.9 11 to 21 Fast not measured yet Estimated
DeepSeek-R1 Distill Qwen 14B
deepseek-r1:14b · thinks first
Fits 13.7 10 to 18 Unclear could be Moderate or Fast not measured yet Estimated
Phi-4 14B
phi4:14b
Fits 13.9 9.8 to 18 Unclear could be Moderate or Fast not measured yet Estimated
Qwen3 14B
qwen3:14b · thinks first
Fits 13.4 9.6 to 18 Unclear could be Moderate or Fast not measured yet Estimated
gpt-oss 20B
gpt-oss:20b · thinks first
Fits 16.5 22 to 41 Fast not measured yet Estimated
Gemma 3 27B
gemma3:27b
Fits 27.2 5.3 to 9.8 Moderate not measured yet Estimated
Qwen3 30B (A3B)
qwen3:30b · thinks first
Fits 22.6 26 to 49 Fast 9.3/10 Estimated
Qwen3 32B
qwen3:32b · thinks first
Fits 26.3 4.5 to 8.3 Unclear could be Slow or Moderate not measured yet Estimated
Llama 3.1 70B
llama3.1:70b
Fits 51.5 2.1 to 3.9 Slow not measured yet Estimated
gpt-oss 120B
gpt-oss:120b · thinks first
Fits 70.5 19 to 35 Fast not measured yet Estimated

Models marked "thinks first" can write out their reasoning before they answer, so a job takes longer than the speed alone suggests.

Speed is the speed of writing an answer, shown as a range. Reading speed is not estimated. How the estimates work · How close they are · Back to the finder