GPUs / Intel

Iris Plus

Memory
LPDDR4X / DDR4
Bus
128-bit
Published peak
59.7 to 68.3 GB/s
Capacity
from 8 GB

Measured

Read
20.9GB/s
Write
19.5GB/s
Copy
21.4GB/s
Random
3.10GB/s
Latency
841ns
Sustain
22.6GB/s

Medians of 1 verified run from 1 operator.

Local AI speed

ModelParameters4-bit, tokens/s8-bit, tokens/s
Llama 3.1 8B8.03B≤ 5.2Does not fit
Qwen 2.5 14B14.7B≤ 2.8Does not fit
gpt-oss-20b21B, 3.6B activeDoes not fitDoes not fit
Mistral Small 24B24BDoes not fitDoes not fit
Qwen 2.5 32B32.5BDoes not fitDoes not fit
Llama 3.3 70B70.6BDoes not fitDoes not fit
gpt-oss-120b117B, 5.1B activeDoes not fitDoes not fit
Upper bounds: median read ÷ bytes read per token. Mixture-of-experts models read only their active experts. Fit checks weights against the smallest 8 GB configuration; the KV cache needs room on top.

Top operators

#OperatorScoreReadDate
1OP-55FEA816.920.9 GB/s2026-09-27

Full Iris Plus leaderboard