GPUs / Qualcomm
Adreno 830
- Memory
- LPDDR5X-10600
- Bus
- Not published
- Published peak
- Not published
- Capacity
- from 12 GB
Measured
- Read
- 64.7GB/s
- Write
- 55.9GB/s
- Copy
- 61.1GB/s
- Random
- 0.44GB/s
- Latency
- 313ns
- Sustain
- 63.5GB/s
Medians of 1 verified run from 1 operator.
Local AI speed
| Model | Parameters | 4-bit, tokens/s | 8-bit, tokens/s |
|---|---|---|---|
| Llama 3.1 8B | 8.03B | ≤ 16 | ≤ 8.1 |
| Qwen 2.5 14B | 14.7B | ≤ 8.8 | Does not fit |
| gpt-oss-20b | 21B, 3.6B active | ≤ 36 | Does not fit |
| Mistral Small 24B | 24B | ≤ 5.4 | Does not fit |
| Qwen 2.5 32B | 32.5B | Does not fit | Does not fit |
| Llama 3.3 70B | 70.6B | Does not fit | Does not fit |
| gpt-oss-120b | 117B, 5.1B active | Does not fit | Does not fit |