GPUs / Apple
Apple M1
- Memory
- LPDDR4X
- Bus
- 128-bit
- Published peak
- 68.25 GB/s
- Capacity
- 8 GB
Measured
- Read
- 51.0GB/s
- Write
- 52.5GB/s
- Copy
- 52.1GB/s
- Random
- 3.82GB/s
- Latency
- 373ns
- Sustain
- 51.5GB/s
Medians of 1 verified run from 1 operator. Read reaches 75 % of the published peak.
Local AI speed
| Model | Parameters | 4-bit, tokens/s | 8-bit, tokens/s |
|---|---|---|---|
| Llama 3.1 8B | 8.03B | ≤ 13 | Does not fit |
| Qwen 2.5 14B | 14.7B | ≤ 6.9 | Does not fit |
| gpt-oss-20b | 21B, 3.6B active | Does not fit | Does not fit |
| Mistral Small 24B | 24B | Does not fit | Does not fit |
| Qwen 2.5 32B | 32.5B | Does not fit | Does not fit |
| Llama 3.3 70B | 70.6B | Does not fit | Does not fit |
| gpt-oss-120b | 117B, 5.1B active | Does not fit | Does not fit |