Apple Silicon trains LLMs 3 to 4 times slower than similarly priced NVIDIA GPUs when memory fits, but can beat NVIDIA when VRAM is exceeded and ZeRO-Offload is required.
https://pytorch .org/get- started/previous-versions/, 2024
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.PF 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
support 1representative citing papers
citing papers explorer
-
Profiling Apple Silicon Performance for ML Training
Apple Silicon trains LLMs 3 to 4 times slower than similarly priced NVIDIA GPUs when memory fits, but can beat NVIDIA when VRAM is exceeded and ZeRO-Offload is required.