00:00
2026-09-25
rocm.blogs.amd.com
large-language-models
Model Weight Profiles: Where Do the Parameters Go?
AMD detailed model weight profiles that break down how many parameters go to embeddings, attention, and dense layers, explaining why the amd/Llama-3.1-8B-Instruct-FP8-KV checkpoint is 9.08 GB rather t…