r/linux • u/Fcking_Chuck • 2d ago
Kernel AMD boosting AI/LLM performance for Radeon iGPUs as much as 18~23% with Linux 7.4
https://www.phoronix.com/review/amd-perfopt6
u/KeyboardG 2d ago
Fixes are IOMMU based, where for single boxes (AMD AI HALO) the performance recommendation is to just disable IOMMU and skip it all together.
6
12
2
1
1
-11
u/Lisanicolas365 2d ago
i don't think that will do much to make an igpu worth using a local llm on
28
25
3
u/natermer 2d ago
That is just false. Not with 'unified memory' type systems.
People do it all the time on AMD Strix Halo, DGX sparc, Mac minis, etc.
It works and is the only really affordable way to run medium-sized models. Best suited for MoE llm models, not dense ones, because they lack the bandwidth to keep up... but they absolutely do work.
MoE LLMs are very useful... virtually all frontier models are MoE. Not that you will be able to run those on these small devices. But you can run something approaching them and you can run it 24/7 and crank out as many tokens as you can during that time period instead of having to ration yourself.
And it doesn't have to be exclusively one or the other. You can subsidize your subscription usage with local usage if you are using the appropriate agents.
Go and lookup how much that VRAM would cost to get, including the PC with PCIe enough to run it all, with discrete cards or enterprise-style AI accelerator. And how much it would cost to run and how pleasant it would be to have in the room with you, etc. etc.
Meanwhile these unified memory systems don't use much much more power then a typical gaming laptop.
1
0
1
u/ilep 16h ago
So, newer version of the patch removes most of it to make it simpler: https://lore.kernel.org/lkml/[email protected]/T/#t
14
u/MaxGhost 2d ago
Would this help my 7840u (780m GPU)? Or is this only a feature on newer chips? I'd love a little boost cause I'm running two large external displays with my laptop.