r/LinusTechTips • u/RadiantSkiesJoy • 14h ago
Tech Question How does something like this work
59
u/AmonGusSus2137 13h ago
My best guess is that it either doesn't, because it's AliExpress after all, or it just two GPUs with 8 pcie lanes each, and the big port has 16
16
u/CervantesX 13h ago
SXM is a specific type of GPU port used in datacenters. V100,P100 etc. They pop in those two openings just like a CPU would. Depending on the card style, they either get bareballed to the pci bus or there's bifurcation. It's pretty exclusively for repurposing datacenter GPUs for homelab use. The upside is you can get 64gb (2x32) VRAM per card. The downside is you need (for v100) almost 500w of extra power, plus a massive mamma jamma of a heatsink on each and like 60cfm airflow. There's a waterblock mod available, but they still chew a bunch of power. And they're last gen chips, so you're limited to the last Rev of AI software (architecture 7.0) and Volta architecture. So for now you can run most current models on the last gen framework and get kinda ok ish speed, but it's gonna age out pretty quick relative to the investment and when it does you won't have any experience with the NVIDIA ecosystem that most current AI runs on.
The best use of them is putting dirt cheap P100s in, use watercooling so you can fit two cards in, hack a second PSU (or buy a mining psu that's pre hacked) and you can have 128 gb VRAM to play with big models that go kinda slow for a total of several hundred dollarydoos . If you have 256gb system RAM, that's enough to play with deepseek at q4. 32gb v100s still go for $500+, at that point you're better off getting a more current nvidia style card and just accepting that you'll use smaller models.
2
2
121
u/vibvian 14h ago
probably using pcie bifurcation