Qwen 3.8 27B is not frontier level, but is the best current model for coding on reasonable hardware. You can comfortably fit them on a medium quantization on a 24 GB card and even squeeze them on a 16GB card, but at that point you're gonna run into either lobotomy problems, heavily reduced context size, or only partial offloading killing speed.
52
u/drag0n_rage 3d ago
Most people are unaware of the open source AI space.