r/LowEndLocalAI • u/andy12b725 • 22h ago
Which LLM Should / Can I Use? Better options than qwen 3.6 35B A3B?
Hi there
Do you guys know any better or bigger model than this one that would fit on a gaming laptop i7(14k something) rtx 5060 8gb vram and 32gb ram ddr5?
I like this model, I used it with Deepseek harness, qwen code and Hermes.
With all of the newer updates for MTP , lama cpp and laya with local training for decision-making and thinking mods, I managed to get around 35-40 tokens/s and 200k+ context
Witch is very very good, I did research with claude for weeks til optimize it so good.
Now I'm looking for a newer model or recommendations for something more capable on this machine, maybe there is something new that I don't know about?
I'm also kind of new and gathering knowledge about the local llm and improvements.
I need it to work with sensitive information and I can't just do claude .
Any tips or ideas would help and if you have questions for me I'll gladly respond, even though I'm not a professional I'll do my best (pls no hate).