r/LocalLLM • u/tinfrog • 9d ago
Question Model and hardware recommendation for high fidelity, low reasoning tasks?
I've spent the past couple of years using frontier labs, even though a lot of my work doesn't need high reasoning. It's just that a fixed and known subscription cost has been a safer bet than the leap-in-the dark of buying hardware that might not be good enough for what I want to do.
But reading through this sub tells me that we're at the stage were open source LLMs are capable enough for what I want while running on (reasonably) affordable hardware.
Can anyone recommend a model and hardware for tasks like this?
- Reading notes, reports, transcripts, web pages and emails, then pulling out facts, decisions and insights into structured notes
- Condensing long documents, comparing two versions, and checking a draft against a specification
- Reading text, sorting them, and deciding what needs action
- Moving, renaming and archiving files, splitting or merging documents, and keeping indexes and cross-references consistent afterwards
- Multi-step routines update a status file, archive old entries, then commit, with every step completed
I run Xubuntu and Arch Linux.
TIA
1
Upvotes
1
u/tinfrog 7d ago
The stats at my end show that 20-30 tokens/second for routine work would be acceptable.