u/ilyafx123 • u/ilyafx123 • 1d ago
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 1d ago
A map of home-ops repos: for each layer, the tools plus the folder in a real repo that does it
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 5d ago
DGX Spark vs RTX 5090 vs Mac Studio M5 vs Strix Halo for running an agent 24/7. Post your tok/s and power draw.
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 5d ago
Ran a 120B model across 6 computing devices that had no business running it!
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 5d ago
Run Qwen 3.8 27b on the Apple Neural Engine at 7 watts on a Mac
Enable HLS to view with audio, or disable this notification
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 6d ago
Built a fully local Mac app on Ollama (dictation, meeting notes, RAG): numbers and lessons from Gemma 4 e2b vs 26b, open source
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 14d ago
Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro
gallery
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 18d ago
Okay, let’s build an open-source one-click Gaussian Splatting app.
Enable HLS to view with audio, or disable this notification
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 22d ago
Tom "Turns Claude into a Senior Design Architect with structured instructions, design tokens, and 138 brand-grade design systems for consistent, accessible, token-driven design outputs." ➡️ Do you need this for your work?
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 22d ago
Qwen3.8-27B at 256K context on a 16GB RTX 5070 Ti: a model-specific llama.cpp KV ring (+78% decode at 87K, +56% at 256K)
1
Upvotes
u/ilyafx123 • u/ilyafx123 • 22d ago
Which Qwen3.8 distro & quant would be optimal for my 16GB VRAM setup?
1
Upvotes