Some users report local setup already justified by hybrid approach
5 Sep 14 10:54 PM · 13d ago · 5 comments · 1 source · development 5 of 5
Commenters describe owning local hardware for specific tasks (automation, sensitive data, personal wikis) while maintaining separate API subscriptions for other work—a pragmatic split that sidesteps the pure cost-payback question.
“I have a local model monitoring my finances and personal wiki - things I wouldn't want Claude to touch - and the Qwen 3.5 9b handles it all just perfectly.”
jrecyclebinrlindsey123 Developer
The whole story articlespostscomments the bright band is this development · numbered dots are the others · click one to jump
What people said 11 voices · verbatim
-
I'm curious what people are sending to Claude that is so secret. Claude knows about my interior decorating, questions about light bulbs, curiosity about what the Galactic Empire was even trying to do, unpacking SCOTUS decisions, shoe trees, Fed inflation history, etc.What part of my brain is contained here? Sure, the conversations have back and…
-
Also, whatever your doing won't be at the whims of cloud providers; it won't fail because they decided to quantize your customer $ into a shittier model.Some how, _instability_ has gained valuable currency, so now we all act like the constant change of whatever is actually good for us. FOMO is just like breathing guys. That anxiety induced by tech…
-
The math is wrong, the tok/s is at least 2x that, at least with MTP and Q8 KV which you should always use. And the default tokens a day is ridiculously low at least for coding.Having said that, it will never pay for itself. A simpler more absolute math is, if I buy a Mac and use it to sell tokens on OpenRouter, will I make a profit? And the answer…
-
> If I am offloading some of my thought processes to a machine"Offloading thought" sounds a lot better than "outsourcing thought", but the latter is what we're really doing. Offloading implies you thought it first and then gave it to the LLM, but we're only giving it the minimun so it can do most of the work in our place,
-
On a purely monetary basis it probably never will.You're competing against companies that get tax breaks, locate themselves optimally, and have large economies of scale.Also, if it did, the hardware would be bought up, raising the price until there was no economic profit again.If you can find a unique application for it then maybe?
-
This was my thought as well. I have a local model monitoring my finances and personal wiki - things I wouldn't want Claude to touch - and the Qwen 3.5 9b handles it all just perfectly.I also needed a new device anyway - and having this much system memory to run virtual machines has been amazing.Am paying subscriptions as well tho lol.
-
I want the autonomy but local models of the size I would have the means to host wouldn't be capable enough. What usecases tend to suit these smaller models that tend to produce incorrect or otherwise flawed responses often? Could they work for anomaly detection and what would a rough architecture look like?
-
The real question should be why anyone would voluntarily continue to spend money on a software service that costs as much as an expensive computer when they could just buy an (upgradeable) expensive computer and use it as much as they want, approximately forever.
-
This tells me that the max throughput for the models I'm running on my hardware is lower than it actually is. Please allow us to tweak all the variables instead of locking me in to whatever rate you found by searching
-
For me, I'm glad they train on my stuff if it improves the model. Hell, I've been using tons of muse-spark-1.3-contributor for this very reason (and because it's a decent model for a bargain basement price)
-
Agree. As the meme/old-ad goes, "Running it on my own machine? Priceless!"Some of us get a weird thrill that we can actually do this. Mind-boggling time we live in.
All 5 developments of Sunk Cost calculator shows when local LLM hardware pays for… →
Hacker NewsMastodonNewswires