35 points rlindsey123 1 hour ago 56 comments
Obviously there are other reasons to buy your own hardware aside from just saving money on llms but this is just looking at it from a raw cost saving perspective.
If you have any ideas on how I can make this more helpful lmk!
hyperhello 35 minutes ago | parent
gruez 26 minutes ago | parent
-- Warren Buffett
taraindara 13 minutes ago | parent
rlindsey123 13 minutes ago | parent
tyre 7 minutes ago | parent
I don't know when we'll have an open equivalent to Fable, let alone whatever (insane) hardware you'd need to run it locally.
jrflo 34 minutes ago | parent
rlindsey123 10 minutes ago | parent
itake 34 minutes ago | parent
Mac mini can also build iOS applications. I think if you’re a mobile dev, you can have concurrent builds for your agents instead of everyone waiting on a single machine to finish.
rlindsey123 15 minutes ago | parent
ProjectArcturis 31 minutes ago | parent
mcone 30 minutes ago | parent
usernomdeguerre 6 minutes ago | parent
txrx0000 29 minutes ago | parent
no-name-here 22 minutes ago | parent
txrx0000 18 minutes ago | parent
selectodude 14 minutes ago | parent
koito17 12 minutes ago | parent
In May of this year, I was running qwen3.6:35b-a3b on my MacBook (bought in 2024). Obviously not as fast as, say, running a model on Cerebras, but a year ago it wasn't really feasible to have a local model running on my 2024 laptop with vision support. (Concretely, I was passing apartment diagram pictures to Qwen and making it compare different apartments for which ones would feel the most spacious while optimizing for initial moving costs and other factors.)
This was back in May and I wouldn't be surprised if there have been significant improvements since then.
Overall, I think it's fair to compare a workflow like "use llama.cpp locally to upload some pictures and ask questions" to "open the ChatGPT app, upload pictures from your phone, and ask questions". Sure, you can't run a model like GPT-5.4 locally, but the model is mostly an implementation detail here. What a user will care about is: "when I go with the llama.cpp option, am I getting useful information from my conversations?"
no-name-here 6 minutes ago | parent
tyre 15 minutes ago | parent
What part of my brain is contained here? Sure, the conversations have back and forth (some have dozens of exchanges), but, like, that's not the secret to me. I don't think it can replicate me, and even if it could… okay?
Are you worried they're going to target ads? That the government will steal something? What?
Claude Code has information about my home server, but google or DDG would also have the broad strokes (torrents). I don't know. Maybe others are working on more sensitive things at home.
poincareball 14 minutes ago | parent
octoberfranklin 12 minutes ago | parent
The proof to the Navier-Stokes problem.
truncate 11 minutes ago | parent
Its the same point used against privacy. What's so secret you are doing that you need privacy. I think in the end, its about privacy and not trusting these model companies with your data. Facebook manipulated people behaviors with all the data they had, no reason AI companies wont someday decide to do that same, and they have far more intimate knowledge.
When it comes to coding, I also don't like the idea of them taking my money and potentially at same time potentially using as dataset generator.
jrecyclebin 14 minutes ago | parent
I also needed a new device anyway - and having this much system memory to run virtual machines has been amazing.
Am paying subscriptions as well tho lol.
catchnear4321 8 minutes ago | parent
Local isn’t strictly about NOT lab. It’s rapidly becoming apples (though not just macs) to oranges to compare the to.
Which is why the premise is silly. To be underwater it would need to be a real comparison. It’s not, and the claude fartifact doesn’t make it so.
throwaway894345 13 minutes ago | parent
ChickeNES 6 minutes ago | parent
shadowpho 25 minutes ago | parent
rlindsey123 16 minutes ago | parent
chasd00 23 minutes ago | parent
Idk about the quality of this setup but just pasting it here as an example. https://explainx.ai/blog/heretic-llm-abliteration-guide-2026
ChickeNES 19 minutes ago | parent
When does the average person actually need to do that?
jerf 11 minutes ago | parent
I expect this is only going to get worse. "Censorship" isn't just going to be about who you vote for and which political party the model will say nice things about and which it is more likely to say bad things about. It's going to become about whether the hoi polloi are allowed to have effective AIs at all. Like the 1990s internet, AI has outrun a lot of power structures but that is not going to continue indefinitely.
ChickeNES 5 minutes ago | parent
beachy 10 minutes ago | parent
So I can certainly understand why someone would want the guardrails gone.
kees99 5 minutes ago | parent
Case in point, last week I was poking Opus 5 into writing me some RPi-pico firmware for driving a small e-paper screen. Font was built in right into C code as hex constants. Space being tight, I asked if there is some clever compression that could be applied. Claude thought for good 10 minutes, then guardrail kicked in telling me that was "cyber", and refused to continue.
ThunderSizzle 22 minutes ago | parent
I paid $1350 and threw an R9700 in an existing machine. That's a 4 month pay off or so.
Plus, I can feed it sensitive data all day and not be worried where it's going.
no-name-here 18 minutes ago | parent
bix6 21 minutes ago | parent
QwenGlazer9000 19 minutes ago | parent
The premium is not having your million dollar prize and career stolen by billionaires.
rlindsey123 14 minutes ago | parent
monksy 15 minutes ago | parent
rlindsey123 11 minutes ago | parent
bpbp-mango 13 minutes ago | parent
gfody 12 minutes ago | parent
serial_dev 10 minutes ago | parent
01HNNWZ0MV43FF 10 minutes ago | parent
harhargange 9 minutes ago | parent
Zetaphor 7 minutes ago | parent