This aged well. You could’ve bought a DGX Spark for $3500-4000, and Codex's $200 plan was actually pretty generous. 3 month later we have a 64 GB DGX Spark retailing for $4950, the 128 GB retailing for $6950, and a new $500 Codex plan. How will it look like in 1 year from now?
"It just doesn’t math." I keep seeing this take. But why run AI on your own hardware? Because for many of us, the cloud just doesn’t math either. Personally, I replaced all my Claude + Codex usage with DeepSeek-v4-Flash running locally on two DGX Sparks. For reference: $8,000 spent on the Claude Sonnet API would currently buy you roughly 800 million output tokens — about 5 months of continuous generation at 60 tokens/sec. The "It just doesn't math" claim: "For $8000 + electricity you could get over 4 years of Claude Max $200/mo sub plans, which would give you more Sonnet usage than your local setup." Who says Claude Max stays at $200/mo? They’re literally losing money on every sub right now — this price won’t last. They also nerf the limits constantly, & you’re STILL rate-limited even on the top tier. I know you can’t run actual Sonnet locally. That’s not the point. The point is: for my workflows, the local models I can run are good enough to fully replace it. A lot of people (including me) simply prefer not to send work through certain cloud providers, whether for privacy, trust, or other reasons. So the real question is: if the hardware can replace what you’re already paying for in API costs, is $8k justifiable? For me, it absolutely is. Plus… you actually own it. You don’t like owning things?

Oct 3, 2026 · 11:14 AM UTC

32
21
230
36,031
Sort replies: Relevant Recent Liked
Replying to @MiaAI_lab
So are you now doing your primary work with local or cloud AI?
1
3
897
Both
1
712
Replying to @MiaAI_lab
I got a small cluster of "cheap" 16GB VRAM gaming systems in case local models can what I need or if I'm getting completely priced out by subs
21
Replying to @MiaAI_lab
true, and then companies claiming that tokens price had dropped significantly, but pattern show otherwise
326
Replying to @MiaAI_lab
I'm gutted I missed out on what now seems "cheap" hardware.
359
Replying to @MiaAI_lab
por eso es tan importante hacer correr un LLM decente en una RTX3050. Hay que lograr que las GPU de gaming hagan el 80% del trabajo. @ashxhart Si, ya sé que lo "cool" es M5 pero es el 0.00001% del parque hw en los hogares.
175
Replying to @MiaAI_lab
I ordered a binned M5 ultra 256gb today. As a SWE I needed to lock in prices now before the subsidized subscriptions end.
1
70
Replying to @MiaAI_lab
I don't think those prices are coming down either, those numbers are going to keep going up
249
Replying to @MiaAI_lab
The additional cost up front for your own hardware it’s going to seem like chump change once these corporations want to actually start making money from retailers.
2
291
Replying to @MiaAI_lab
Post tagged. And those are NOW prices. I’ll revisit this in 6 months.
2
361
Replying to @MiaAI_lab
It’s all going to feel that much worse after the IPOs and end of subsidization
4
437
Replying to @MiaAI_lab
Secure your compute. It won’t be enough, but get what you can folks
139
Replying to @MiaAI_lab
Whats a good argument against buying 1tb ddr4 right now for $8k-$11k and just figuring out what to do with it?
34
Replying to @MiaAI_lab
Ordered a Spark yesterday 🫡
230
Replying to @MiaAI_lab
Competition from the frontier and open-source near-frontier as well as the ability to shrink models and keep intelligence will keep plans worthwhile for many users, but we’ll probably still see spikes of Einstein-tier models that are extremely expensive for a few months at a time
188
Replying to @MiaAI_lab
You forgot to mention how unbelievably slow gpt-6.1-sol is right now too. 14-20 tok/s is unbearable. Fuck dots. I've switched to Qwen3.8-flash-next for most things, 100+ tok/s on my 5080. I wish I had better hardware for local, but this is getting what I needed done just fine.
112
Replying to @MiaAI_lab
I do that math per decision. My 0.5B extractor scores inline on a local GPU in about 106ms, so routine picks cost nothing past the card. Subscriptions earn it back on calls that need frontier quality. Where do most of yours land, local or frontier?
165
Replying to @MiaAI_lab
At this point we need start owning our own compute or prices will start getting ridiculous.
30
Replying to @MiaAI_lab
El incremento de precio en el hardware de consumo y entrada para IA en pocos meses demuestra que la demanda supera con mucho a la oferta inicial de los fabricantes.
53
Replying to @MiaAI_lab
We will own nothing and be happy about it.... 🥺
1
180
Replying to @MiaAI_lab
Well said local AI it's winning in so many ways it's not even comparable, the only thing that it's already getting fixed is having the best open source model, beating closed ones
39
Replying to @MiaAI_lab
The 64 GB jump hurts the single-box entry more than dual-rig TCO. Against the new $500 Codex plan, where's your 1-year break-even now for heavy codegen?
1
280
Replying to @MiaAI_lab
I will cling to my 3090 ... and hunt for more. Cloud inference getting more expensive ... timed with private inference getting more expensive too. Suspiciously bad smelling! But it opens opportunity for "alternative new players" ... lets see if they can take advantage!
58
Replying to @MiaAI_lab
I’d also count the value of keeping a model version that already works for a job. Local for repeatable tasks, cloud for the rest. That gives you room to try new models without changing every workflow at once.
1
92
Replying to @MiaAI_lab
I did, followed your advice, thanks!
3
274
Replying to @MiaAI_lab
It looks like people cobbling together their own local lab, from various and diverse parts. The Spark was too much when a month was $20. Now it pays for itself inaa significantly shorter time.
27
Replying to @MiaAI_lab
Subscription based hardware at “ownership” pricing
71
Replying to @MiaAI_lab
The hardware price rises fast. Software value stays high. Smart choice to use the code now.
173
Replying to @MiaAI_lab
64KB is enough!
65
Replying to @MiaAI_lab
I dont know what lind of price surges we have coming on the mac studios
65