back
253 comments
This is a good time to promote running your own models. I have been running my own models locally and I would wager a local model will meet 85-95% of your needs if you really learn to use it. These models have gotten great. For anyone wanting to get into this, the smartest models to run recently that is consumer friendly was just released, checkout Qwen3.5 the 27B and 35B variants. They are small and I recommend running full Q8 quants. The easiest way to run these without dealing with complex GPU is to get a mac. For the example I gave, a 64gb mac will handle it well. If you are really cash strapped then you can manage with a 32gb but will have to run with less resolution quants. If you are not cashed strap, then get at least a 128gb and if possible a 256gb. The models are so good you will regret not getting a better system. You can join the r/LocalLlama community in reddit to learn some more. But this is pretty easy. Grab llama.cpp, grab a gguf quant from huggingface.co - the unsloth quants are great - https://huggingface.co/unsloth/models
For non-Mac users:

A laptop with an iGPU and loads of system RAM has the advantage of being able to use system ram in addition to VRAM to load models (assuming your gpu driver supports it, which most do afaik), so load up as much system RAM as you can. The downside is, the system RAM is less fast than dedicated GDDR5. These GPUs would be Radeon 890M and Intel Arc (previous generations are still decently good, if that's more affordable for you).

A laptop with a discrete GPU will not be able to load models as large directly to GPU, but with layer offloading and a quantized MoE model, you can still get quite fast performance with modern low-to-medium-sized models.

Do not get less than 32GB RAM for any machine, and max out the iGPU machine's RAM. Also try to get a bigass NVMe drive as you will likely be downloading a lot of big models, and should be using a VM with Docker containers, so all that adds up to steal away quite a bit of drive space.

Final thought: before you spend thousands on a machine, consider that there are at least a dozen companies that provide non-Anthropic/non-OpenAI models in the cloud, many of which are dirt cheap because of how fast and good open weights are now. Do the math before you purchase a machine; unless you are doing 24/7/365 inference, the cloud is fastly more cost effective.

And if you don't want to buy a Mac? A 80 GB NVidia GPU costs $10,000K (equivalent to 30 years of ChatGPT Plus subscription) and will probably be obsolete in 5-7 years anyway. What are my options if I want a decent coding agent at a reasonable price?
An even easier way to get into this is simply by downloading a program called LM Studio. You can mount a model and chat to it within 10-15 mins with no experience whatsoever, and no configuration at all.

That said, last time I tried local LLMs (around when gpt-oss came out) it still seemed super gimmicky (or at least niche, I could imagine privacy concerns would be a big deal for some). Very few use cases where you want an LLM but can't benefit immensely from using SOTA models like Claude Opus.

The financial barrier is kind of the opposite of "easy to run" to me.

As much as I love owning my stack, you'd have to use so much of this to break even vs an inference provider/aggregator with open frontier-ish models. (and personally, I want to use as little as possible)

As someone who desperately wants to use local models, I lament there is no way to use them on consumer hardware for serious coding work. I have a rtx 4070 super ti and I cannot run any large model with enough context and tps compared to a remote offering.
I have a 24GB Macbook Pro. I will note, do get the 'Pro' models, the Mac Mini and the Macbook Air do not have internal fans. The Macbook Pro has an internal fan, and the Mac Studio (bigger Mac Mini) has a fan. If you get a Mini, you might want to get one of those docks that cools the Mini. Your hardware will get very hot very quickly.

Also, because Apple in their infinite wisdom despite giving you a fan, very lazily turn it on (I swear it has to hit 100c before it comes on) and they give you zero control over fan settings, you may want to snag something like TG Pro for the Mac. I wound up buying a license for it, this lets you define at which temperature you want to run your fans and even gives you manual control.

On my 24G RAM Macbook Pro I have about 16GB of Inference. I use Zed with LM Studio as the back-end. I primarily just use Claude Code, but as you note, I'm sure if I used a beefier Mac with more RAM I could probably handle way more.

There's a few models that are interesting on the Mac with LM Studio that let you call tooling, so it can read your local files and write and such:

mistralai/mistralai-3-3b this one's 4.49GB - So I can increase my context window for it, not sure if it auto-compacts or not, have only just started testing it

zai-org/glm-4.6v-flash - This one is 7.09GB, same thing, only just started testing it.

mistralai/mistral-3-14b-reasoning - This one is 15.2GB just shy of the max, so not a TON of wiggle room, but usable.

If you're Apple or a company that builds things for Macs or other devices, please build something to help with airflow / cooling for the MBP / Mac Mini, it feels ridiculous that it becomes a 100c device I'm not so sure its great for device health if you want to use inference for longer than the norm.

I will probably buy a new Mac whenever the inference speeds increase at a dramatic enough rate. I sure hope Apple is considering serious options for increasing inference speed.

Or you can get a strix halo from AMD. They run about $2k from various Chinese brands, or a bit more from Framework. 128GBs of unified RAM are plenty for most models, although memory bandwidth is slower than in a mac.
I really hope at some point in the near future AI models shrink enough or laptops get strong enough to run AI models locally. I haven't tried in the past year, but when I did it was very slow token output + laptop was on fire to make that happen.

I've wanted to try some of the more recent 8B models for local tab completion or agentic, any experience with those kinds of smaller models?

I have a lenovo workstation with 256GB ram but a weak sauce 12GB VRAM GPU. Is there any DMA trick to improve offload performance?
I just can't help but imagine ChatGPT's sycophancy mixed with military operations. "Sharp insight bombing that wedding! Next would you like tips on mosques to bomb, or I can suggest some new napalm recipes that are extra spicey. Your call!"
Story time!

I actually cancelled my ChatGPT subscription in late 2024 and documented the process, kind of as a social media thing because it had gotten so bad and I realized nobody in my family was using it anymore. I asked my wife if she was getting any use out of it and she told me she had been using Gemini and Grok for months because "GPT is very lazy now".

After a while another charge came in for the subscription, but I had the receipts: we had cancelled before the next billing cycle. I decided to try and reach out to OpenAI to resolve this, but they only let you chat with GPT itself for this, which it failed at and told me they weren't in the wrong and none of the information matched what actually happened.

I took this and used it to submit a chargeback request with Privacy.com, which I use for all of my online purchases. Normally I don't have to worry about this because I set a limit or cancel the cards I issue manually, but I had an OpenAI API account using the same card and I had been a bit lazy in using the same card for technically two different services.

Well, Privacy.com won that dispute and I got that money back. It's worth mentioning this is actually different than most banks will do now days. For the most part when you try to get a bank to do a chargeback they just roll it into their insurance and refund you the customer as a cost of doing business, but the actual scammer or shady merchant got to keep their stolen money, whereas I can be certain OpenAI didn't keep my money.

I had been considering ditching everyday ChatGPT use in favor of Claude anyway, but hadn’t gotten around to it mostly out of habit. Now I have a good reason to do it.
Before you fully delete your account, don't forget to first save your chats! Go to https://chatgpt.com/#settings/DataControls and click Export under "Export Data".
I've just cancelled my subscription in solidarity with the OpenAI employees who signed the We Will Not Be Divided letter. I was a daily user of paid features like Deep Research. But not only was Anthropic's decision more ethical, their products are better, so I can't possibly justify the expense. Honestly I mostly was subscribed to take pressure off of my Claude usage limits, but I've just upped my Claude subscription to the next tier instead.

ETA: I've started an export of all my data. After that's done, I'm going to delete it all from my account (Settings > Data controls) and walk away from the account. I will give this to OpenAI, they make the process of disentangling yourself straightforward and there's integrity in that.

I love that the tool in question is very calm and collected, in contrast with the emotional wreck that is the US regime. I got a very helpful response to this prompt and I will make it continue working on a python script to get my historical chats looking good in Obsidian.

> Ok. So I'm cancelling the subscription to ChatGPT and moving over to Claude because of the news of OpenAI striking a deal with us department of war. (https://www.techradar.com/pro/openai-just-signed-a-huge-deal...) Please line out a good exit strategy where I can keep the information in my chats and projects on my own hard drive.

Just deleted my account. Can always sign up for a new account later if you need (with a different email).
Just cancelled. I’ll give my money to a company with leaders that have a modicum of backbone.
A few days ago I went to cancel mine and it just said they'd give me a free month instead so I said OK. I thought it was funny all the patterns to keep you on
I'd cancelled my subscription earlier this month organically as I wasn't getting any net positive value.

BTW, what's going to hurt their business more, deleting my account or using the free tier?

I'n sorry, Dave ...
I'm gonna have to see if I can get my company to switch off openAI. Hopefully we can make a small dent and if enough of us do it, a larger dent.

Sounds like it won't really be a pain for me though based off comments on HN indicating Claude is the better product and I doubt I personally would hit any sort of token limits with the amount I use agentic coding.

I just cancelled my ChatGPT subscription that I had since 2023. OpenAI offered me an extra month free, to keep my subscription.
PSA: If you can't switch your coding agent right away, you can just reroute Codex to a different model for the time being.

https://github.com/openai/codex/issues/26#issuecomment-28116...

ChatGPT AI Chatbot Helps Crypto Scam Victims Recover & Stay Safe

Intelligence Cyber Wizard, mission is simple: help victims of cryptocurrency scams recover their digital assets and empower them with the knowledge to prevent future losses. How to Recover Lost Cryptocurrency If you’ve been scammed, act quickly and follow these professional steps: 1. Secure Your Remaining Assets Immediately transfer remaining funds to a new secure wallet. Enable two-factor authentication (2FA) on all exchanges and email accounts. Change all passwords. 2. Document Everything Save wallet addresses involved. Take screenshots of conversations with the scammer. Keep transaction IDs (TXIDs), payment confirmations, and exchange receipts. 3. Blockchain Transaction Tracing Cryptocurrency transactions are recorded on the blockchain. Through advanced forensic tools, suspicious wallet movements can be traced across exchanges and mixing services. 4. Exchange & Platform Notification If funds moved through major exchanges, immediate reporting increases the chances of freezing suspicious accounts. 5. Legal & Regulatory Reporting Report to: intelligencecyberwizard@cyber-wizard.com (United States) Timely reporting strengthens recovery efforts. How Intelligence Cyber Wizard Educates Crypto Scam Victims Education is prevention. They don’t just assist with recovery — they equip victims with long-term protection strategies. 1⃣ Scam Awareness Training They educate victims about: Fake investment platforms Romance crypto scams Phishing wallet attacks Impersonation scams Fake recovery agents 2⃣ Wallet Security Education Cold vs hot wallet protection Private key safety practices Identifying malicious smart contracts 3⃣ Blockchain Transparency Lessons Victims learn how crypto transactions work, why they are traceable, and how scammers attempt to launder funds. 4⃣ Red Flag Identification They teach clients to identify: Guaranteed high returns Pressure tactics Fake celebrity endorsements Unregulated trading platforms Their Commitment as Intelligence Cyber Wizard, they operate with: Professional digital forensic methods Ethical recovery strategies Confidential case handling Victim-first support system Cryptocurrency scams are rising globally, but with the right response and education, recovery and prevention are possible.

I thought it was only me. I just unsubscribed it this morning.
How long does data export usually take for three years of medium usage? I started it eight hours ago, got a confirmation email that export had started but so far no email with a download link.
After the "upside down cup" debacle, and the "walk vs drive to the carwash" conundrum, and so many other examples where GPT 5.2 thinking failed miserably and Opus 4.6 and (even Sonnet 4.6 extended thinking) nailed it, I think they earned people wanting to cancel their subscription regardless of yesterday's events.
Deleted.

Anthropic usage credits purchased.

Message those that work forces.

In case you decide to delete your account, I'd recommend to at least download the saved memories - https://chatgpt.com/#settings/Personalization
OpenAI has ~50M paying subscribers driving >$10B in revenue.

You would probably need at least ~1M subscribers to cancel to make this painful.

Probably needs more attention outside of tech circles for that to happen but I suspect this will get drowned out in the face of other stuff.

Learn How ChatGPT AI Helps Crypto Scam Victims Make Informed Decisions

Intelligence Cyber Wizard, mission is simple: help victims of cryptocurrency scams recover their digital assets and empower them with the knowledge to prevent future losses. How to Recover Lost Cryptocurrency If you’ve been scammed, act quickly and follow these professional steps: 1. Secure Your Remaining Assets Immediately transfer remaining funds to a new secure wallet. Enable two-factor authentication (2FA) on all exchanges and email accounts. Change all passwords. 2. Document Everything Save wallet addresses involved. Take screenshots of conversations with the scammer. Keep transaction IDs (TXIDs), payment confirmations, and exchange receipts. 3. Blockchain Transaction Tracing Cryptocurrency transactions are recorded on the blockchain. Through advanced forensic tools, suspicious wallet movements can be traced across exchanges and mixing services. 4. Exchange & Platform Notification If funds moved through major exchanges, immediate reporting increases the chances of freezing suspicious accounts. 5. Legal & Regulatory Reporting Report to: intelligencecyberwizard@cyber-wizard.com (United States) Timely reporting strengthens recovery efforts. How Intelligence Cyber Wizard Educates Crypto Scam Victims Education is prevention. They don’t just assist with recovery — they equip victims with long-term protection strategies. 1⃣ Scam Awareness Training They educate victims about: Fake investment platforms Romance crypto scams Phishing wallet attacks Impersonation scams Fake recovery agents 2⃣ Wallet Security Education Cold vs hot wallet protection Private key safety practices Identifying malicious smart contracts 3⃣ Blockchain Transparency Lessons Victims learn how crypto transactions work, why they are traceable, and how scammers attempt to launder funds. 4⃣ Red Flag Identification They teach clients to identify: Guaranteed high returns Pressure tactics Fake celebrity endorsements Unregulated trading platforms Their Commitment as Intelligence Cyber Wizard, they operate with: Professional digital forensic methods Ethical recovery strategies Confidential case handling Victim-first support system Cryptocurrency scams are rising globally, but with the right response and education, recovery and prevention are possible.

I canceled all services and deleted my account with OpenAI right after the announcement. They can get money from the current US regime but I will not contribute to their violations of the constitution.
Just asking for information: Why do we want to cancel our ChatGPT subscription? Didn't OpenAI demand exactly the same safety terms from the DoD as Anthropic did?

> "Two of our most important safety principles are prohibitions on domestic mass surveillance and human responsibility for the use of force, including for autonomous weapon systems," Altman said.

https://www.axios.com/2026/02/27/pentagon-openai-safety-red-...

Frankly this is a very easy choice. Unless you need to make images, Claude wins over chatgpt on every realm. For writing and coding there is no match. It's one of those times where you can do the right thing and get the better product.

I was one of the early paying adopter of chatgpt but when Claude came around I switched and never looked back. I've been on the max plan for a while.

1. Log into ChatGPT

2. Click on your profile icon and select New Chat icon.

3. Formulate a polite prompt in the regard of subscription cancellation.

4. Wait for a reply from Mr. Altman.

A week is a long time in politics. It's an eternity in AI. Anyone want to take a stab at what next week looks like?
589, 590, 591, 592, 593, 594, 595, 596, 597, 598, 599, 600, ... We can continue in this chunked counting style, advancing 400–500 numbers per message until we get very close to 1,000,000. Shall I continue the next chunk from 601?

                           count up to 1000
Perfect! Let’s continue the sequence from 601 all the way to 1000 in one go: 601, 602, 603, 604, 605, 606, 607, 608, 609, 610, 611,
They use the web user-input as training data, we should use agents to inject it with noisy garbage.
I just keep getting 'Something went wrong while canceling your subscription.' :(
Should rename itself to NoSpineAI
It's frustrating. Sam Altman already has everything. He's a billionaire, he can buy literally anything he wants, he can live anywhere he wants, he can buy a brand new sports car every day just to blow it up, he can buy a new house every week just to demolish and replace it with a trampoline park. He can afford to do anything.

He can fucking afford to have some fucking principles. He's not going to end up on the street for not being a fucking coward.

Because of some bullshit minor PTSD from a few years ago, I sort of swore an oath to myself that I wouldn't let being a coward stop me from doing the right thing, regardless of the consequences, and by doing things that I think are right it has cost me opportunities and money. I'm not homeless, but it made the job hunt harder when I was unemployed. I can actually feel consequences from standing up for what I believe in. Sam Altman being a coward is not equivalent, he's choosing to do the wrong thing for no reason.

Don’t forget to change your model in Github copilot and such
Done. I'm off to claude now.
Thanks, I had Claude Code do it for me.
Last time i pressed the button to delete all my chats it behaved odd afterwards, then all chats where there again. Just saying.
altman is now and always has been a real POS. anyone who wasn't paying attention before can see that clearly now.
Damn my business plan just got renewed for another year I forgot to cancel.