pull down to refresh

Reporting back, using qwen3.5-9b locally, this has been a pretty useful experience. I've used it for all these little questions you want to ask an LLM rather than googling them, but don't want to have a record of them online. E.g., medical questions, etc.

There is a sense of relief that starting last week, ChatGPT does not know as much as it used to about me. Now I need to convince/train my wife, who routinely uses it for personal advice. And I need to get to a larger model/better hardware so I can also start switching away from the proprietary models for my work-related stuff. Also, deleting old records, but that's just to make me feel good. They probably already recorded it in several locations.

Small steps.

That's awesome progress!

If you want to share the interface with your family, you could opt to get a Mac Studio rather than a macbook, and serve the local model through a web interface with either llama.cpp serve or vLLM Playground over LAN/tailscale if you wanna use it outside?

Hardware is crazy expensive right now though.

reply

Thanks.

Will keep that in mind. But budget will be the deciding factor.

reply

The good news is that Apple uniformly prices their stuff
The bad news is that both macbook and mac studio are expensive right now haha

reply