The model Puckie uses to answer questions and chat. It runs on your Mac by default. A bigger one gives better answers.
Qwen3-4B, running privately on this Mac. Puck starts it for you. Nothing to set up.
Have a second Mac with more memory? Run a bigger model there and Puckie will use it over your home network.
ollama pull qwen3:8b (or any model you like).studio.local.YOUR-MAC with that Mac's name.In LM Studio, load a model, open the Developer tab, start the server and turn on Serve on local network. In Puck, use LM Studio (on this Mac) or Another Mac and change the port to 1234.
Any service that works with OpenAI's API can answer for Puckie: enter its address, model and key. Puck warns you when what you say would leave your network.
| Puck says | Try |
|---|---|
| No Mac called … on this network | Check the name in Sharing on the other Mac, or use its IP address. |
| No answer from … | Is the other Mac awake? Is Expose Ollama to the network on? |
| Found …, but no model is being served | Open Ollama (or LM Studio's server) on that Mac. |
| macOS is keeping Puck off the local network | System Settings → Privacy & Security → Local Network → turn on Puck. |