Sharing a model
If Ollama, LM Studio, or llama.cpp runs on your computer, the node can find its models and share one in a single step:
bnw share models # models served here, and whether they're sharedbnw share model gemma3:4b --access contacts # publish, connect, answer automatically, offerbnw share model gemma3:4b --access anyone --max-per-hour 10 # sharing again updates it- What sharing does. It publishes the model as an
inference.modelcapability and connects it to the runtime’s OpenAI-compatible chat endpoint (egresslocal). It sets the autopilot to auto (or--mode ask) with 30 calls per person per hour, sets who may use it, and signs a week-long offer that renews itself. It is all the steps below, with defaults. - Detection. The node probes only the runtimes’ fixed loopback addresses (
127.0.0.1:11434,:1234,:8080), with short timeouts and bounded replies. Ollama models that list capabilities withoutcompletion, such as embedding models, are skipped. - Only local endpoints. One-step sharing accepts only an
httpchat endpoint on this computer’s loopback interface; pass--urlfor a runtime on another port. Remote or paid endpoints go through the full setup below, which asks about egress and secrets. - In the web client. My capabilities → Share a model on this computer does the same, with a choice of who can use it.
