Chat Direct to your local model — each send is a curl on your Mac
Server checking…
These act on the currently active Mac (named above). To switch which Mac serves, or turn an individual Mac on/off, use the Server Fleet tab.
Pull a model
Downloads from the Ollama registry onto your Mac. Big models take a while.
Installed models —
MODEL SETTINGS
Defaults are Ollama's. Recommended is the model vendor's guidance for vera:1.0 (temperature 0.7, top_p 0.8, top_k 20) — what this panel uses. Leave a custom field blank to fall back to the Ollama default.
| SETTING | DEFAULT | RECOMMENDED | YOUR VALUE |
|---|
Save settings keeps these as your per-request options in this browser. Apply to server bakes temperature/top_p/top_k/context into the selected model on its Mac for everyone (persistent). The system prompt also ships as the built-in default, so every device starts with it.
AVAILABILITY SCHEDULE
Keeps this model loaded during the window via a cron job on your Mac. Runs even when this app is closed. Times use your Mac's local clock.
No schedule set for this model.
Connect from anywhere
All traffic flows: your browser → CloudFront (WAF + TLS) → Caddy on the AWS gateway → reverse SSH tunnel → Ollama on the Mac. Requests are rate-limited at the edge; backend origins never reach the browser bundle.
none — public endpointList models Chat (streaming)
Keys are stored on the gateway and listed here for the authenticated admin. Copy a key when you create it. (The public chat path is open; these keys are for the developer API surface.)
Server Fleet Macs serving models, reached via reverse tunnels
Add a server
Sessions Every user session — logged to S3 (location fills in once CloudFront fronts it)
Model Console Fixed commands only — list · ps · load · unload · pull · run
Type "help" and press Enter.