MacBook Pro
Laptop24 GB
Ollamallama3.2
Set up local models on the Macs you already own. Then reach them from your iPhone, a laptop, or any device with a browser, all from one private workspace.
macOS firstoMLX and OllamaNo paid inference API required
Laptop24 GB
Ollamallama3.2
Mac mini16 GB
oMLXQwen3.5-9B-MLX-4bit
Mac mini32 GB
Ollamagemma3 27B
Illustrative. Model names are examples. Pairing several Macs and using them from any browser works today. The macOS Connector, model discovery and choosing between Macs are in development.
Why it matters
Most households have Macs that spend much of the day idle. Stratumus puts them behind a single private workspace, so any device you own can reach the Mac that has the model you want.
Pair the Macs you own, then reach them from an iPhone, a laptop, or anything with a browser, including as an installed web app.
oMLX and Ollama today, behind one chat. Switch runtimes without changing how you work.
Models run on hardware you already own, so you are not paying an inference provider per request.
Choose which Mac answers first and which steps in when one is offline. Rename or unpair any Mac.
In developmentSetup
This is the installation experience currently being productized for early access.
Open one macOS app. No Node.js, ports, or configuration files.
Approve the connection from stratumus.com.
Stratumus discovers oMLX, Ollama, and compatible installed models automatically.
Hardware
Apple silicon shares one pool of memory between macOS and the model, so RAM decides how large a model a Mac can run comfortably. Here is a sample household.
Mac mini
Comfortable up toabout 9B models
Laptop
Comfortable up toabout 14B models
Mac mini
Comfortable up toabout 30B models
Sample hardware. Sizes are approximate for 4-bit models, and real memory use depends on the model, quantization and context length. As a rule of thumb, keep a model under about two-thirds of a Mac's memory. Model names are examples, not a supported-model list.
How it works
Your selected model runs on hardware you own. Stratumus securely relays requests between the web app and your Mac, so you can use it from anywhere. Each chat is answered by one Mac.
Where you type, on any device.
Your deviceEncrypted connectionSigns you in, keeps your conversations, and relays requests to your Mac.
Hosted by StratumusAuthenticated outbound relayRuns on your Mac and connects out to the control plane.
Your MacLoopback onlyRuns the model you picked, on your hardware.
Your MacWhat Stratumus handles
Prompts and responses pass through the Stratumus control plane, currently hosted on Railway, on their way between your browser and your Mac. Direct or end-to-end encrypted transport is planned.
Conversations are managed by the Stratumus application, which is how your history persists between sessions.
Web search is optional. When you turn it on, your search query goes to an external search service to fetch results.
Status
FAQ
Not yet. Today, each chat is answered by one Mac running one model. Turning the extra hardware you already own into a shared server for bigger models is where we're heading. It's a concept for now, not part of the current build. We're passionate about putting the compute power we already have to work as our own AI infrastructure. We like the way you think!
Inference stays on your Mac. Prompts and responses pass through the Stratumus relay so you can reach your Mac from anywhere. See exactly what Stratumus handles.
A Mac running oMLX or Ollama, and a browser on the device you are using. Your Mac connects out to Stratumus, so there are no ports to open, and it works away from home. Stratumus is macOS first.
Stratumus is starting on macOS with oMLX and Ollama. Leave your email and we'll let you know when early access opens.