The promiseLocal when you can. Cloud when you want.
Caelum Stack will not pretend that every machine can run a good model — that depends on your hardware, and it will tell you the truth about yours rather than sell you a fantasy. What it does promise is the choice. If your machine can run a model, you never have to hand your work to anyone. If it cannot, nothing about the product is withheld from you.
↺
If your machine can run one
Serve it from the Models page and the whole loop — voice in, reasoning, answer out — stays on your hardware. No key, no provider account, no request over the wire.
If it cannot, or you would rather not
Connect Anthropic, OpenAI, Google or anyone else and you get the same voice, the same tools, the same interface. Only the conversation leaves — your audio was already transcribed on your machine.
Switching between the two is one click, whenever you like.