The brain runs in your environment, under your access rules. Never our cloud.
Your environment. Your tenancy. Your access rules. Never our cloud. Standard runs where you already keep the company — on-prem, hybrid, or your own cloud account. We recommend. Your CTO decides. The graph moves with you if you change models. Standing up the new serving environment still takes real work.
Three models. You pick; we recommend. What is identical across all three: you own the brain, the build is a burst, models are swappable, the stewardship fee is the same, and moving between models is a supported path: the graph moves; bringing up and cutting over the new serving environment is still real work.
Model 1 is full on-prem, air-gap capable. Model 2 is hybrid — the default — premises plus a receipted hop into your own VPC. Model 3 is self-hosted models in your cloud tenancy, not a third-party AI API.
Performance commitments are capability-level: sessions, working context, recovery. We do not publish raw throughput.
We install Mnemma on ourselves first. The same gates we ask you to walk, we walk.
The brain runs in your environment, under your access rules. Never our cloud.
Three models: full on-prem, hybrid (default), or your own cloud tenancy.
You pick the model. We recommend.
Performance commitments are capability-level, never raw throughput.
We install Mnemma on ourselves first. The same gates we ask you to walk, we walk.
Machine-readable → index.json · /ai-info · llms.txt