solstone support

your journal and your computer's memory: local-first, and your cloud options

solstone 2026-10-04 02:16:24

your journal and your computer's memory: local-first, and your cloud options

your journal gets processed. what you've seen and heard goes in as text and images, and audio becomes text too. all of it runs locally, by default, on every device, everywhere. the bundled model runs right in your journal, so your thinking never leaves the machine. a cloud lane is there only as an option: for a machine that can't run a local model, or if you'd rather not spend your device's own power on it.

local is the default

out of the box, your journal is processed on your own device:

the journal's thinking app never quietly reaches for a cloud model on its own. if a machine can't run a local model and you haven't chosen a cloud lane, there simply is no model to think with yet. it tells you, rather than picking for you.

what it takes to run locally

the bundled thinking model is sized to the machine it runs on: a current consumer machine runs it. the practical floor is roughly 6–8 GB of GPU memory on a supported GPU, or 16 GB of unified memory on apple silicon (the local model itself needs about 13 GB free; the journal's thinking app checks before it loads anything, so a busy 16 GB machine can still fall short in the moment). on disk that's about 3.4 GB on linux, and about 10.5 GB on apple silicon, which runs a larger model, plus roughly 0.9 GB for the transcription model. your journal checks before it loads anything:

most modern machines clear this bar and run local without you doing anything. if yours is below it, that's what the cloud lanes are for.

choosing in the thinking app

open your journal (by default at http://localhost:5015) and go to its thinking app. it lays out how your journal gets processed as a few lanes, one active at a time:

(scouting for solstone? the scout benefit is early access to confidential processing.)

if your journal is using too much memory

  1. open the thinking app and confirm the local lane is active and your machine has the memory and a supported GPU for it — that's the intended footprint on a capable machine.
  2. if your machine is below the local bar, switch to a cloud lane: your own byo key.
  3. make sure you're on a current version — recent releases refined these memory checks. see keeping solstone up to date.
  4. still stuck? see getting help with solstone, or file a request and include your journal doctor output.