← Darkbloom help

Can you run several Darkbloom models at once?

Updated

Yes: by default the Darkbloom provider can keep up to three models loaded at once. Whether it should is another question. For most Macs one well-chosen model earns more than several. Facts below come from Darkbloom’s open-source docs and issues (links at the end); memory figures are worked out from Darkbloom’s own load formula.

How it works

What a second model costs in memory

Each loaded model adds its full weights (with a 20% loading allowance). A working reserve for activations is charged once, not per model, plus at least 1 GiB for concurrent requests. By default the provider also holds back 4 GB and uses at most 90% of memory, and macOS and your apps need their share. Free memory needed to load each set, from Darkbloom’s formula:

Models loaded togetherFree memory neededSmallest Mac to try
gpt-oss-20b alone18.0 GiB24 GB (tight)
Gemma 4 26B QAT alone23.9 GiB32 GB
Qwen3.6 35B A3B alone30.3 GiB36 GB
gpt-oss-20b + Gemma 4 26B QATabout 37.5 GiB64 GB (48 GB too tight)
gpt-oss-20b + Qwen3.6 35Babout 43.8 GiB64 GB
Gemma 4 26B QAT + Qwen3.6 35Babout 47.7 GiB64 GB (tight)
Qwen3.5 35B + Qwen3.6 35Babout 53.7 GiB96 GB
gpt-oss-20b + Gemma 4 26B QAT + Qwen3.6 35Babout 61.3 GiB96 GB

Why one model usually earns more

Practical combinations

Sources

How BloomGauge helps

BloomGauge’s Manager runs one model at a time. With Manager on, it holds the model that has paid best on your Mac over the last 30 days (or, on a new Mac, what pays best on Macs with the same chip and memory), and moves only when at least five Macs like yours have clearly earned more on another model for two hours. Switches go through Darkbloom’s own CLI and must pass memory, idle and temperature checks. It starts observe-only.

Download BloomGauge, free for Mac

Questions

Can Darkbloom run more than one model at a time?

Yes. The provider keeps up to max_model_slots models loaded (default 3), memory permitting. Each extra model adds its full weights to memory, and Gemma 4 only gets public work when it is the only model the Mac advertises, so one model is usually the better choice.

How much memory do two Darkbloom models need?

Worked out from Darkbloom’s load formula: about 37.5 GiB free for gpt-oss-20b plus Gemma 4 26B QAT, about 44 GiB for gpt-oss-20b plus a Qwen 35B, and about 48 GiB for Gemma plus a Qwen 35B. In practice that means a 64 GB Mac or larger.

Should I advertise several models on Darkbloom?

Usually not. Advertising Gemma with anything else loses public Gemma work, a second model takes memory from concurrent requests, and advertising more models than slots can cause constant loading and unloading. Pick one model, keep others downloaded and switch when demand moves.

Related

Updated 2026-09-29. Still stuck? Ask in #bloomgauge on the Darkbloom Slack or contact us. BloomGauge is independent and not affiliated with Darkbloom.