Skip to content
Models

Models

The model sheet lists what can run on this device, with speed, download size and context on every row. Two kinds of model appear: OIOXO models, which we build and which count against the free window, and open models, which are stock releases from Hugging Face and never count against anything.

OIOXO models

ModelDownloadNeeds free memoryNotes
OIOXO Light540 MB750 MBFastest; the default on most laptops.
OIOXO Prime1.16 GB1.5 GBThe stronger OIOXO model.
OIOXO Ultra2.5 GB2.8 GBDesktops and strong GPUs.
OIOXO Light (Mobile)545 MB750 MBThe phone build of Light.

On Free, OIOXO models are available for a 5-hour window with weekly and monthly caps; on Pro they are unlimited. The meter is always visible before a limit is hit.

Open models

ModelDownloadContextNotes
Qwen2.5 Coder 1.5B986 MB (Q4_K_M)32kAgent-capable; what Free falls back to after the window.
Qwen2.5 Coder 3B1.93 GB (Q4_K_M)32kStronger; wants 8 GB RAM.
Llama 3.2 3B Instruct2.02 GB (Q4_K_M)128kGood at chat, weaker at tool calls.

Open models are unlimited on every plan. On the desktop, larger open GGUFs appear when the machine can hold them; the web sheet hides anything over 2.2 GB or 9B parameters because a browser tab cannot run it.

Helpers that download only when used

OIOXO Vision (260 MB) and Vision Prime (511 MB) read an image you attach to the composer. They are fetched the first time a feature needs them, never on boot.

Consent and fit

  • Nothing downloads without a click. A model that does not fit is shown as such and not started.
  • A downloaded model stays in browser storage and loads from disk on the next visit. Settings → Models & keys lists and deletes them.
  • A first run can take minutes on CPU while the model reads the prompt; the strip shows the estimate and, past it, the elapsed time.
Which one?

Start with what the sheet recommends. If a run writes a file that does not work, OIOXO offers a bigger model you already have or one of your keys — see Modes.