Avalos AI
← Back to chat
Effort tiers

Five levels of thinking.
Each one is a real budget, not a label.

Changing tier changes what actually happens on the server: how many tokens the answer may use, whether the model deliberates before replying, how many times it checks its own work, and which engine answers at all. The icon is a gauge — the higher the tier, the more is moving inside it.

What each tier actually turns on

These are the real settings behind the dial, not marketing tiers. Higher tiers cost more time by design — deliberation is the product, not a side effect. Your tier sets a ceiling, not a fixed spend: the question you actually ask decides where inside that ceiling it lands. Ask something trivial and every tier answers in reflex — you are never made to wait for “what is the capital of Portugal”. The figures below are the CEILING each tier is allowed, not what it spends: each reasoning tier spends inside its band, low on an easy question and high on a hard one. Reasoning starts at Galaxy — Base and Star answer directly.

TierToken bandDeliberatesSelf-check passes Answer samplesLive researchResearch depthToolsEngine
T1 · Base4,000no01no225Free — Avalos CPU
T2 · Star4,000no01no325Paid — GPU
T3 · Galaxy4,000–12,000yes11yes435Paid — GPU
T4 · Cosmos8,000–16,000yes13yes535Paid — GPU, largest first
T5 · Quantum16,000–21,000yes25yes635Paid — GPU, largest first

Which should you pick?

A simple rule: match the tier to the cost of being wrong.

If you are…UseBecause
Trying Avalos AI for the first timeT1Free. Five messages, then a quick sign-in. Runs entirely on Avalos hardware, so nothing leaves the box — and it is the slowest tier, not the fastest. Star answers the same question in seconds.
Chatting day to day, and you want it quick and sharpT2The everyday default on a paid plan. The same direct answer as Base, moved onto GPU — it does not deliberate; that starts at Galaxy. Same 4,000-token ceiling, same 25 tools, answered faster.
Asking something with a right answer you can't easily checkT3Reasoning turns on and the answer is checked once before you see it.
Doing real analysis, or something you'll act onT4Three independent attempts; it only settles when they agree.
Working on something where a mistake is expensiveT5Five attempts, two verification passes, deepest research. Slowest on purpose.

Plans — how much you can use

A plan is your allowance. A tier is how hard Avalos AI thinks about one question. They share names because a plan is named for the top tier it unlocks — but they are two different dials, and you change tiers freely inside whatever plan you are on. Allowances are counted in tokens, not messages, because a Quantum answer costs roughly fifty times a Base one; counting messages would price them the same and that would be dishonest.

Every allowance below is set to be at least 15% more generous than the comparable plan at the big labs, at the same price point. Two windows run at once: a rolling 5-hour session and an ISO week that resets Monday 00:00 UTC.

PlanPricePer 5h sessionPer weekAnswers per session
Guest—15,000120,000~30 at Base-T1
Free—35,000714,000~14 at Star
Base$560,0001,190,000~120 at Base-T1
Star$15130,0001,785,000~52 at Star
Galaxy$50520,0008,925,000~130 at Galaxy
Cosmos$601,400,00015,000,000~107 at Cosmos
Quantum$752,900,00019,000,000~111 at Quantum

The “answers per session” column is an estimate until real traffic calibrates it — a short question costs far less than a hard one, at every tier. Inside the product the meter shows your own measured average instead, and that figure is computed on your device and never leaves it.