Skip to main content

How do limits work?

Each plan includes a monthly agent usage budget denominated in dollars. Each Notis message deducts from that budget using Notis model rates, which include a 20% service markup over the provider’s published API price for OpenAI models. You can view your billing cycle, included usage, On-Demand usage, and request breakdowns on the Usage page of your portal. In a Manager thread, the usage indicator also shows the thread’s completed cost and how much context the latest completed run used. Each paid plan has its own cap: $20/month on Pro, $59 on Pro+, and $149 on Ultra. If you use Notis with a team, everyone shares one monthly usage pool. For example, a 10-person Ultra team, at $149 of included monthly usage per seat, gets a shared $1,490 pool for the month. Any teammate can use it, and all team activity counts toward the same total. The Usage page shows the team’s shared usage for the current billing cycle, and billing settings are managed by the team owner. Notis manages model routing automatically based on the task. Prices below are shown per 1M tokens and reflect the current Notis rates after markup.

Auto mode

By default, Notis runs on Auto (recommended). Instead of a fixed model, Notis’s smart router reads each turn and picks the intelligence level — model and reasoning effort together — that fits the task: light, inexpensive work for a quick reply, and a stronger model for a hard, multi-step task. Because the choice is made per turn, a single conversation can move up and down levels as the work changes, so you only pay for the capability each step actually needs.

Intelligence levels

You can pin an intelligence level — per conversation in the Desktop and Web app, as a per-channel default in Channels, or per automation in the automation editor — instead of letting Auto choose. Each level sets the model and its reasoning effort together: The multiplier compares each model’s regular per-token rates with Medium. Reasoning can change how many tokens a message uses, so actual billing always uses the real token counts and rates below.

Processing tiers (priority)

Notis chooses the processing tier automatically, but on a plan with priority access you can pin one per thread or per channel alongside the intelligence level (it shows as Standard / Fast / Economy in the picker): If a model does not support Economy or Fast, Notis uses its Standard rate instead. The Economy and Fast token tables below show the resulting per-token rates.
Automations always run on Economy (Flex) — about half the base price, at a lower (slower) priority. That’s fixed and separate from the per-automation Intelligence setting: pinning an automation to High or Low changes only its model and effort, never its priority, so it keeps the Flex discount either way.
For unusually large GPT-5.6 requests with more than 272,000 input tokens in total (including cached input), the whole request uses long-context rates: input and cached input are charged at 2× the table rate, and output at 1.5×. A request with exactly 272,000 input tokens still uses the regular rates above. Specialized tools use their own model rates: Deep Research uses gpt-5.6-terra with public-web search. Its cost is based on the input and output tokens reported for the finished research, using the Terra rates in the tables above; there is no separate per-report list price. Reading text aloud and reading a web page are counted in your usage like everything else. Both used to be free, so if you use them a lot you will now see them in your usage breakdown. Pay-per-use tools are counted the same way. Each one has its own price per use, shown on its page, and it comes out of the same monthly usage.

What happens when I reach my limit?

Once your monthly usage is exhausted, Notis notifies you and prompts you to either upgrade to a higher tier, activate On-Demand Usage, or wait for your billing cycle to reset. On teams, that means the shared team pool is exhausted. On-Demand is optional. If you leave it disabled, Notis pauses paid work once your plan cap is exhausted. If you enable it, Notis can keep working past your included credits and bill the extra usage separately. You can choose a fixed monthly On-Demand limit or use Unlimited mode. Read the full guide here: On-Demand Usage.