How do limits work?
Each plan includes a monthly agent usage budget denominated in dollars. Each Notis message deducts from that budget using Notis model rates, which include a 20% service markup over the provider’s published API price for OpenAI models. You can view your billing cycle, included usage, On-Demand usage, and request breakdowns on the Usage page of your portal. In a Manager thread, the usage indicator also shows the thread’s completed cost and how much context the latest completed run used. Each paid plan has its own cap: $20/month on Pro, $59 on Pro+, and $149 on Ultra. If you use Notis with a team, everyone shares one monthly usage pool. For example, a 10-person Ultra team, at $149 of included monthly usage per seat, gets a shared $1,490 pool for the month. Any teammate can use it, and all team activity counts toward the same total. The Usage page shows the team’s shared usage for the current billing cycle, and billing settings are managed by the team owner. Notis manages model routing automatically based on the task. Prices below are shown per 1M tokens and reflect the current Notis rates after markup.Auto mode
By default, Notis runs on Auto (recommended). Instead of a fixed model, Notis’s smart router reads each turn and picks the intelligence level — model and reasoning effort together — that fits the task: light, inexpensive work for a quick reply, and a stronger model for a hard, multi-step task. Because the choice is made per turn, a single conversation can move up and down levels as the work changes, so you only pay for the capability each step actually needs.Intelligence levels
You can pin an intelligence level — per conversation in the Desktop and Web app, as a per-channel default in Channels, or per automation in the automation editor — instead of letting Auto choose. Each level sets the model and its reasoning effort together:
The multiplier compares each model’s regular per-token rates with Medium. Reasoning can change how many tokens a message uses, so actual billing always uses the real token counts and rates below.
Processing tiers (priority)
Notis chooses the processing tier automatically, but on a plan with priority access you can pin one per thread or per channel alongside the intelligence level (it shows as Standard / Fast / Economy in the picker):
If a model does not support Economy or Fast, Notis uses its Standard rate instead. The Economy and Fast token tables below show the resulting per-token rates.
Automations always run on Economy (Flex) — about half the base price, at a lower (slower) priority. That’s fixed and separate from the per-automation Intelligence setting: pinning an automation to High or Low changes only its model and effort, never its priority, so it keeps the Flex discount either way.
- Normal
- Flex
- Fast
Reading text aloud and reading a web page are counted in your usage like everything else. Both used to be free, so if you use them a lot you will now see them in your usage breakdown.
Pay-per-use tools are counted the same way. Each one has its own price per use, shown on its page, and it comes out of the same monthly usage.

