Skip to content
SpecifyDocs

Models

Usage and AI budget

How AI usage is calculated, budget warnings and top-ups, and ways to use less.

AI usage in Specify is managed as a monthly AI budget per workspace. Every task that uses AI, including chat, document generation, automations, and knowledge search, draws from the same budget.

Check your usage

See this month's usage period and AI budget usage under Settings → Plans & Billing → Usage. Usage also appears in the sidebar account menu.

The Usage page showing this month's usage period, unlimited documents, and an AI Budget usage bar
  • Usage is scoped to the current workspace. Personal and team workspace budgets are separate.
  • Usage resets at the end of each monthly usage period.

How usage is calculated

Each request's actual cost is calculated from the token prices of the model used and deducted from the budget. The basis is actual token usage, not credit multipliers or request counts.

  • Usage grows with more expensive models, higher reasoning effort, and reading many long documents. For per-model prices, see Models overview.
  • When the same content is read repeatedly, caching applies and input costs drop significantly.
  • For GPT models, when a single request's input exceeds 272,000 tokens, that request's price goes up (2× for input, 1.5× for output).
  • Knowledge search also uses a small amount of usage.

Budget by plan

Higher plans come with a larger monthly AI budget.

PlanAI usage
FreeBasic AI usage
ProMore AI usage than the Free plan
Max5× the AI usage of Pro
TeamA shared pool where the team uses the per-seat budget together
EnterpriseCustomized for your organization

For exact limits and prices, see the pricing page.

Warnings and top-ups

  • A warning appears when you've used 70% and 90% of the budget.
  • When the budget is used up, an AI budget exhausted banner appears, and you can add usage with Top up $5. Top-up credits are deducted only after the base budget is used up.
  • Even with a remaining balance, each run has round and time limits, so a task may finish only partially.

Ways to use less

  1. Default to GPT-6 Luna. The least expensive model is enough for search, summaries, and short questions.
  2. Raise reasoning effort only when needed. Low or Medium is enough for most document work.
  3. Narrow your questions. Specifying targets with document names or @ mentions reduces unnecessary searching and reading.
  4. Start a new conversation instead of continuing a long one. The longer a conversation gets, the more each request has to read.
  5. Review automation frequency. Switch automations that don't need to run daily to a weekly schedule.