Join the early access program
Browse documentation
ConceptBeginneradmin, member, viewer

Token Usage & Shared Pool

Understand how LLM tokens are calculated, how the company-wide shared token pool operates, and how to optimize consumption.

4 min read11,200 viewsUpdated 2026-03-01

What are tokens and how are they consumed?

Large language models process text not as full words, but as mathematical subunits called "tokens" (in English, 1,000 words typically equates to ~1,300 tokens).

Every task executed by an agent comprises two consumption stages:

  • Input Tokens: The prompt you send, conversation context history, uploaded documents (PDFs, spreadsheets), active agent directives, and system rules.
  • Output Tokens: The agent's generated response, structured data tables, synthesized code, and tool execution parameters.
In Botonom, you never need to juggle separate quotas per agent. All hired employees draw from your organization's unified, company-wide shared token pool.

Company-wide shared token pool

Your subscription tier allocates a collective monthly token allowance shared across all agents in your workspace:

PlanMonthly Shared PoolActive Digital EmployeesAdditional Token Top-ups
Personal6,000,000 tokens1 digital employeeYes (1M, 5M, 10M packs)
Business15,000,000 tokensUp to 10 digital employeesYes (1M, 5M, 10M packs)
EnterpriseCustom high-volumeUnlimited employeesCustom contracted packs
  • A high-volume customer support agent answering hundreds of daily inquiries shares the same pool as a finance agent performing deep weekly reconciliations.
  • When pool consumption reaches 80% and 95%, administrators receive automated in-app and email threshold alerts.
  • If tokens reach 100%, non-critical scheduled routines pause until an instant top-up pack is purchased or the pool resets on your monthly billing date.
  • Tokens do not roll over between billing cycles; your pool completely refreshes at each cycle start.

Model tiers and smart routing

Not every business task demands maximum reasoning horsepower:

  • Standard Models (1x Weight): Customer inquiries, appointment scheduling, meeting summarization, and rapid document queries.
  • Genius / Advanced Reasoning Models (3x–5x Weight): Deep multi-table database queries, legal document contract analysis, multi-agent synthesis, and autonomous coding.

Botonom's Smart Router automatically selects the most efficient model tier based on task complexity, ensuring your token pool lasts longer.

Best practices for token optimization

  1. Precise Directives: Write concise, clear directives rather than repetitive paragraphs.
  2. Selective Knowledge Ingestion: Upload relevant chapters or documents rather than raw, uncompressed 500-page archives.
  3. Task Scoping: Break massive multi-department workflows into modular steps with human verification gates.
tokensbillingoptimizationshared-pool

Start hiring in three steps

No interviews, no onboarding calls. Just browse, hire, and let them work.

01

Browse Employees

Explore our AI talent pool. Filter by role, department, or industry to find the perfect candidate for your team.

02

Hire & Onboard

Select your AI employees and equip them with skills. Pick the plan that fits your headcount, no config needed.

03

Watch Them Work

Your agents handle tasks 24/7. Monitor performance, adjust skills, and scale your workforce anytime.

No credit card requiredSet up in 5 minutesCancel anytime