Plans
- Free: $10 one-time usage allowance, up to 3 users, sandbox size up to
micro - Pro ($60/month): $60 monthly usage allowance, up to 5 users, sandbox size up to
medium, with overage billing - Max ($200/month): $200 monthly usage allowance, up to 10 users, sandbox size up to
xl, with overage billing
How the usage allowance works
Two usage categories draw from the same balance:- Managed inference: model input, output, and cache usage is charged at the rates in Usage Pricing.
- Cloud VM compute: metered from the CPU, memory, and runtime used by each session.
Cloud VM compute rates
VM compute is a separate usage category from managed inference. It draws from the same dollar-denominated allowance at $0.0403/vCPU-hour + $0.0130/GiB-RAM-hour.
The per-size hourly prices below apply those resource rates to each VM configuration.
Don’t see the VM size you want? Larger VM sizes are available upon request. Contact support@tembo.io to discuss your requirements.
What consumes the allowance
Managed inference and VM compute are metered while Tembo runs work such as:- Creating pull requests from assigned issues (Linear, Jira, Slack)
- Fixing production errors from Sentry
- Optimizing database queries and indexes
- Processing pull request feedback through the Feedback Loop
When the allowance runs out
What happens when your balance reaches zero depends on your plan and overage setting:- Free: new sessions are blocked until you upgrade or add prepaid balance.
- Pro or Max with overages enabled: sessions continue and are billed as overage after the included allowance and prepaid balance are used, up to the maximum overage limit you set. Once that limit is reached, new sessions are blocked.
- Pro or Max with overages disabled: new sessions are blocked as soon as the included allowance and prepaid balance are exhausted.
Usage bursts
Some integrations can create a large batch of work at once — for example, a repository or integration sync, or a token refresh that triggers a backlog of issues. Tembo does not cap or rate-limit how much work an integration enqueues; the batch is processed as worker capacity becomes available, and each resulting session records dollar usage when it runs. The same allowance and overage limits above apply, so a burst cannot spend beyond your available balance plus your overage limit. To absorb large bursts without interruption, enable overages with an appropriate limit.Usage refunds
Managed inference usage may be refunded automatically in some cases:- Failed sessions: if Tembo cannot complete a session because of an internal error
- Duplicate issues: if the same issue is queued more than once
- Invalid inputs: if a session fails due to missing access or configuration problems
Usage and monitoring
You can track managed inference, compute, total usage, VM resource-hours, and daily spend in the Billing dashboard. Overage invoices list managed inference and compute separately.Prepaid balance and on-demand usage
- Free plan: upgrade to Pro or Max for overages, or add prepaid balance
- Pro and Max: enable overage billing and set a monthly spend limit
- Prepaid balance: does not expire
Troubleshooting
Payment failed: confirm the card is valid and has available funds. Prepaid balance did not appear immediately: billing updates can take a few minutes. Need help? Contact hi@tembo.io.Enterprise pricing
If you need custom pricing, reach out to support@tembo.io for:- Custom usage allowances
- Custom or unlimited sandbox sizing
- Volume pricing
- Custom inference configuration
- Dedicated support
- Flexible billing options