Skip to main content
Everything the agent does with AI is paid for with AI credits: prepaid funds that belong to your organization. The principle is simple, and it is stated in the product:
Credits pay for the agent’s work: reading documents, drafting, and browsing the web. You pay what the AI providers charge us, with EU hosting adding 10% on Claude models.
Model usage is billed at the provider’s list price, with no margin added by Casa Conect. The 10% on Claude models is what Amazon Bedrock charges for keeping inference inside EU regions. It is the price of EU processing, passed through unchanged.

Buy credits

In Settings → Access & billing, choose a pack or enter a custom amount. Credits are charged to the organization’s saved card and are available immediately. The page shows your balance and an estimate: “About N days left at the current rate.”
When credits run out, AI features pause until you top up. Conversations, workflows and indexing stop, including work the agent was doing unattended. Your data is unaffected. Use auto-refill to avoid this.

Auto-refill

Charge the saved card for a credit pack whenever the balance drops below the threshold. Choose the threshold and the pack. Two safeguards protect you from a runaway charge:
  • At most five automatic charges per day. If the limit is reached, auto-refill pauses until an admin switches it on again.
  • If the card is declined, auto-refill pauses and the page tells you so.

Control spending

Monthly budget

An optional cap on the organization’s AI spend per calendar month, in USD. When it is reached, AI work stops until the next month, or until an admin raises the cap.

Per-member limits

See what each member has spent this month, and set a monthly limit for any of them. A member who reaches their limit is paused. The rest of the team is not.

Model policies

Block the most expensive models for everyone.
Before any model is called, Casa Conect checks, in this order: an active subscription, the model policy, the credit balance, then the organization budget and the member’s limit. A refusal at any step costs nothing, and the member is told which one applied.

What consumes credits

Usage is broken down by activity under Settings → Access & billing → Usage: The same page breaks usage down By model, showing calls, tokens and cost, and by member.

What each model costs

Model usage is measured in tokens, which are pieces of words. As a rule of thumb, 1,000 tokens is about 750 English words, and somewhat fewer in Spanish. Input is everything the model reads: your message, the conversation so far, and the documents and search results it looked at. Output is what it writes. Output costs more than input. Provider list prices in US dollars per million tokens, as of September 2026: EU Includes Amazon Bedrock’s 10% premium for EU regional inference.
Providers change their prices, and Casa Conect follows them. Your Usage page always shows what you were actually charged. This table is a guide.

A sense of scale

Cached input is why long conversations cost less than you might expect. When the model re-reads material it saw moments ago, the provider charges about a tenth of the normal price.

Long conversations

On the ChatGPT frontier models, a single request that contains more than 272,000 input tokens, which is several hundred pages, is priced by the provider at twice the input rate and one and a half times the output rate. In practice this affects only very long conversations over very large documents. Starting a new conversation for a new subject avoids it.

Other usage

Keeping costs down

  • Leave Auto on for everyday work, and pick a stronger model only when the task needs it. See Choosing a model.
  • One subject per conversation. The whole conversation is re-read on every turn.
  • Reference documents instead of pasting long text. The agent then reads only the sections it needs.
  • Use workflows for template filling. They use an economical model wherever that is sufficient.
  • Close browser sessions when you have finished with them.
  • Upload clean scans. Poor scans cost more to read and give worse results.