Skip to main content

Q-Hub AI — how it works and how AI usage will be charged

Updated 10 September 2026 · 3 min read

Summary

Q-Hub AI is available to customers now and we are not charging for it. Every account gets an allowance of AI tokens, and we are topping accounts up for free on request while we tune efficiency and understand our own cost base.

This page sets out what the AI system is, how usage is measured, what the current position is, and what the projected charging model looks like when we do move to a paid model.


What the Q-Hub AI system is

Q-Hub AI sits across the platform rather than being a single feature. In broad terms it does three things:

  1. Ingests and indexes customer content — documents, forms, processes, registers and other records held in the customer's hubs are processed and turned into a searchable knowledge index for that company.

  2. Retrieves the right context for a question — when a user asks something, the system finds the relevant records for that user, within their access permissions, rather than sending everything to a model.

  3. Reasons over that context — a large language model produces the answer, summary, draft or suggestion, grounded in the retrieved records.

This is a retrieval-augmented generation (RAG) design. It matters commercially because the cost is driven by how much content is ingested and how much reasoning is done, not by seat count.

Where users meet it

Surface

What it does

AI search / deep search

Natural-language search across the customer's Q-Hub data

Q-Hub Assistant

General chat assistant available across the app

Chat with…

Contextual chat against a specific document, form, process or record

AI builders

Assisted creation of forms, processes and requirements

AI actions

AI-assisted steps inside workflows

Compliance gap analysis

In development — analysis of a customer's system against a standard

Platform and data handling

  • Hosted on AWS (eu-west-2) alongside the rest of Q-Hub, with MongoDB Atlas as the vector store.

  • Reasoning runs on Google Vertex AI in the London region.

  • Retrieval respects existing Q-Hub access controls — a user cannot retrieve content they could not otherwise open.

  • Company admins control whether AI is on, which users have it, and which content is excluded.


How usage is measured

Usage is measured in tokens. A token is the unit we deduct for AI work carried out on the customer's behalf. Tokens are consumed by activity, not by the number of licensed users, so a small team running heavy analysis can use more than a large team doing occasional searches.

What tokens are actually used for

Indicative costs for common tasks. Ranges are real — the cost of a task depends on how much content the AI has to read to do it.

Task

Typical cost

What drives the range

Build a simple form from a PDF

7–20 tokens

Length and complexity of the source PDF

Build an audit template

10–20 tokens

Number of questions and depth of the standard

Analyse a process for customer trends — initial request

20–80 tokens

Volume of data being analysed

Analyse a process for customer trends — follow-up questions

5–10 tokens

Context is already loaded, so follow-ups are cheap

Assess audit requirement compliance

10–50 tokens per requirement

Volume of connected evidence and content

Two things worth understanding from these figures:

  • Follow-up questions are cheap. Once the AI has loaded the context for an analysis, further questions on the same topic cost a fraction of the initial request. Users get better value by working through a topic in one session than by starting fresh each time.

  • Per-requirement work scales. Compliance assessment is charged per requirement, so a full standard is a large job.


Current position — no charge

Item

Position today

Charge to customer

None. AI usage is not billed.

Standard allowance

300 tokens per account

Additional tokens

Granted free on request, subject to review

Review basis

Usage pattern, account size, and our own cost exposure

Why

We are still improving the efficiency of the AI and balancing our underlying cost. We would rather learn from real usage than cap it early.


Projected charging model

The following is the model we expect to move to but is not yet confirmed.

Monthly refreshing allowance by plan

Each customer receives a token allowance that refreshes monthly.

Plan

Refreshing tokens per month

Equivalent list value (excl VAT)

Pro

1,000

£50

Essentials

500

£25

Basic

300

£15

Additional token purchase

Item

Value

Price per token

£0.05 (5p)

Minimum order

1,000 tokens

Cost of minimum order

£50

Additional tokens are bought in blocks of 1,000. Smaller top-ups are not offered under this model.

Explore the Q-Hub platform

Was this article helpful?

Ready to try it? Get started

Related articles