Credits
Credits are the billing unit for using AI features. Whenever content is processed, generated, transcribed, analysed, or structured, credits are deducted.
How many credits are consumed depends on:
- which model is in use
- how much is processed
- which type of processing takes place – e.g. text processing, reasoning, OCR, embedding, or audio processing
The current credit values and model assignments can be found in the central credit overview in the Google Sheet .
Overview
| Modality | Typical billing |
|---|---|
| Text | by input, output, and optionally reasoning effort |
| Audio | per voice input or per minute (Konversa) |
| Images | flat rate per action (Image-to-Text) or by quality (image generation) |
| Documents | depending on text processing and OCR usage |
| Embeddings | by input tokens |
How the system works
Credits are not billed as a flat rate per product, but according to the type of processing.
Depending on the use case, different billing logics may apply:
- Text – when processing and generating text
- Reasoning Effort – when a model uses additional computational effort for in-depth reasoning
- Audio – for voice inputs, transcriptions, or meeting processing
- Images – for image generation or extracting content from images
- OCR – when scanned documents or images first need to be made machine-readable
- Embeddings – when text is converted into vectors, e.g. for search or knowledge access
Text
For text-based AI features, credits are calculated based on the amount of text processed and the computational logic used.
Three components are particularly relevant:
- Input – the text passed to the model
- Output – the text the model returns
- Reasoning Effort – additional computational effort for models that use in-depth reasoning
Not every model uses reasoning. When it is used, it can additionally increase credit consumption.
Audio
In the audio area, a distinction is made between dictation in chat and Konversa.
Dictation in chat
When speech is converted directly into text in the chat, billing is currently per voice input.
Konversa
Konversa is used for audio and meeting processing, e.g. for in-person meetings, uploaded recordings, or online meetings with a meeting bot.
Depending on usage, billing can be:
- per transcription minute
- per bot minute
- additionally via text processing, when further content such as summaries is generated or processed from a transcription
→ Learn more: Dictations, In-Person Meetings, Online Meetings
Images
Image generation
For image generation, credit consumption depends on the selected quality.
The general rule is:
- higher quality = higher credit consumption
- lower quality = lower credit consumption
The detailed values for image generation will be added here shortly.
Image-to-Text
When information is extracted from images, credit consumption is billed as a flat rate per action.
Documents
For documents, the decisive factor is which type of processing actually takes place.
- If a document already contains a readable text layer, the normal text costs apply for the processed text.
- If content first needs to be extracted from a scanned document or image, OCR is additionally used.
Depending on the use case, text costs and OCR costs may therefore be combined.
Embeddings
Embeddings are used to convert text into vectors – for example, for search, knowledge access, or semantic processing.
The input text is processed. Billing is therefore based on input tokens.
Further details
The complete and current overview of modalities, models, and credit values can be found in the Google Sheet with the credit overview .