Skip to content
Platform

Credits

Credits are the billing unit for using AI features. Whenever content is processed, generated, transcribed, analysed, or structured, credits are deducted.

How many credits are consumed depends on:

  • which model is in use
  • how much is processed
  • which type of processing takes place – e.g. text processing, reasoning, OCR, embedding, or audio processing

The current credit values and model assignments can be found in the central credit overview in the Google Sheet .


Overview

Modality Typical billing
Text by input, output, and optionally reasoning effort
Audio per voice input or per minute (Konversa)
Images flat rate per action (Image-to-Text) or by quality (image generation)
Documents depending on text processing and OCR usage
Embeddings by input tokens

How the system works

Credits are not billed as a flat rate per product, but according to the type of processing.

Depending on the use case, different billing logics may apply:

  • Text – when processing and generating text
  • Reasoning Effort – when a model uses additional computational effort for in-depth reasoning
  • Audio – for voice inputs, transcriptions, or meeting processing
  • Images – for image generation or extracting content from images
  • OCR – when scanned documents or images first need to be made machine-readable
  • Embeddings – when text is converted into vectors, e.g. for search or knowledge access

Text

For text-based AI features, credits are calculated based on the amount of text processed and the computational logic used.

Three components are particularly relevant:

  • Input – the text passed to the model
  • Output – the text the model returns
  • Reasoning Effort – additional computational effort for models that use in-depth reasoning

Not every model uses reasoning. When it is used, it can additionally increase credit consumption.


Audio

In the audio area, a distinction is made between dictation in chat and Konversa.

Dictation in chat

When speech is converted directly into text in the chat, billing is currently per voice input.

Konversa

Konversa is used for audio and meeting processing, e.g. for in-person meetings, uploaded recordings, or online meetings with a meeting bot.

Depending on usage, billing can be:

  • per transcription minute
  • per bot minute
  • additionally via text processing, when further content such as summaries is generated or processed from a transcription

→ Learn more: Dictations, In-Person Meetings, Online Meetings


Images

Image generation

For image generation, credit consumption depends on the selected quality.

The general rule is:

  • higher quality = higher credit consumption
  • lower quality = lower credit consumption

The detailed values for image generation will be added here shortly.

Image-to-Text

When information is extracted from images, credit consumption is billed as a flat rate per action.


Documents

For documents, the decisive factor is which type of processing actually takes place.

  • If a document already contains a readable text layer, the normal text costs apply for the processed text.
  • If content first needs to be extracted from a scanned document or image, OCR is additionally used.

Depending on the use case, text costs and OCR costs may therefore be combined.


Embeddings

Embeddings are used to convert text into vectors – for example, for search, knowledge access, or semantic processing.

The input text is processed. Billing is therefore based on input tokens.


Further details

The complete and current overview of modalities, models, and credit values can be found in the Google Sheet with the credit overview .