Usage Limits & Fair Use Policy

innoGPT offers a fair flat-rate price—to ensure this works for everyone, transparent usage limits are in place to guarantee that the platform remains stable and fast for every user.

Why are there limits at all?

Every request incurs real costs behind the scenes: innoGPT pays the respective model providers (e.g., OpenAI or Anthropic) for each token used. At the same time, we offer you a fixed flat-rate price. The limits are set generously. To ensure this remains sustainable and the platform stays stable for all users, we have a Fair Use Policy in place.

What uses up a lot of tokens?

Not all requests cost the same. Two factors are particularly significant:

  • Ultra and Premium models, such as Claude Opus or GPT-5.5, cost significantly more per request than leaner models. If you work exclusively with these models, you’ll reach the limit much faster.

  • Deep Research performs many individual requests in the background and therefore consumes a particularly large number of tokens per run—which adds up quickly.

What happens when the limit is reached?

innoGPT does not take immediate, drastic action. Instead, there are tiered measures: First, a temporary rate limit may be imposed. In the next step, computationally intensive models are temporarily restricted and redirected to more resource-efficient alternatives. If the limit is consistently exceeded, innoGPT will proactively reach out to work with you on finding a more suitable plan.

Tip for smart usage

For simple tasks like text summaries, short answers, or standard research, leaner models are perfectly sufficient. You should use Ultra models specifically for complex queries—that way, you’ll get the most out of your quota.


Available Plans & Models

Depending on your plan, different model categories are available to you. As a general rule: The higher the plan tier, the more powerful (and computationally intensive) models are unlocked.

Which plan includes which model categories?

  • Personal / Pro / Business / Partner / Family: Standard, Premium & Ultra

  • Go: Smart Select with Standard models only

  • Trial — 7-day trial period: Standard


Usage Scope: What does “unlimited messages” mean?

innoGPT Does not use strict message quotas per user, but rather a pooled usage budget per workspace across all users.

What’s included:

  • Standard messages

What is billed separately:

  • API usage

  • Add-ons such as PII, videos, and podcasts

How does the Workspace budget work?

  • More expensive models (Premium/Reasoning models) consume more budget per request

  • Lower-cost models consume correspondingly less

  • When the budget is reached, a soft cap kicks in: Premium models are restricted, while Standard models remain fully available

💡 What does this mean in practice? No one “runs out.” As soon as the Premium budget is exhausted, users can seamlessly continue working with the more efficient models—no hard cutoff, no blocked workflow.