Skip to main content

AI Studio Billing

How AI Studio bills for services across its token and per-patch pricing models.

This guide provides a detailed overview of how billing works in Hyperstack AI Studio. It explains which services incur charges, how those charges are calculated, and where to find usage and cost data within the platform. It also introduces the two billing models used in AI Studio: token-based pricing for text-processing services and per-patch pricing for image generation.

Overview of AI Studio Billing Model​

Hyperstack AI Studio uses two billing models to accommodate different types of services:

  • Token-Based Pricing: Applies to services that process text, such as inference and the Text Playground. Billing is based on the number of tokens used:

    • Input Tokens and Output Tokens are priced separately.
    • Usage is calculated per 1 million tokens.
  • Per-Patch Pricing: Applies to image generation using vision models in the Image Playground. Billing is based on the number of 16×16 patches produced, where 1 patch = 256 pixels (a 16×16 pixel region) of generated image output. Larger image sizes produce more patches and therefore cost more.

Billed AI Studio Services​

The table below summarizes each billable service in AI Studio, its purpose, and billing model:

ServiceDescriptionBilling Model
Serverless Inference & Text PlaygroundRun inference through the UI Playground or API using any model in the catalog.Token-based
Image PlaygroundGenerate or edit images using vision models (text-to-image, image-to-image).Per 16×16 patch
Dedicated InferenceServe an open-weight model on GPUs reserved for your organization, behind a private endpoint.Hourly (flavor rate)

Accessing Billing Information​

To view your AI Studio billing data:

  1. Go to the Billing page in Hyperstack.
  2. On the Overview tab, view a summary of your current usage and total AI Studio costs.
  3. For service-specific details, navigate to the Resource Activity tab.
  4. Select any listed service Usage Report to view details (e.g., Serverless Inference, Image Generation).

Within the usage reports you will see:

  • For Token-Based Services (Serverless Inference & Text Playground):

    • Resource Name – The inference request.
    • Model – The model used during the request.
    • Price Per 1M Tokens – The price per 1 million tokens for the request, across its input and output tokens.
    • Total Tokens – Combined total of input and output tokens used.
    • Total Cost – Total cost for that usage entry.
  • For Image-Based Services (Image Playground):

    • Resource Name – The resource or job name.
    • Model – The vision model used.
    • Unit Metric – The billing unit used for image generation. For image generation, the unit is a 16×16 Patch (1 patch = 256 pixels, or a 16×16 pixel region of generated output).
    • Price Per 16×16 Patch – The charge applied per 16×16 patch.
    • Total 16×16 Patches – Total patches generated.
    • Total Cost – Total cost for that usage entry.
Past usage reports

Usage reports for services that are no longer offered in AI Studio remain available for previous billing cycles, so you can still review what you were charged.

Pricing and Account Balance​

  • Token and patch usage are calculated every minute and billed after use.
  • Your AI Studio usage draws from a shared account balance used across Hyperstack services.
  • Access to services is suspended if your balance falls below the required minimum.
General Hyperstack Billing Policies

AI Studio services follow the same account, billing, and payment policies as the rest of Hyperstack. Learn more.