Skip to content

IBM watsonx.ai

Cloud model platform · fee not published · checked Sep 23, 2026

Models on IBM watsonx.ai

3 model routes.

Premium: blended price (3 input to 1 output) on this platform against the lab's own API. Regions are separate routes.

About IBM watsonx.ai

IBM's AI studio and inference platform serving IBM Granite and third-party models (Meta, Mistral, OpenAI gpt-oss, others) with pay-as-you-go per-token pricing.

Fee on tokens
Not published
Pricing model
Per 1M tokens (billed as Resource Units, 1 RU = 1,000 tokens incl. input and output); Essentials plan from USD 0/month, Standard from USD 1110/month
Bring your own key
No
Entry plan
Usage based
Hosting
Hosted

Features stated: none found. Not stated: Fallbacks, Load balancing, Caching, Guardrails, Observability, Rate limits, Budgets, Prompts, OpenAI-compatible, Bring your own key, Data residency, Private networking.

The terms, in their words

Markup
Per 1M tokens (billed as Resource Units, 1 RU = 1,000 tokens incl. input and output); Essentials plan from USD 0/month, Standard from USD 1110/month
Free tier
Free Toolbox playground: up to 300,000 tokens per month
Providers
IBM Granite family plus third-party models from Meta, Google, DeepSeek, Mistral, OpenAI and more
  • For foundation model inference, charges are based on a Resource Unit (RU) metric equivalent to 1000 tokens (including both input and output tokens).
  • Foundation Models : Up to 300,000 tokens per month
  • Starting at USD 1110/month*

Source: www.ibm.com/products/watsonx-ai/pricing

IBM watsonx.ai vsAWS BedrockAzure AI FoundryGoogle Vertex AIOCI Generative AIHeliconeKong AI Gateway

Sources: each gateway's own pricing page and docs, read by hand and fetched again every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.