Explore the IBM library of AI models available in the watsonx.ai studio
Select the IBM® Granite®, open-source or third-party model best suited for your business and deploy on-prem or in the cloud.
Choose the model that best fits your specific use case, budget considerations, regional interests and risk profile.
Tailored for business, IBM Granite family of open, performant and trusted models deliver exceptional performance at a competitive price, without compromising safety.
Llama models are open, efficient large language models designed for versatility and strong performance across a wide range of natural language tasks.
Mistral models are fast, performant, open-weight language models designed for modularity and optimized for text generation, reasoning and multilingual applications.
There are several foundation models from other providers available on watsonx.ai.
granite-speech-4-1-2b
IBM
Deploy on demand only
granite-4-1-3b
IBM
Deploy on demand only
granite-4-1-8b
IBM
Deploy on demand only
granite-4-1-30b
IBM
Deploy on demand only
granite-vision-4-1-4b
IBM
Deploy on demand only
granite-4h-micro
IBM
Deploy on demand only
granite-4h-tiny
IBM
Deploy on demand only
granite-4h-small
IBM
Both Pay as you go and Deploy on demand
granite-vision-3-3-2b
IBM
Deploy on demand only
granite-3-3-2b-instruct
IBM
Deploy on demand only
granite-3-3-8b-instruct
IBM
Deploy on demand only
granite-3-2-8b-instruct
IBM
Deploy on demand only
granite-3-1-8b-base
IBM
Deploy on demand only
granite-3-1-8b-instruct
IBM
Deploy on demand only
granite-3-8b-base
IBM
Deploy on demand only
granite-13b-chat-v2
IBM
Deploy on demand only
granite-3b-code-instruct
IBM
Deploy on demand only
granite-8b-code-instruct
IBM
Both Pay as you go and Deploy on demand
granite-20b-code-base-sql-gen
IBM
Deploy on demand only
granite-20b-code-instruct
IBM
Deploy on demand only
granite-20b-multilingual
IBM
Deploy on demand only
granite-20b-code-base-schema-linking
IBM
Deploy on demand only
granite-34b-code-instruct
IBM
Deploy on demand only
granite-7b-lab
IBM
Deploy on demand only
granite-8b-japanese
IBM
Deploy on demand only
granite-guardian-3-8b
IBM
Pay as you go only
*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.
llama-4-maverick-17b-128e-instruct-int4
Meta
Deploy on demand only
llama-4-maverick-17b-128e-instruct-fp8
Meta
Both Pay as you go and Deploy on demand
llama-4-scout-17b-16e-instruct-fp8-dynamic
Meta
Deploy on demand only
llama-3-3-70b-instruct
Meta
Both Pay as you go and Deploy on demand
llama-3-3-70b-instruct-hf
Meta
Deploy on demand only
llama-3-2-11b-vision-instruct
Meta
Pay as you go only
llama-3-2-90b-vision-instruct
Meta
Deploy on demand only
llama-3-1-405b-instruct-fp8
Meta
Deploy on demand only
llama-3-1-8b
Meta
Deploy on demand only
llama-3-1-8b-instruct
Meta
Deploy on demand only
llama-3-1-70b
Meta
Deploy on demand only
llama-3-1-70b-gptq
Meta
Deploy on demand only
llama-3-1-70b-instruct
Meta
Deploy on demand only
llama-3-1-405b-instruct-fp8
Meta
Deploy on demand only
llama-3-8b-instruct
Meta
Deploy on demand only
llama-3-70b-instruct
Meta
Deploy on demand only
llama-2-70b-chat
Meta
Deploy on demand only
codellama-34b-instruct-hf
Meta
Deploy on demand only
llama-guard-3-11b-vision
Meta
Pay as you go only
*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.
mistral-medium-3-5-0
Mistral
Deploy on demand only
devstral-small-2512
Mistral
Deploy on demand only
mistral-large-2512
Mistral
Both Pay as you go and Deploy on demand
ministral-14b-instruct-2512
Mistral
Deploy on demand only
ministral-8b-instruct-2512
Mistral
Deploy on demand only
ministral-3b-instruct-2512
Mistral
Deploy on demand only
mistral-medium-2508
Mistral
Deploy on demand only
devstral-medium-2507
Mistral
Deploy on demand only
devstral-small-2507
Mistral
Deploy on demand only
mistral-small-3-2-24b-instruct-2506
Mistral
Deploy on demand only
mistral-medium-2505
Mistral
Both Pay as you go and Deploy on demand
mistral-small-3-1-24b-instruct-2503
Mistral
Both Pay as you go and Deploy on demand
codestral-2501
Mistral
Deploy on demand only
mistral-large-instruct-2411
Mistral
Deploy on demand only
ministral-8b-instruct-2410
Mistral
Deploy on demand only
mistral-large-instruct-2407
Mistral
Deploy on demand only
mistral-nemo-instruct-2407
Mistral
Deploy on demand only
mixtral-8x7b-base
Mistral
Deploy on demand only
mixtral-8x7b-instruct-v01
Mistral
Deploy on demand only
pixtral-12b
Mistral
Deploy on demand only
*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.
nvidia-nemotron-3-ultra-550b-a55b-quantized
Red Hat
Deploy on demand only
gpt-oss-120b
Open AI
Both Pay as you go and Deploy on demand
gpt-oss-20b
Open AI
Deploy on demand only
deepseek-r1-distill-llama-8b
DeepSeek
Deploy on demand only
deepseek-r1-distill-llama-70b
DeepSeek
Deploy on demand only
allam-1-13b-instruct
SDAIA
Deploy on demand only
llama-3-1-nemotron-ultra-253b-v1-fp8
NVIDIA
Deploy on demand only
nvidia-nemotron-3-super-120b-a12b-fp8
NVIDIA
Deploy on demand only
nvidia-nemotron-nano-12b-v2-vl-fp8
NVIDIA
Deploy on demand only
eurollm-1-7b-instruct
Utter Project
Deploy on demand only
eurollm-9b-instruct
Utter Project
Deploy on demand only
mt0-xxl-13b
BigScience
Deploy on demand only
poro-34b-chat
LumiOpen
Deploy on demand only
*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.
*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.
*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.
What happens when you train a powerful AI model with your own unique data? Better customer experiences and faster value with AI. Explore these stories and see how.
Wimbledon used watsonx.ai foundation models to train its AI to create tennis commentary.
The Recording Academy used AI Stories with IBM watsonx to generate and scale editorial content around GRAMMY nominees.
The Masters uses watsonx.ai to bring AI-powered hole insights combined with expert opinions to digital platforms.
AddAI.Life uses watsonx.ai to access selected open-source large language models to build higher quality virtual assistants.
IBM believes in the creation, deployment and utilization of AI models that advance innovation across the enterprise responsibly. IBM watsonx AI portfolio has an end-to-end process for building and testing foundation models and generative AI. For IBM-developed models, we search for and remove duplication, and we employ URL blocklists, filters for objectionable content and document quality, sentence splitting and tokenization techniques, all before model training.
During the data training process, we work to prevent misalignments in the model outputs and use supervised fine-tuning to enable better instruction following so that the model can be used to complete enterprise tasks through prompt engineering. We are continuing to develop the Granite models in several directions, including other modalities, industry-specific content and more data annotations for training, while also deploying regular, ongoing data protection safeguards for IBM developed-models.
Given the rapidly changing generative AI technology landscape, our end-to-end processes are expected to continuously evolve and improve. As a testament to the rigor IBM puts into the development and testing of its foundation models, the company provides its standard contractual intellectual property indemnification for IBM-developed models, similar to those it provides for IBM hardware and software products.
Moreover, contrary to some other providers of large language models and consistent with the IBM standard approach on indemnification, IBM does not require its customers to indemnify IBM for a customer’s use of IBM-developed models. Also, consistent with the IBM approach to its indemnification obligation, IBM does not cap its indemnification liability for the IBM-developed models.
The current watsonx models now under these protections include:
(1) Slate family of encoder-only models
(2) Granite family of a decoder-only model
* Supported context length by model provider, but actual context length on platform is limited. For more information, please see Documentation.
Inference is billed in Resource Units. 1 Resource Unit is 1,000 tokens. Input and completion tokens are charged at the same rate. 1,000 tokens are generally about 750 words.
Not all models are available in all regions. See our documentation for details.
Context length is expressed in tokens.
The IBM statements regarding its plans, directions and intent are subject to change or withdrawal without notice at its sole discretion. See Pricing for more details. Unless otherwise specified under Software pricing, all features, capabilities and potential updates refer exclusively to SaaS. IBM makes no representation that SaaS and software features and capabilities are the same.