Foundation models in watsonx.ai 

Explore the IBM library of AI models available in the watsonx.ai studio

Product screenshot of watsonx.ai foundation models UI dashboard

Choose the model you need

Select the IBM® Granite®, open-source or third-party model best suited for your business and deploy on-prem or in the cloud.

IBM's POV on AI models Choose the right foundation model

What’s new?

New model
Granite-4-1-3b is now available in the watsonx® foundation model library.
New model feature
Granite-4-1-8b is now available in the watsonx® foundation model library.
New model feature
Granite-4-1-30b is now available in the watsonx® foundation model library.
New model feature
Granite-vision-4-1-4b is now available in the watsonx® foundation model library.
New model feature

Foundation model library

Choose the model that best fits your specific use case, budget considerations, regional interests and risk profile.

View the embedding model library
IBM logo
IBM models

Tailored for business, IBM Granite family of open, performant and trusted models deliver exceptional performance at a competitive price, without compromising safety.

View the IBM model library Learn more about Granite
Meta logo
Meta Llama models

Llama models are open, efficient large language models designed for versatility and strong performance across a wide range of natural language tasks.

View the Meta model library Learn more about our partnership
Mistral logo
Mistral AI models

Mistral models are fast, performant, open-weight language models designed for modularity and optimized for text generation, reasoning and multilingual applications.

View the Mistral model library
Illustration of magnifying glass
Other third-party model providers

There are several foundation models from other providers available on watsonx.ai.

View the model library
Model Name Model Provider Model Consumption and Deployment Availability

granite-speech-4-1-2b

New

IBM

Deploy on demand only

granite-4-1-3b

New

IBM

Deploy on demand only

granite-4-1-8b

New

IBM

Deploy on demand only

granite-4-1-30b

New

IBM

Deploy on demand only

granite-vision-4-1-4b

New

IBM

Deploy on demand only

granite-4h-micro

IBM

Deploy on demand only

granite-4h-tiny

IBM

Deploy on demand only

 

granite-4h-small

IBM

Both Pay as you go and Deploy on demand

granite-vision-3-3-2b

IBM

Deploy on demand only

granite-3-3-2b-instruct

IBM

Deploy on demand only

granite-3-3-8b-instruct

IBM

Deploy on demand only

granite-3-2-8b-instruct

IBM

Deploy on demand only

granite-3-1-8b-base

IBM

Deploy on demand only

granite-3-1-8b-instruct

IBM

Deploy on demand only

granite-3-8b-base

IBM

Deploy on demand only

granite-13b-chat-v2

IBM

Deploy on demand only

granite-3b-code-instruct

IBM

Deploy on demand only

granite-8b-code-instruct

IBM

Both Pay as you go and Deploy on demand

granite-20b-code-base-sql-gen

IBM

Deploy on demand only

granite-20b-code-instruct

IBM

Deploy on demand only

granite-20b-multilingual

IBM

Deploy on demand only

granite-20b-code-base-schema-linking

IBM

Deploy on demand only

granite-34b-code-instruct

IBM

Deploy on demand only

granite-7b-lab

IBM

Deploy on demand only

granite-8b-japanese

IBM

Deploy on demand only

granite-guardian-3-8b

IBM

Pay as you go only

*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.

Model Name Model Provider Model Consumption and Deployment Availability

llama-4-maverick-17b-128e-instruct-int4

Meta

Deploy on demand only

llama-4-maverick-17b-128e-instruct-fp8

Meta

Both Pay as you go and Deploy on demand

llama-4-scout-17b-16e-instruct-fp8-dynamic

Meta

Deploy on demand only

llama-3-3-70b-instruct

Meta

Both Pay as you go and Deploy on demand

llama-3-3-70b-instruct-hf

Meta

Deploy on demand only

llama-3-2-11b-vision-instruct

Meta

Pay as you go only

llama-3-2-90b-vision-instruct

Meta

Deploy on demand only

llama-3-1-405b-instruct-fp8

Meta

Deploy on demand only

llama-3-1-8b

Meta

Deploy on demand only

llama-3-1-8b-instruct

Meta

Deploy on demand only

llama-3-1-70b

Meta

Deploy on demand only

llama-3-1-70b-gptq

Meta

Deploy on demand only

llama-3-1-70b-instruct

Meta

Deploy on demand only

llama-3-1-405b-instruct-fp8

Meta

Deploy on demand only

llama-3-8b-instruct

Meta

Deploy on demand only

llama-3-70b-instruct

Meta

Deploy on demand only

llama-2-70b-chat

Meta

Deploy on demand only

codellama-34b-instruct-hf

Meta

Deploy on demand only

llama-guard-3-11b-vision

Meta

Pay as you go only

*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.

Mistral models

Model Name Model Provider Model Consumption and Deployment Availability

mistral-medium-3-5-0

New

Mistral    

Deploy on demand only

devstral-small-2512

Mistral    

Deploy on demand only

mistral-large-2512

Mistral    

Both Pay as you go and Deploy on demand

ministral-14b-instruct-2512

Mistral    

Deploy on demand only

ministral-8b-instruct-2512

Mistral    

Deploy on demand only

ministral-3b-instruct-2512

Mistral    

Deploy on demand only

mistral-medium-2508

Mistral    

Deploy on demand only

devstral-medium-2507

Mistral    

Deploy on demand only

devstral-small-2507

Mistral    

Deploy on demand only

mistral-small-3-2-24b-instruct-2506

Mistral    

Deploy on demand only

mistral-medium-2505

Mistral    

Both Pay as you go and Deploy on demand

mistral-small-3-1-24b-instruct-2503

Mistral    

Both Pay as you go and Deploy on demand

codestral-2501

Mistral    

Deploy on demand only

mistral-large-instruct-2411

Mistral    

Deploy on demand only

ministral-8b-instruct-2410

Mistral    

Deploy on demand only

mistral-large-instruct-2407

Mistral    

Deploy on demand only

mistral-nemo-instruct-2407

Mistral    

Deploy on demand only

mixtral-8x7b-base

Mistral    

Deploy on demand only

mixtral-8x7b-instruct-v01

Mistral    

Deploy on demand only

pixtral-12b

Mistral

Deploy on demand only

*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.

Other third-party foundation models

Model Name Model Provider Model Consumption and Deployment Availability

nvidia-nemotron-3-ultra-550b-a55b-quantized

New

Red Hat

Deploy on demand only

gpt-oss-120b

Open AI

Both Pay as you go and Deploy on demand

gpt-oss-20b

Open AI

Deploy on demand only

deepseek-r1-distill-llama-8b

DeepSeek

Deploy on demand only

deepseek-r1-distill-llama-70b

DeepSeek

Deploy on demand only

allam-1-13b-instruct

SDAIA

Deploy on demand only

llama-3-1-nemotron-ultra-253b-v1-fp8

NVIDIA

Deploy on demand only

nvidia-nemotron-3-super-120b-a12b-fp8

NVIDIA

Deploy on demand only

nvidia-nemotron-nano-12b-v2-vl-fp8

NVIDIA

Deploy on demand only

eurollm-1-7b-instruct

Utter Project

Deploy on demand only

eurollm-9b-instruct

Utter Project

Deploy on demand only

mt0-xxl-13b

BigScience

Deploy on demand only

poro-34b-chat

LumiOpen

Deploy on demand only

*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.

Embedding model library

Model Name Model Consumption and Deployment Availability

granite-embedding-278m-multilingual

Pay as you go only

slate-125m-english-rtrvr-v2

Pay as you go only

slate-30m-english-rtrvr-v2

Pay as you go only

*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.

Third-party embedding models

Model Name Model Provider Model Consumption and Deployment Availability

all-mini-l6-v2

Microsoft

Pay as you go only

multilingual-e5-large

Intel

Pay as you go only

*Prices shown are indicative, may vary by country, exclude any applicable taxes and duties, and are subject to product offering availability in a locale.

Client stories

What happens when you train a powerful AI model with your own unique data? Better customer experiences and faster value with AI. Explore these stories and see how.

Wimbledon logo
Wimbledon

Wimbledon used watsonx.ai foundation models to train its AI to create tennis commentary.

Read the case study
Recording Academy logo
The Recording Academy

The Recording Academy used AI Stories with IBM watsonx to generate and scale editorial content around GRAMMY nominees.

Read the announcement
Masters logo
The Masters

The Masters uses watsonx.ai to bring AI-powered hole insights combined with expert opinions to digital platforms.

Read the announcement
AddAI.Life logo
AddAI.Life

AddAI.Life uses watsonx.ai to access selected open-source large language models to build higher quality virtual assistants.

Read the case study
Take the next step

Start operationalizing and scaling generative AI and machine learning for business by exploring our free trial or booking a live demo.

  1. Start your free trial
  2. Book a live demo
Intellectual property protection for IBM-developed watsonx.ai models

IBM believes in the creation, deployment and utilization of AI models that advance innovation across the enterprise responsibly. IBM watsonx AI portfolio has an end-to-end process for building and testing foundation models and generative AI. For IBM-developed models, we search for and remove duplication, and we employ URL blocklists, filters for objectionable content and document quality, sentence splitting and tokenization techniques, all before model training.

During the data training process, we work to prevent misalignments in the model outputs and use supervised fine-tuning to enable better instruction following so that the model can be used to complete enterprise tasks through prompt engineering. We are continuing to develop the Granite models in several directions, including other modalities, industry-specific content and more data annotations for training, while also deploying regular, ongoing data protection safeguards for IBM developed-models.  

Given the rapidly changing generative AI technology landscape, our end-to-end processes are expected to continuously evolve and improve. As a testament to the rigor IBM puts into the development and testing of its foundation models, the company provides its standard contractual intellectual property indemnification for IBM-developed models, similar to those it provides for IBM hardware and software products.

Moreover, contrary to some other providers of large language models and consistent with the IBM standard approach on indemnification, IBM does not require its customers to indemnify IBM for a customer’s use of IBM-developed models. Also, consistent with the IBM approach to its indemnification obligation, IBM does not cap its indemnification liability for the IBM-developed models.

The current watsonx models now under these protections include:

(1) Slate family of encoder-only models

(2) Granite family of a decoder-only model

Learn more about licensing for Granite models (PDF)

Footnotes

* Supported context length by model provider, but actual context length on platform is limited. For more information, please see Documentation.

Inference is billed in Resource Units. 1 Resource Unit is 1,000 tokens. Input and completion tokens are charged at the same rate. 1,000 tokens are generally about 750 words.

Not all models are available in all regions. See our documentation for details.

Context length is expressed in tokens.

The IBM statements regarding its plans, directions and intent are subject to change or withdrawal without notice at its sole discretion. See Pricing for more details. Unless otherwise specified under Software pricing, all features, capabilities and potential updates refer exclusively to SaaS. IBM makes no representation that SaaS and software features and capabilities are the same.