Skip to main content

Supported LLMs

Use the Model Hub to securely configure and manage access to commercial and open-source LLMs for Glean Assistant and Agents.

The Model Hub offers a curated set of leading models from providers like OpenAI, Google, Anthropic, and Amazon. You can bring your own keys for providers or use Glean’s Universal Model Key to access pre‑procured models with guardrails.

To configure which models are available and choose models for workflows, see Configure LLMs in the Model Hub.

For details on each deployment model, see Glean deployment models.

note

GCP hosted customers cannot use AWS hosted models, for example, Amazon Bedrock. This change expands model choice on AWS while keeping cross‑cloud data-path restrictions in place.

Supported models

Glean supports the following model creators. Availability depends on your hosting environment and provider access:

  • OpenAI
  • Google Gemini
  • Anthropic
  • Amazon

Supported models for Glean Agents

Glean provides basic, standard, and premium models for Glean Agents. The model tier depends on the model's reasoning capabilities and overall cost to run.

Basic

OpenAI (via Azure or OpenAI)
  • GPT-5 mini
  • GPT-5 nano
  • GPT-4.1 mini
  • GPT-4.1 nano
  • GPT-4o mini
  • o3-mini
Google Gemini (via Google Vertex AI)
  • Gemini Flash Preview 3
  • Gemini Flash 2.5
  • Gemini Flash Lite Preview 2.5
Anthropic (via Google Vertex AI or Amazon Bedrock)
  • Claude Haiku 4.5
Amazon
  • Amazon Nova Pro 1.0

Standard

OpenAI (via Azure or OpenAI)
  • GPT-5.2
  • GPT-5.1
  • GPT-5
  • GPT-4.1
  • o3
  • o4-mini
Google Gemini (via Google Vertex AI)
  • Gemini Flash 3.5
  • Gemini Pro 3.1
  • Gemini Pro 2.5

Premium

OpenAI (via Azure or OpenAI)
  • GPT-5.6 Terra
  • GPT-5.6 Sol
  • GPT-5.6 Luna
  • GPT-5.5
  • GPT-5.4
  • o1
  • GPT Image 1.5
  • GPT Image 2
Anthropic (via Google Vertex AI or Amazon Bedrock)
  • Claude Opus 4.8
  • Claude Opus 4.7
  • Claude Opus 4.6
  • Claude Sonnet 4.6
Google Gemini (via Google Vertex AI)
  • Nano Banana Pro (Gemini Pro Image 3)
  • Gemini Flash Image 3.1
  • Gemini Flash Image 2.5

Supported Models for Glean Assistant

Glean provides basic, standard, and premium models for Glean Assistant. The model tier depends on the model's reasoning capabilities and overall cost to run. If your organization uses the Glean Universal Model Key, Glean optimizes your experience by using the best-in-class basic and standard models by default. You can also choose the model for Glean Assistant when you enable models in the Model Hub.

note

If you set one of the following premium models for Glean Assistant, Glean uses these models for both regular and premium requests:

  • GPT-5.4
  • Claude Sonnet 4.6
  • Claude Opus 4.6

Basic

Glean
  • Waldo (Glean Universal Model Key)
OpenAI (via Azure or OpenAI)
  • GPT-5 mini
  • GPT-5 nano
  • GPT-4.1 mini
  • GPT-4.1 nano
  • GPT-4o mini
Google Gemini (via Google Vertex AI)
  • Gemini Flash Preview 3
  • Gemini Flash 2.5
  • Gemini Flash Lite Preview 2.5
Anthropic (via Google Vertex AI or Amazon Bedrock)
  • Claude Haiku 4.5

Standard

OpenAI (via Azure or OpenAI)
  • GPT-5.2
  • GPT-5.1
  • GPT-5
  • GPT-4.1
  • o3
Google Gemini (via Google Vertex AI)
  • Gemini Flash 3.5 (Glean Universal Model Key)
  • Gemini Pro 3.1 (Glean Universal Model Key)
  • Gemini Pro Custom Tools 3.1 (Customer Key)
  • Gemini Pro 2.5
note

For Glean Assistant on Customer Key, Gemini Pro 3.1 is not available. Use Gemini Pro Custom Tools 3.1 instead. Gemini Pro 3.1 remains available for Glean Assistant through the Glean Universal Model Key.

Premium

Glean
  • Waldo (Customer Key)
OpenAI (via Azure or OpenAI)
  • GPT-5.6 Terra
  • GPT-5.6 Sol
  • GPT-5.5
  • GPT-5.4
  • o1
  • GPT Image 1.5
  • GPT Realtime 1.5
  • GPT-4o Transcribe
Anthropic (via Google Vertex AI or Amazon Bedrock)
  • Claude Opus 4.8
  • Claude Opus 4.7
  • Claude Opus 4.6
  • Claude Sonnet 4.6
Google Gemini (via Google Vertex AI)
  • Nano Banana Pro (Gemini Pro Image 3)
  • Gemini Flash Image 3.1
  • Gemini Flash Image 2.5
note

GPT-5.6 Luna is available for Glean Agents only. It isn't yet available for Glean Assistant model choice or Auto mode.

Availability by hosting environment and provider

Which LLM hosting providers you can configure depends on your deployment mode and key type:

  • Glean Universal Model Key: Glean manages connectivity to all supported providers regardless of your deployment's cloud environment.
  • Customer Key (BYOK): Provider access is limited to cloud-native providers because cross-cloud access is not supported.
Hosting providerGlean Universal Model KeyCustomer Key — GCP-based deploymentCustomer Key — AWS-based deployment
OpenAI
Azure OpenAI
Google Vertex AI (Gemini, Claude)
Amazon Bedrock (Claude, Amazon Nova)
  • GCP-based deployment includes Glean Hosted and Customer Hosted on GCP.
  • AWS-based deployment refers to Customer Hosted on AWS.
  • Anthropic (Claude) models are accessible through either Google Vertex AI or Amazon Bedrock, depending on your available providers.
  • For model-level differences between key modes, see the notes under Supported models for Glean Assistant.

Pricing

With Enterprise Flex pricing, each agent run uses an amount of FlexCredits determined by the complexity of an agent. This complexity includes how many connectors the agent searches, how many steps it takes, how much memory it maintains, how many tools it executes, and the model used for each step. Agents that use higher-tier models consume more credits than agents that use lower-tier models.

note

With Enterprise Flex pricing, everyday Glean Assistant queries that leverage basic and standard models don't use credits. If you enable premium models for advanced queries, the advanced queries consume credits.

See the following documentation to learn about pricing dashboards:

See also

  • To configure which models are available and choose models for workflows, see Configure LLMs in the Model Hub.
  • To learn how Glean handles model deprecation, including notification timelines, migration paths, and tools needed for assistants and agents, see Model deprecation.
  • To monitor LLM usage and reliability for Customer Key deployments, see LLM Insights.