Supported LLMs
Use the Model Hub to securely configure and manage access to commercial and open-source LLMs for Glean Assistant and Agents.
The Model Hub offers a curated set of leading models from providers like OpenAI, Google, Anthropic, and Amazon. You can bring your own keys for providers or use Glean’s Universal Model Key to access pre‑procured models with guardrails.
To configure which models are available and choose models for workflows, see Configure LLMs in the Model Hub.
For details on each deployment model, see Glean deployment models.
GCP hosted customers cannot use AWS hosted models, for example, Amazon Bedrock. This change expands model choice on AWS while keeping cross‑cloud data-path restrictions in place.
Model pricing information
For model pricing and tier details, see Glean Core Suite pricing and Glean Enterprise Flex pricing.
Supported models
Model availability depends on your hosting environment and provider access. See Availability by hosting environment and provider.
If your organization uses the Glean Universal Model Key, Glean optimizes your experience by automatically selecting high-performance models by default. You can also choose a model in the Model Hub.
Text models
The following table lists available text models in Glean:
| Provider | Models |
|---|---|
| OpenAI (via Azure or OpenAI) |
|
| Google Gemini (via Google Vertex AI) |
|
| Anthropic (via Google Vertex AI or Amazon Bedrock) |
|
| Amazon |
|
| Glean |
|
| Fireworks/Baseten |
|
| Baseten |
|
Image models
The following table lists supported image models:
| Provider | Models |
|---|---|
| OpenAI (via Azure or OpenAI) |
|
| Google Gemini (via Google Vertex AI) |
|
Availability and key-mode notes
Unless noted below, a model is available for both Glean Assistant and Glean Agents.
- Waldo (Glean Assistant) is available with both the Glean Universal Model Key and a Customer Key.
- Gemini 3.1 Pro Custom Tools is the Customer Key variant of Gemini 3.1 Pro for Glean Assistant; Gemini 3.1 Pro is available through the Glean Universal Model Key.
- Nemotron 3 Ultra is available only through the Glean Universal Model Key.
- Open models, such as GLM 5.2 and Nemotron 3 Ultra, are controlled by creator region on Glean Universal Model Key deployments. See Manage open models by region.
Waldo doesn't currently run with Gemini models.
If you set one of the following models for Glean Assistant, Glean uses it for both regular and advanced requests: GPT-5.4, Claude Sonnet 4.6, and Claude Opus 4.6.
Availability by hosting environment and provider
Which LLM hosting providers you can configure depends on your deployment mode and key type:
- Glean Universal Model Key: Glean manages connectivity to all supported providers regardless of your deployment's cloud environment.
- Customer Key (BYOK): Access to cloud-hosted providers is limited to your own cloud because cross-cloud access is not supported. Anthropic is the exception: you connect to it directly with your own API key, so it works from either cloud.
| Hosting provider | Glean Universal Model Key | Customer Key — GCP-based deployment | Customer Key — AWS-based deployment |
|---|---|---|---|
| OpenAI | ✅ | ✅ | ✅ |
| Azure OpenAI | ✅ | ✅ | ✅ |
| Google Vertex AI | ✅ | ✅ | ❌ |
| Amazon Bedrock | ✅ | ❌ | ✅ |
| Anthropic | ❌ | ✅ | ✅ |
- GCP-based deployment includes Glean Hosted and Customer Hosted on GCP.
- AWS-based deployment refers to Customer Hosted on AWS.
- Anthropic (Claude) models are accessible through Google Vertex AI or Amazon Bedrock. With Customer Key, you can also connect directly to Anthropic using your own Anthropic API key, from either a GCP-based or AWS-based deployment.
- For model-level differences between key modes, see the notes under Availability and key-mode notes.
Pricing
With Enterprise Flex pricing, each agent run uses an amount of FlexCredits determined by the complexity of an agent. This complexity includes how many connectors the agent searches, how many steps it takes, how much memory it maintains, how many tools it executes, and the model used for each step. Agents that use higher-tier models consume more credits than agents that use lower-tier models.
With Enterprise Flex pricing, everyday Glean Assistant queries that leverage basic and standard models don't use credits. If you enable premium models for advanced queries, the advanced queries consume credits.
See the following documentation to learn about pricing dashboards:
See also
- To configure which models are available and choose models for workflows, see Configure LLMs in the Model Hub.
- To control open-model availability by creator region, see Manage open models by region.
- To learn how Glean handles model deprecation, including notification timelines, migration paths, and actions needed for assistants and agents, see Model deprecation.
- To monitor LLM usage and reliability for Customer Key deployments, see LLM Insights.
- To change your default model on a Customer Key deployment, see Change your default model on Customer Key.