Skip to content

Choosing an AI Model ​

Quick Answer ​

Use the model selector dropdown in the chat composer to choose which AI model powers your conversation. Zeus.ai Secure (Third-Party) is the default model — a cached, SiteZeus-paid DeepSeek model hosted by Fireworks AI under no-retention terms. The catalog covers twelve models across four providers: DeepSeek, OpenAI (GPT), Claude (Anthropic), and Gemini (Google). Accounts with the Clean Room subscription also have access to two data-resident Zeus.ai Clean Room models that run entirely inside the Zeus.ai Azure tenant (US data zone).

Model availability is governed company-wide. By default every company has access to the two Clean Room models and Zeus.ai Secure (Third-Party). Additional public models are opt-in and must be enabled by a company admin or SiteZeus support.

The chat composer's model selector open, listing the available AI models with their context sizes. The active model is highlighted.


Model Availability and Locks ​

Some models may appear with a lock icon in the selector. There are two reasons a model can be locked:

  • Clean Room subscription required — the two Zeus.ai Clean Room models are gated behind the Clean Room add-on. Clicking a locked Clean Room model opens an upgrade prompt. The Clean Room is free while in beta; contact support@sitezeus.com to enable access.
  • Policy: ask your administrator — the model is not on your company's allowed list. A company admin or SiteZeus support can enable it from Company Settings (profile menu) or the support-admin panel.

Available Models ​

Zeus.ai Secure (Third-Party) — Default ​

  • Description: Cached — third-party host
  • Context: 1M tokens (1,000,000)
  • File support: Text and code only — no images, PDFs, audio, video, or URLs
  • Default: Yes — this is the model selected when you open a new thread
  • Data handling: Hosted by Fireworks AI via Azure Foundry, outside the Zeus.ai Azure tenant. Fireworks applies zero data retention by default and does not train on customer data. SOC 2 Type II and HIPAA certified.
  • Use when: You want a fast, capable default model for everyday Zeus.ai tasks

Zeus.ai Clean Room (subscription required) ​

Zeus.ai Clean Room models run entirely inside the Zeus.ai Azure tenant (US data zone). Your prompts are processed without leaving the Azure environment and are never sent to the model vendor's public cloud.

Access: Requires the Clean Room add-on. The Clean Room is free while in beta. When your account is not entitled, a lock icon appears next to each Clean Room model; clicking it opens an upgrade prompt. Contact support@sitezeus.com to enable access. Company admins can also toggle the Clean Room on or off from Company Settings.

Zeus.ai Clean Room ​

  • Description: Private — US data zone
  • Context: 1M tokens (1,000,000)
  • File support: Text and code only — no images, PDFs, audio, video, or URLs
  • Data handling: DeepSeek V4 Pro on Azure AI Foundry (first-party US data zone). No prompt caching. No training or retention.
  • Use when: You need data-resident chat for text-based analytics tasks

Zeus.ai Clean Room (GPT-5) ​

  • Description: Private — US data zone
  • Context: 272K tokens
  • File support: Images, PDFs, URLs, and text
  • Data handling: Azure OpenAI GPT-5 hosted in the US Data Zone (DataZoneStandard). Data resident in the US; no external egress.
  • Use when: You need document or image analysis in a data-resident environment

DeepSeek V4 Pro ​

  • Description: Max Thinking
  • Context: 1M tokens (1,000,000)
  • File support: Text and code only — no images, PDFs, audio, video, or URLs
  • Availability: Requires admin opt-in (policy-gated) or a company DeepSeek API key
  • Best for: Deep reasoning tasks when you want the direct DeepSeek API
  • Use when: Your admin has enabled it and you want max-thinking reasoning on text

Claude Models (Anthropic) ​

All Claude models require admin opt-in unless your company has added an Anthropic API key or a SiteZeus admin has enabled them.

Claude Opus 4.8 ​

  • Description: Most Intelligent Claude
  • Context: 1M tokens (1,000,000)
  • File support: Images, PDFs, URLs, and text
  • Best for: Complex multi-step analysis requiring deep Claude reasoning
  • Use when: You need the highest-capability Claude model

Claude Sonnet 4.6 ​

  • Description: Speed + Intelligence
  • Context: 1M tokens (1,000,000)
  • File support: Images, PDFs, URLs, and text
  • Best for: General-purpose chat, analytics workflows, tool-backed tasks
  • Use when: You need a balanced Claude model for image or document analysis

Claude Haiku 4.5 ​

  • Description: Fast & Lightweight
  • Context: 200K tokens
  • File support: Images, PDFs, URLs, and text
  • Best for: Quick lookups and short questions
  • Use when: Speed matters more than depth

OpenAI GPT Models ​

All OpenAI models require admin opt-in unless your company has added an OpenAI API key or a SiteZeus admin has enabled them.

GPT-5.5 ​

  • Description: Flagship Intelligence
  • Context: 1M tokens (1,000,000)
  • File support: Images, PDFs, URLs, and text
  • Best for: Demanding analytics workflows where maximum OpenAI capability is needed
  • Use when: You want OpenAI's most capable model

GPT-5.4 ​

  • Description: Flagship Intelligence
  • Context: 1M tokens (1,000,000)
  • File support: Images, PDFs, URLs, and text
  • Best for: Strong general-purpose analytics
  • Use when: You want a capable OpenAI model at lower cost than GPT-5.5

GPT-5.4 Mini ​

  • Description: Fast + Affordable
  • Context: 1M tokens (1,000,000)
  • File support: Images, PDFs, URLs, and text
  • Best for: Quicker tasks where response speed and cost matter
  • Use when: You need OpenAI's fast, lightweight option at 1M context

Gemini Models (Google) ​

Gemini models are fully multimodal — they also support audio and video files. All Gemini models require admin opt-in unless your company has added a Google API key or a SiteZeus admin has enabled them.

Gemini 3.1 Pro (Preview) ​

  • Description: Advanced Reasoning (Preview)
  • Context: 1M tokens (1,000,000)
  • File support: Images, audio, video, PDFs, URLs, and text
  • Best for: Large document analysis, deep multimodal reasoning
  • Use when: You need extensive context plus the newest Gemini reasoning

Gemini 3.5 Flash ​

  • Description: Agentic + Coding (GA)
  • Context: 1M tokens (1,000,000)
  • File support: Images, audio, video, PDFs, URLs, and text
  • Best for: Quick multimodal tasks, agentic workflows, and coding at large context
  • Use when: You want a stable, fast Gemini model at 1M context

Reasoning Effort ​

The Reasoning Effort control lets you tune how much the model "thinks" before responding, independently of which model you have chosen.

Where to Find It ​

The Effort lever sits next to the model selector in the chat composer. Click it to open a popover with a 5-notch slider ranging from Off on the left (Faster) to Max on the right (Smarter).

Effort Levels ​

LevelDescription
OffFastest -- minimal reasoning
LowQuick answers, light reasoning
MediumBalanced speed and depth (default)
HighDeeper reasoning, slower
MaxMost thorough -- highest latency

Your selection is saved automatically and applies to every new message until you change it. Click any notch label or drag the slider thumb to switch levels. The tooltip on the control reads: "Higher effort makes the model think longer -- smarter, more thorough answers, but slower. Lower effort responds faster."

When to Adjust Effort ​

  • Off or Low -- Quick factual lookups, short location searches, or simple follow-up questions where speed matters more than depth.
  • Medium (default) -- Everyday Zeus.ai tasks: analytics, scorecards, reports, Site Sonar.
  • High or Max -- Complex multi-step analysis, ambiguous or nuanced site comparisons, or when the model's first answer was less thorough than you expected.

How to Change Models ​

Step 1: Find the Model Selector ​

  1. Click in the chat input to expand the composer
  2. Look for the model dropdown at the top-left of the composer
  3. It shows the current model name (default: Zeus.ai Secure (Third-Party))

Step 2: Select a Model ​

  1. Click the dropdown to see all available models
  2. Each model shows its description and context limit
  3. Capability icons indicate support for images, PDFs, audio, video, and URLs
  4. A lock icon means the model is unavailable — hover for the reason
  5. Click your preferred model to select it

Step 3: Continue Chatting ​

  • Your next message will use the new model
  • Previous messages in the thread stay as they were
  • Model selection persists per conversation thread

Which Model Should I Use? ​

TaskRecommended Model
General use (default)Zeus.ai Secure (Third-Party)
Complex analytics/reasoningDeepSeek V4 Pro, Claude Opus 4.8, or GPT-5.5
Quick simple questionsClaude Haiku 4.5 or GPT-5.4 Mini
Searching locationsZeus.ai Secure, GPT-5.4, or Claude Sonnet 4.6
Analyzing imagesAny Claude, GPT, or Gemini model
Large document analysis (PDF)Claude Sonnet 4.6, GPT-5.5, or any Gemini model
Video / audio analysisGemini 3.1 Pro or Gemini 3.5 Flash
Maximum context neededZeus.ai Secure, Claude Sonnet 4.6, GPT-5.4, or any Gemini
URL / web link analysisGPT-5.4, any Claude, or any Gemini
Data residency (text-only)Zeus.ai Clean Room
Data residency + images/PDFsZeus.ai Clean Room (GPT-5)

Understanding Model Behavior ​

Thinking / Reasoning Indicator ​

DeepSeek V4 Pro and Claude models support visible "thinking":

  • You may see the model's reasoning process expand before the answer
  • The composer border pulses with electric blue while processing
  • The send button changes to a stop button during processing
  • You can click stop or press Escape to cancel

Context Windows Explained ​

The context window determines how much information the model can consider:

  • 200K tokens (~150K words): Claude Haiku 4.5
  • 272K tokens (~200K words): Zeus.ai Clean Room (GPT-5)
  • 1M tokens (~750K words): Zeus.ai Secure (Third-Party), Zeus.ai Clean Room, DeepSeek V4 Pro, Claude Sonnet 4.6, Claude Opus 4.8, GPT-5.5, GPT-5.4, GPT-5.4 Mini, Gemini 3.1 Pro, Gemini 3.5 Flash

Larger context = more conversation history and bigger documents in a single thread.


Model Capabilities Comparison ​

FeatureZeus.ai SecureDeepSeek V4 ProGPT-5.5 / 5.4GPT-5.4 MiniClaudeGemini
Location searchStrongStrongStrongStrongStrongStrong
Complex reasoningStrongStrongStrongGoodStrongGood
Image analysisNoNoYesYesYesYes
PDF / DocumentsNoNoYesYesYesYes
Audio / VideoNoNoNoNoNoYes
URL analysisNoNoYesYesYesYes
Context size1M1M1M1M200K–1M1M

Zeus.ai Clean Room capabilities (subscription required — both models run in the US Azure tenant):

FeatureClean RoomClean Room (GPT-5)
Image analysisNoYes
PDF / DocumentsNoYes
URL analysisNoYes
Audio / VideoNoNo
Context size1M272K

Tips ​

Default Company Access ​

By default every company has access to three models: the two Zeus.ai Clean Room models (subscription required to unlock) and Zeus.ai Secure (Third-Party). Other public models must be enabled by a company admin from Company Settings or by SiteZeus support.

Match Model to Your Content ​

  • Text and code: Zeus.ai Secure (default) works well
  • Images or PDFs: switch to any Claude, GPT, or Gemini model
  • Audio or video: use a Gemini model
  • Documents over 100 pages: use Claude Sonnet 4.6, GPT-5.5, or any Gemini
  • Data residency required: use a Zeus.ai Clean Room model

Consider the Trade-offs ​

  • Reasoning models (DeepSeek V4 Pro, Claude Opus 4.8) "think" before responding, so they take longer
  • Gemini 3.1 Pro sits on the preview channel; Gemini 3.5 Flash is the stable GA option
  • Larger context windows do not always mean better answers — pick the model that fits the task
  • Clean Room (GPT-5) has a hard 272K context ceiling; use the plain Clean Room for longer threads

SiteZeus Location Intelligence Platform — powered by Zeus.ai