---
title: "Choosing a Model — AVCodex Docs"
description: "Choosing a Model — AVCodex documentation for AV integrators, programmers, and ops teams."
lang: en
json-ld:
---

[](/)

Solutions

[Pricing](/pricing)[The Signal](/blog)[Resources](/resources)

Learn

[Free AI Assessment](/scorecard)[Get Started →](/pricing)

[Documentation Home](/docs)

Guides 

Getting Started

-   [The Alchemist Copilot](/docs/guides/the-alchemist-copilot)
-   [Choosing a Model](/docs/guides/choosing-a-model)
-   [Skills & Templates](/docs/guides/skills-and-templates)
-   [Pricing & Usage](/docs/guides/pricing-and-usage)
-   [Understanding Tokens](/docs/guides/understanding-tokens)
-   [Maximize AVCodex Capabilities](/docs/guides/maximize-avcodex-capabilities)

Knowledge & Memory

-   [How Knowledge Sources Work](/docs/guides/how-knowledge-sources-work)
-   [Knowledge Retrieval Settings](/docs/guides/knowledge-retrieval-settings)
-   [User Memory](/docs/guides/user-memory)
-   [Consumer Brain](/docs/guides/consumer-brain)

Agent Capabilities

-   [Image Recognition](/docs/guides/image-recognition)
-   [Image Generation](/docs/guides/image-generation)
-   [Video Generation](/docs/guides/video-generation)
-   [Deep Research and Deep Thinking](/docs/guides/deep-research-and-deep-thinking)
-   [Heartbeat (Proactive AI Outreach)](/docs/guides/heartbeat-proactive-ai-outreach)
-   [Database Connections](/docs/guides/database-connections)
-   [Agent-to-Agent Links](/docs/guides/agent-to-agent-links)
-   [Message Tagging](/docs/guides/message-tagging)
-   [Lead Generation Forms](/docs/guides/lead-generation-forms)
-   [Multilingual Apps](/docs/guides/multilingual-apps)
-   [Understanding Evaluations](/docs/guides/understanding-evaluations)

Design & Experience

-   [Style Studio](/docs/guides/style-studio)
-   [Component Studio](/docs/guides/component-studio)
-   [HQ Profile](/docs/guides/hq-profile)
-   [Multiplayer Chat](/docs/guides/multiplayer-chat)
-   [Circles](/docs/guides/circles)
-   [Desktop Agent](/docs/guides/desktop-agent)

Voice & Phone

-   [Phone Numbers](/docs/guides/phone-numbers)
-   [Outbound Calling](/docs/guides/outbound-calling)
-   [Voice Cloning](/docs/guides/voice-cloning)

Publish & Share

-   [Embed Chat Widget](/docs/guides/embed-chat-widget)
-   [Custom Domains](/docs/guides/custom-domains)
-   [PWA Installation](/docs/guides/pwa-installation)
-   [AVCodex Sites](/docs/guides/avcodex-sites)
-   [Embed on Kajabi](/docs/guides/embed-on-kajabi)
-   [How to Use AVCodex with Claude Code](/docs/guides/how-to-use-avcodex-with-claude-code)

Monetization & Access

-   [Selling Access](/docs/guides/selling-access)
-   [Consumer Monetization](/docs/guides/consumer-monetization)
-   [Access Control](/docs/guides/access-control)
-   [Bring Your Own Auth](/docs/guides/bring-your-own-auth)
-   [Clever SSO for Schools](/docs/guides/clever-sso-for-schools)

Analytics & Operations

-   [Analytics & Chat History](/docs/guides/analytics-and-chat-history)
-   [Performance Dashboard](/docs/guides/performance-dashboard)
-   [Programmatic Usage Stats](/docs/guides/programmatic-usage-stats)
-   [Session Lifecycle Webhooks](/docs/guides/session-lifecycle-webhooks)
-   [Audit Logs](/docs/guides/audit-logs)

Teams & White-Label

-   [Team Management](/docs/guides/team-management)
-   [Enterprise Whitelabel](/docs/guides/enterprise-whitelabel)

Alchemist Platform

-   [Alchemist Tickets](/docs/guides/alchemist-tickets)
-   [Alchemist Getting Started](/docs/guides/alchemist-getting-started)
-   [Alchemist Working with Tickets](/docs/guides/alchemist-working-with-tickets)
-   [Alchemist Local Development](/docs/guides/alchemist-local-development)

Alchemist Operations

-   [Alchemist Environment Variables](/docs/guides/alchemist-environment-variables)
-   [Alchemist Deploys and Domains](/docs/guides/alchemist-deploys-and-domains)
-   [Alchemist Self-Healing](/docs/guides/alchemist-self-healing)

Alchemist API & Automation

-   [Alchemist API Keys](/docs/guides/alchemist-api-keys)
-   [Alchemist MCP Server](/docs/guides/alchemist-mcp-server)
-   [Alchemist Pipeline Configuration](/docs/guides/alchemist-pipeline-configuration)
-   [Alchemist Pipeline Permutations](/docs/guides/alchemist-pipeline-permutations)

Developer Platform

-   [Building Custom MCP Servers](/docs/guides/building-custom-mcp-servers)
-   [Consumer OAuth for Custom MCP Servers](/docs/guides/consumer-oauth-for-custom-mcp-servers)

AVCodex MCP Server

-   [Overview](/docs/guides/overview)
-   [MCP Reference](/docs/guides/mcp-reference)
-   [Setup & Installation](/docs/guides/setup-and-installation)
-   [Authentication](/docs/guides/authentication)
-   [Tools Reference](/docs/guides/tools-reference)
-   [Common Workflows](/docs/guides/common-workflows)
-   [Rate Limits](/docs/guides/rate-limits)

Custom Actions 

Pro Actions 

API 

Builder API 

Agentic Commerce (ACP) 

Integrations 

[Docs](/docs)/ Guides / Getting Started 

# Choosing a Model

Last updated · MAR 2026 · [Read as Markdown](/docs/guides/choosing-a-model.md)

AVCodex supports multiple AI models from OpenAI, Anthropic, Google, and more. Each has different strengths, speeds, and costs. This guide helps you pick the right one for the job, whether that is a Crestron programming helper, a field tech triage agent, or an RFP assistant pulling spec from manufacturer PDFs.

## [Quick recommendations# ](#quick-recommendations)

Use case

Recommended model

General purpose (programmer help, internal SOP agent)

GPT-5.4 or Claude Sonnet 4.6

Image analysis (rack photos, error code screen captures, signal flow drawings)

Gemini 2.5 Pro, GPT-5.4, or Claude Sonnet 4.6

Complex reasoning (RFP responses, design engineering math, control system logic)

o3, Claude Opus 4.6, or GPT-5.4 Pro

Fast responses (QR-code room support, quick lookups)

Gemini 2.5 Flash, GPT-5.4 Mini, or Claude Haiku 4.5

Long documents (manufacturer manuals, AS-built packets, multi-room programming guides)

Gemini 2.5 Pro or GPT-4.1 (1M token context)

Cost-sensitive (high volume, simple FAQ-style agents)

Gemini 2.5 Flash Lite, Gemini 2.5 Flash, or GPT-5.4 Nano

## [Estimate your costs# ](#estimate-your-costs)

Use the calculator in the dashboard to estimate monthly AI costs based on your expected usage volume.

## [Featured models# ](#featured-models)

### [Best for general purpose# ](#best-for-general-purpose)

These models handle a wide spread of tasks: SOP writing, code assistance, conversational support, light analysis.

### [Best for speed# ](#best-for-speed)

When response time matters most (think a tech in a rack closet on a phone), these models return near-instant results without giving up much quality.

### [Best for reasoning# ](#best-for-reasoning)

For multi-step analysis, design problems, and tasks that need careful thinking through specs and constraints.

### [Best value for long documents# ](#best-value-for-long-documents)

Read entire programming guides, manufacturer manuals, or full RFP documents in one shot using massive context windows.

## [Model deep dives# ](#model-deep-dives)

### [OpenAI GPT-5.4# ](#openai-gpt-5-4)

### [Claude Sonnet 4.6# ](#claude-sonnet-4-6)

### [Gemini 2.5 Pro# ](#gemini-2-5-pro)

### [OpenAI o3# ](#openai-o3)

### [Claude Haiku 4.5# ](#claude-haiku-4-5)

### [Gemini 2.5 Flash# ](#gemini-2-5-flash)

## [Cost-effective options# ](#cost-effective-options)

If you are optimizing for cost (for example, a public-facing room support agent that fields hundreds of touch panel questions a day), these models give strong value.

## [Key considerations# ](#key-considerations)

### [Vision support# ](#vision-support)

If your agent analyzes images (a tech sending a phone photo of a Crestron error screen, a project manager uploading a rack drawing), pick a model with native vision support. Models without vision fall back to a workaround that can be less accurate.

**Vision-capable models:**

-   All GPT-4.1 and GPT-5.x variants (not the o-series reasoning models).
-   All Claude models.
-   All Gemini models.

**No vision support:**

-   OpenAI o-series (o1, o3, o4-mini, etc.).

### [Response speed# ](#response-speed)

Speed shapes the user experience. A field tech standing in a closet wants the answer in two seconds, not twenty.

-   **Fastest**: GPT-5.4 Nano, Claude Haiku 4.5, Gemini Flash Lite.
-   **Medium**: GPT-5.4, Claude Sonnet 4.6, Gemini 2.5 Flash.
-   **Slower**: Claude Opus 4.6, o1, o3 (reasoning takes time).

### [Context window# ](#context-window)

For processing long documents (full manufacturer manuals, multi-room AS-built packets), pick a model with a large context window.

-   **1M tokens**: GPT-4.1 variants, GPT-5.x variants, Claude Opus 4.6, Claude Sonnet 4.6, all Gemini models.
-   **200k tokens**: Claude Haiku 4.5, OpenAI o-series.

### [Reasoning quality# ](#reasoning-quality)

For complex tasks (RFP scoring, design engineering, multi-system signal flow analysis):

-   **Best reasoning**: o3 Pro, Claude Opus 4.6, GPT-5.4 Pro.
-   **Very good**: GPT-5.4, Claude Sonnet 4.6, o3, o4-mini.
-   **Good**: GPT-5.4 Mini, Claude Haiku 4.5, Gemini 2.5 Pro.

## [Changing your model# ](#changing-your-model)

1.  Open your agent in the AVCodex dashboard.
2.  Go to **Build** > **Configure**.
3.  Under **Model**, pick your preferred model.
4.  Click **Save**.

Changes take effect immediately for new conversations.

## [Testing different models# ](#testing-different-models)

Not sure which works best for your use case? A few approaches:

1.  **A/B testing**: Spin up two versions of your agent (one Sonnet, one GPT-5.4 for example) and compare feedback from your team.
2.  **Specific prompts**: Test your most common queries (a sample Crestron SIMPL question, a Q-SYS Lua snippet, a Biamp Tesira config question) against multiple models to compare quality.
3.  **Speed vs quality**: Start with a fast model. If your users need deeper answers, step up.

## [Using with MCP# ](#using-with-mcp)

You can manage AI models through code using the AVCodex MCP Server, including AI-powered model recommendations.

### [Available tools# ](#available-tools)

Tool

Description

`get_model`

Get the current model for an agent.

`update_model`

Change the AI model for an agent.

`list_available_models`

List all available AI models.

`recommend_model`

Get an AI-powered model recommendation.

### [Example prompts# ](#example-prompts)

code 

```
"What model is my Crestron programming agent using?"
"Switch my field tech support agent to GPT-5.4"
"Which model would you recommend for a high-volume room support agent?"
"List all available models with their pricing"
```

### [Workflow: AI-powered model selection# ](#workflow-ai-powered-model-selection)

code 

```
You: Recommend a model for my field tech support agent that needs to be fast
and handle simple Crestron error code lookups.

AI: [Calls recommend_model with priority="speed"]

Based on your use case, I recommend:

1. **Gemini 2.5 Flash** (Best choice)
   - Very fast responses (~1.5s)
   - Cost: $0.10/M input, $0.39/M output
   - Great for short lookups and FAQ-style traffic

2. **GPT-5.4 Nano** (Budget option)
   - Fastest response time
   - Cost: $0.26/M input, $1.63/M output
   - Good for high-volume, simple tasks

Want me to update your agent to use Gemini 2.5 Flash?

You: Yes, update it.

AI: [Calls update_model]
Updated Field Tech Support to use Gemini 2.5 Flash. Changes take effect
for new conversations.
```

> **Note:** MCP server access requires a Builder plan or higher. See the [MCP Setup Guide](/docs/guides/overview) to get started.

\*AVCodex · Your AV expertise. Amplified by AI.\*

Was this helpful? 

[Edit this page →](#)

[

Previous

The Alchemist Copilot

](/docs/guides/the-alchemist-copilot)[

Next

Skills & Templates

](/docs/guides/skills-and-templates)

On this page

-   [Quick recommendations](#quick-recommendations)
-   [Estimate your costs](#estimate-your-costs)
-   [Featured models](#featured-models)
-   [Best for general purpose](#best-for-general-purpose)
-   [Best for speed](#best-for-speed)
-   [Best for reasoning](#best-for-reasoning)
-   [Best value for long documents](#best-value-for-long-documents)
-   [Model deep dives](#model-deep-dives)
-   [OpenAI GPT-5.4](#openai-gpt-5-4)
-   [Claude Sonnet 4.6](#claude-sonnet-4-6)
-   [Gemini 2.5 Pro](#gemini-2-5-pro)
-   [OpenAI o3](#openai-o3)
-   [Claude Haiku 4.5](#claude-haiku-4-5)
-   [Gemini 2.5 Flash](#gemini-2-5-flash)
-   [Cost-effective options](#cost-effective-options)
-   [Key considerations](#key-considerations)
-   [Vision support](#vision-support)
-   [Response speed](#response-speed)
-   [Context window](#context-window)
-   [Reasoning quality](#reasoning-quality)
-   [Changing your model](#changing-your-model)
-   [Testing different models](#testing-different-models)
-   [Using with MCP](#using-with-mcp)
-   [Available tools](#available-tools)
-   [Example prompts](#example-prompts)
-   [Workflow: AI-powered model selection](#workflow-ai-powered-model-selection)

[](/)

The AI platform built exclusively for professional AV. Build, deploy, and sell AI tools that understand your industry.

### Platform

-   What You Can Build
-   Templates
-   [Pricing](/pricing)

### Services

-   [Done-For-You](/pricing)
-   [Academy](/academy)
-   [Contact](/contact)

### Company

-   About
-   [The Signal](/blog)
-   [Docs](/docs)
-   [LinkedIn](#)

© 2026 AVCodex. A Future Ready Holdings Inc. product. SOC 2 Type II Certified · HIPAA Compliant