---
title: "Pricing &amp; Usage — AVCodex Docs"
description: "Pricing &amp; Usage — AVCodex documentation for AV integrators, programmers, and ops teams."
lang: en
json-ld:
---

[](/)

Solutions

[Pricing](/pricing)[The Signal](/blog)[Resources](/resources)

Learn

[Free AI Assessment](/scorecard)[Get Started →](/pricing)

[Documentation Home](/docs)

Guides 

Getting Started

-   [The Alchemist Copilot](/docs/guides/the-alchemist-copilot)
-   [Choosing a Model](/docs/guides/choosing-a-model)
-   [Skills & Templates](/docs/guides/skills-and-templates)
-   [Pricing & Usage](/docs/guides/pricing-and-usage)
-   [Understanding Tokens](/docs/guides/understanding-tokens)
-   [Maximize AVCodex Capabilities](/docs/guides/maximize-avcodex-capabilities)

Knowledge & Memory

-   [How Knowledge Sources Work](/docs/guides/how-knowledge-sources-work)
-   [Knowledge Retrieval Settings](/docs/guides/knowledge-retrieval-settings)
-   [User Memory](/docs/guides/user-memory)
-   [Consumer Brain](/docs/guides/consumer-brain)

Agent Capabilities

-   [Image Recognition](/docs/guides/image-recognition)
-   [Image Generation](/docs/guides/image-generation)
-   [Video Generation](/docs/guides/video-generation)
-   [Deep Research and Deep Thinking](/docs/guides/deep-research-and-deep-thinking)
-   [Heartbeat (Proactive AI Outreach)](/docs/guides/heartbeat-proactive-ai-outreach)
-   [Database Connections](/docs/guides/database-connections)
-   [Agent-to-Agent Links](/docs/guides/agent-to-agent-links)
-   [Message Tagging](/docs/guides/message-tagging)
-   [Lead Generation Forms](/docs/guides/lead-generation-forms)
-   [Multilingual Apps](/docs/guides/multilingual-apps)
-   [Understanding Evaluations](/docs/guides/understanding-evaluations)

Design & Experience

-   [Style Studio](/docs/guides/style-studio)
-   [Component Studio](/docs/guides/component-studio)
-   [HQ Profile](/docs/guides/hq-profile)
-   [Multiplayer Chat](/docs/guides/multiplayer-chat)
-   [Circles](/docs/guides/circles)
-   [Desktop Agent](/docs/guides/desktop-agent)

Voice & Phone

-   [Phone Numbers](/docs/guides/phone-numbers)
-   [Outbound Calling](/docs/guides/outbound-calling)
-   [Voice Cloning](/docs/guides/voice-cloning)

Publish & Share

-   [Embed Chat Widget](/docs/guides/embed-chat-widget)
-   [Custom Domains](/docs/guides/custom-domains)
-   [PWA Installation](/docs/guides/pwa-installation)
-   [AVCodex Sites](/docs/guides/avcodex-sites)
-   [Embed on Kajabi](/docs/guides/embed-on-kajabi)
-   [How to Use AVCodex with Claude Code](/docs/guides/how-to-use-avcodex-with-claude-code)

Monetization & Access

-   [Selling Access](/docs/guides/selling-access)
-   [Consumer Monetization](/docs/guides/consumer-monetization)
-   [Access Control](/docs/guides/access-control)
-   [Bring Your Own Auth](/docs/guides/bring-your-own-auth)
-   [Clever SSO for Schools](/docs/guides/clever-sso-for-schools)

Analytics & Operations

-   [Analytics & Chat History](/docs/guides/analytics-and-chat-history)
-   [Performance Dashboard](/docs/guides/performance-dashboard)
-   [Programmatic Usage Stats](/docs/guides/programmatic-usage-stats)
-   [Session Lifecycle Webhooks](/docs/guides/session-lifecycle-webhooks)
-   [Audit Logs](/docs/guides/audit-logs)

Teams & White-Label

-   [Team Management](/docs/guides/team-management)
-   [Enterprise Whitelabel](/docs/guides/enterprise-whitelabel)

Alchemist Platform

-   [Alchemist Tickets](/docs/guides/alchemist-tickets)
-   [Alchemist Getting Started](/docs/guides/alchemist-getting-started)
-   [Alchemist Working with Tickets](/docs/guides/alchemist-working-with-tickets)
-   [Alchemist Local Development](/docs/guides/alchemist-local-development)

Alchemist Operations

-   [Alchemist Environment Variables](/docs/guides/alchemist-environment-variables)
-   [Alchemist Deploys and Domains](/docs/guides/alchemist-deploys-and-domains)
-   [Alchemist Self-Healing](/docs/guides/alchemist-self-healing)

Alchemist API & Automation

-   [Alchemist API Keys](/docs/guides/alchemist-api-keys)
-   [Alchemist MCP Server](/docs/guides/alchemist-mcp-server)
-   [Alchemist Pipeline Configuration](/docs/guides/alchemist-pipeline-configuration)
-   [Alchemist Pipeline Permutations](/docs/guides/alchemist-pipeline-permutations)

Developer Platform

-   [Building Custom MCP Servers](/docs/guides/building-custom-mcp-servers)
-   [Consumer OAuth for Custom MCP Servers](/docs/guides/consumer-oauth-for-custom-mcp-servers)

AVCodex MCP Server

-   [Overview](/docs/guides/overview)
-   [MCP Reference](/docs/guides/mcp-reference)
-   [Setup & Installation](/docs/guides/setup-and-installation)
-   [Authentication](/docs/guides/authentication)
-   [Tools Reference](/docs/guides/tools-reference)
-   [Common Workflows](/docs/guides/common-workflows)
-   [Rate Limits](/docs/guides/rate-limits)

Custom Actions 

Pro Actions 

API 

Builder API 

Agentic Commerce (ACP) 

Integrations 

[Docs](/docs)/ Guides / Getting Started 

# Pricing & Usage

Last updated · MAR 2026 · [Read as Markdown](/docs/guides/pricing-and-usage.md)

# Pricing and Usage

How AVCodex usage-based pricing works, and how to keep AI costs in check on integrator-scale traffic.

* * *

Understanding how AI pricing works helps you build better agents and control costs. This guide explains the pricing model and gives you tools to estimate and optimize your spend.

## [Why usage-based pricing# ](#why-usage-based-pricing)

We get asked about pricing a lot, so we will be straight about it.

**We want every AV team to have access to the best AI models available.** Not watered-down versions. Not artificially capped options. The actual best models from OpenAI, Anthropic, and Google.

These models have real costs. Running GPT-5 costs more than running GPT-5 Nano. Claude Opus 4 costs more than Claude 3.5 Haiku. If we charged a flat rate, we would have to do one of three things:

1.  **Restrict access to premium models**, forcing everyone onto cheaper ones.
2.  **Charge everyone the premium price**, putting AVCodex out of reach for many integrators.
3.  **Lose money**, which means we cannot keep building tools for the AV industry.

Usage-based pricing solves it. It means:

-   **Cost-conscious teams** can stand up high-volume room support agents on fast, cheap models like Gemini 2.5 Flash or GPT-5 Nano and keep monthly costs minimal.
-   **Quality-focused teams** can pull in the heaviest reasoning models (o3, Claude Opus 4) for RFP responses, design engineering, and deep troubleshooting when the work is worth it.
-   **Smart teams mix and match.** Cheap models for triage. Premium models for the hard problems.

This is not about being greedy. It is about running a business that can keep delivering best-in-class AI to every type of AV team, at every budget level.

## [How pricing works# ](#how-pricing-works)

### [Your plan includes usage# ](#your-plan-includes-usage)

Every AVCodex plan includes a monthly usage allowance:

Plan

Monthly price

Included usage

Builder

$29/mo

$10

Studio

$99/mo

$30

Studio Pro

$299/mo

$100

Stay within your allowance and you pay nothing extra. Need more? Buy additional credits in advance.

### [What counts as usage# ](#what-counts-as-usage)

Usage is measured in **tokens**, the units AI models use to process text. Roughly:

-   1 token is about 4 characters or three-quarters of a word.
-   A typical message uses 500 to 1,000 tokens total (input plus output).

Different models have different per-token costs. Premium models cost more, efficient models cost less.

Want to dig in? See the [Understanding Tokens](/docs/guides/understanding-tokens) guide with an interactive tokenizer.

## [Estimate your costs# ](#estimate-your-costs)

Use the calculator in the dashboard to estimate your monthly AI costs based on expected traffic.

## [Model pricing comparison# ](#model-pricing-comparison)

How different models stack up on price. Pick based on the work the agent is doing.

### [Most cost-effective models# ](#most-cost-effective-models)

These models give strong quality at the lowest prices.

### [Premium models# ](#premium-models)

When you need the best reasoning quality (think a senior design engineer agent reviewing a multi-room AV-over-IP build), these deliver.

## [Strategies to optimize cost# ](#strategies-to-optimize-cost)

### [1\. Match the model to the task# ](#1-match-the-model-to-the-task)

Do not run a senior reasoning model for FAQ traffic.

Task type

Recommended approach

Quick Q&A, room support

Gemini 2.5 Flash or GPT-5 Nano

General conversation, internal SOP help

GPT-5 Mini or Claude 3.5 Haiku

Content creation (project narratives, SOWs, drafts)

GPT-5 or Claude Sonnet 4

Complex analysis (design engineering, RFP scoring)

o3 or Claude Opus 4

Long documents (multi-room programming guides, AS-builts)

Gemini 2.5 Pro (best value for 1M context)

### [2\. Optimize your prompts# ](#2-optimize-your-prompts)

Shorter, clearer prompts use fewer tokens.

-   **Be specific** about what you want.
-   **Cut unneeded context** from system prompts.
-   **Use tight instructions** instead of verbose ones.

### [3\. Watch response length# ](#3-watch-response-length)

If your agent does not need long answers (a tech in a closet wants the fix in two sentences), tell the model to be concise. Output tokens cost more than input tokens.

### [4\. Use the right context window# ](#4-use-the-right-context-window)

Models with large context windows (Gemini 2.5 Pro at 1M tokens) shine on long documents. If your agent only handles short Q&A on touch panel issues, you do not need to pay for that capacity.

## [Real-world examples# ](#real-world-examples)

### [Example 1: Room support agent# ](#example-1-room-support-agent)

**Use case**: Answering touch panel and routine room questions through a QR code in every conference room.

**Recommended**: Gemini 2.5 Flash ($0.25/M avg).

**Monthly cost for 10,000 messages**: about $2.50.

### [Example 2: Programmer assistant# ](#example-2-programmer-assistant)

**Use case**: Helping Crestron and Q-SYS programmers draft modules, comments, and code reviews.

**Recommended**: GPT-5 ($16/M avg).

**Monthly cost for 5,000 messages**: about $40.

### [Example 3: RFP and spec analyzer# ](#example-3-rfp-and-spec-analyzer)

**Use case**: Reviewing 60-page RFPs, extracting requirements, summarizing scope.

**Recommended**: Gemini 2.5 Pro ($4/M avg, 1M context).

**Monthly cost for 1,000 documents**: about $20.

### [Example 4: Design engineering assistant# ](#example-4-design-engineering-assistant)

**Use case**: Deep multi-step reasoning on complex AV-over-IP designs, signal flow validation, codec calculations.

**Recommended**: o3 ($33/M avg).

**Monthly cost for 2,000 queries**: about $33.

## [Monitoring your usage# ](#monitoring-your-usage)

You can track usage in real time:

1.  Open **Settings** > **Billing** in your AVCodex dashboard.
2.  View your current credit balance and usage history.
3.  Buy additional credits when you need them.

We notify you when credits are running low, so you have time to top up before your agent pauses.

## [FAQ# ](#faq)

### [What happens if I run out of credits?# ](#what-happens-if-i-run-out-of-credits)

Your agent's AI functionality pauses until you purchase more credits. We send notifications as your balance drops, so you are not caught off guard. Top up anytime from billing settings.

### [How do I add more credits?# ](#how-do-i-add-more-credits)

Go to **Settings** > **Billing** and buy credits in advance. Credits are prepaid. No surprise bills, no overages. You control exactly what you spend.

### [Do unused credits roll over?# ](#do-unused-credits-roll-over)

No, included usage does not roll over month to month. This keeps pricing predictable.

### [Why do output tokens cost more than input tokens?# ](#why-do-output-tokens-cost-more-than-input-tokens)

That reflects how the underlying AI providers price their APIs. Generating new text (output) takes more compute than processing existing text (input).

### [Can I change models mid-conversation?# ](#can-i-change-models-mid-conversation)

An agent uses one model at a time, but you can change it anytime in Build settings. The change applies to new conversations.

## [What is coming# ](#what-is-coming)

Today, usage billing covers chat interactions and tool calls. We are expanding it to unlock more capability for every team:

### [Voice interactions# ](#voice-interactions)

Real-time voice conversations are expensive. Adding voice to usage billing lets us open the feature to everyone (not only specific tiers). You pay for what you use.

### [Knowledge source embeddings# ](#knowledge-source-embeddings)

Bulk uploads of knowledge sources are limited today. Adding embedding costs to usage billing lets us optimize for large-scale uploads. You will be able to load hundreds or thousands of documents (every Biamp datasheet, every Crestron module, every commissioning guide) into your agent's knowledge base.

### [Image and video generation# ](#image-and-video-generation)

AI-generated images and video are resource-intensive (video generation runs $3 to $6 per video). Usage billing lets us offer these tools to every plan. Generate marketing visuals, room walkthroughs, product demos. You only pay for what you create.

### [Billing and invoices# ](#billing-and-invoices)

The **Settings** > **Billing** page will show detailed usage invoices, so you can see exactly where the spend goes and track it over time.

We are building toward every AI capability being available to every team, with transparent, pay-for-what-you-use pricing.

## [Need help choosing# ](#need-help-choosing)

If you are not sure which model fits, see [Choosing a Model](/docs/guides/choosing-a-model) for recommendations by scenario.

## [Using with MCP# ](#using-with-mcp)

You can read billing and usage information through code using the AVCodex MCP Server.

### [Available tools# ](#available-tools)

Tool

Description

`get_credits`

View current credit balance.

`get_subscription`

Get subscription tier and details.

`get_usage`

Get a detailed usage breakdown.

### [Example prompts# ](#example-prompts)

code 

```
"How many credits do I have left?"
"What is my current subscription plan?"
"Show me my usage breakdown for this month"
```

### [Workflow: usage monitoring# ](#workflow-usage-monitoring)

code 

```
You: Show me my current billing status

AI: [Calls get_credits, get_subscription]

Subscription: Studio ($99/mo)
Credit Balance: $18.50 remaining
Monthly Allowance: $30 included

Usage this month:
- GPT-5: $8.20 (410K tokens)
- Claude Sonnet 4: $2.30 (115K tokens)
- Gemini Flash: $1.00 (4M tokens)
```

> **Note:** MCP server access requires a Builder plan or higher. See the [MCP Setup Guide](/docs/guides/mcp/setup) to get started.

\*AVCodex · Your AV expertise. Amplified by AI.\*

Was this helpful? 

[Edit this page →](#)

[

Previous

Skills & Templates

](/docs/guides/skills-and-templates)[

Next

Understanding Tokens

](/docs/guides/understanding-tokens)

On this page

-   [Why usage-based pricing](#why-usage-based-pricing)
-   [How pricing works](#how-pricing-works)
-   [Your plan includes usage](#your-plan-includes-usage)
-   [What counts as usage](#what-counts-as-usage)
-   [Estimate your costs](#estimate-your-costs)
-   [Model pricing comparison](#model-pricing-comparison)
-   [Most cost-effective models](#most-cost-effective-models)
-   [Premium models](#premium-models)
-   [Strategies to optimize cost](#strategies-to-optimize-cost)
-   [1\. Match the model to the task](#1-match-the-model-to-the-task)
-   [2\. Optimize your prompts](#2-optimize-your-prompts)
-   [3\. Watch response length](#3-watch-response-length)
-   [4\. Use the right context window](#4-use-the-right-context-window)
-   [Real-world examples](#real-world-examples)
-   [Example 1: Room support agent](#example-1-room-support-agent)
-   [Example 2: Programmer assistant](#example-2-programmer-assistant)
-   [Example 3: RFP and spec analyzer](#example-3-rfp-and-spec-analyzer)
-   [Example 4: Design engineering assistant](#example-4-design-engineering-assistant)
-   [Monitoring your usage](#monitoring-your-usage)
-   [FAQ](#faq)
-   [What happens if I run out of credits?](#what-happens-if-i-run-out-of-credits)
-   [How do I add more credits?](#how-do-i-add-more-credits)
-   [Do unused credits roll over?](#do-unused-credits-roll-over)
-   [Why do output tokens cost more than input tokens?](#why-do-output-tokens-cost-more-than-input-tokens)
-   [Can I change models mid-conversation?](#can-i-change-models-mid-conversation)
-   [What is coming](#what-is-coming)
-   [Voice interactions](#voice-interactions)
-   [Knowledge source embeddings](#knowledge-source-embeddings)
-   [Image and video generation](#image-and-video-generation)
-   [Billing and invoices](#billing-and-invoices)
-   [Need help choosing](#need-help-choosing)
-   [Using with MCP](#using-with-mcp)
-   [Available tools](#available-tools)
-   [Example prompts](#example-prompts)
-   [Workflow: usage monitoring](#workflow-usage-monitoring)

[](/)

The AI platform built exclusively for professional AV. Build, deploy, and sell AI tools that understand your industry.

### Platform

-   What You Can Build
-   Templates
-   [Pricing](/pricing)

### Services

-   [Done-For-You](/pricing)
-   [Academy](/academy)
-   [Contact](/contact)

### Company

-   About
-   [The Signal](/blog)
-   [Docs](/docs)
-   [LinkedIn](#)

© 2026 AVCodex. A Future Ready Holdings Inc. product. SOC 2 Type II Certified · HIPAA Compliant