> ## Documentation Index
> Fetch the complete documentation index at: https://help.wonka.chat/llms.txt
> Use this file to discover all available pages before exploring further.

# Token Pricing

> Understand AI model token costs and choose the right model for your use case

## What Are Tokens?

<Info>
  Tokens are the units that measure AI model usage. Both your input (prompt) and the AI's response (completion) consume tokens. Think of tokens as roughly 3/4 of a word in English.

  Example: "Hello, how are you today?" ≈ 6 tokens
</Info>

All token prices are listed in USD per 1 million tokens.

***

## Available Models in WonkaChat

<Info>
  The following models are currently available in WonkaChat. For specific pricing information, please contact our Sales Support team.
</Info>

<Tabs>
  <Tab title="By Provider">
    ### OpenAI Models

    <AccordionGroup>
      <Accordion title="GPT-5 Series (Latest)" icon="star">
        | Model       | Best For                                                   |
        | ----------- | ---------------------------------------------------------- |
        | **GPT-5.2** | Most advanced general-purpose AI, cutting-edge performance |
        | **GPT-5**   | High-quality responses, complex reasoning                  |

        <Check>
          GPT-5 series represents the latest in AI capabilities with enhanced reasoning and generation quality.
        </Check>
      </Accordion>

      <Accordion title="GPT-4o Series" icon="circle-check">
        | Model           | Best For                                     |
        | --------------- | -------------------------------------------- |
        | **GPT-4o Mini** | General tasks, best value for most use cases |
        | **GPT-4o**      | High-quality responses, complex reasoning    |

        <Tip>
          GPT-4o Mini is our most popular choice - excellent balance of quality and cost-effectiveness.
        </Tip>
      </Accordion>
    </AccordionGroup>

    ***

    ### Anthropic Claude Models

    <AccordionGroup>
      <Accordion title="Claude Sonnet Series" icon="brain">
        | Model                 | Best For                                                       |
        | --------------------- | -------------------------------------------------------------- |
        | **Claude Sonnet 4.6** | Latest version, enhanced performance and instruction following |
        | **Claude Sonnet 4.5** | Balanced performance for complex tasks                         |
        | **Claude Sonnet 4**   | Complex reasoning, detailed analysis                           |

        <Check>
          Claude Sonnet models are renowned for following complex, nuanced instructions accurately.
        </Check>
      </Accordion>

      <Accordion title="Claude Specialized Models" icon="sparkles">
        | Model                | Best For                                            |
        | -------------------- | --------------------------------------------------- |
        | **Claude Haiku 4.5** | Fast, efficient tasks with quick turnaround         |
        | **Claude Opus 4.5**  | Maximum capability, most complex and critical tasks |

        <Info>
          Claude Haiku is optimized for speed, while Opus provides the highest quality for critical work.
        </Info>
      </Accordion>
    </AccordionGroup>

    ***

    ### Google Gemini Models

    <AccordionGroup>
      <Accordion title="Gemini 3 Series (Preview)" icon="rocket">
        | Model                      | Best For                                    |
        | -------------------------- | ------------------------------------------- |
        | **Gemini 3 Pro Preview**   | Next-generation complex reasoning (preview) |
        | **Gemini 3 Flash Preview** | Next-generation fast processing (preview)   |

        <Warning>
          Gemini 3 models are currently in preview. Features and availability may change.
        </Warning>
      </Accordion>

      <Accordion title="Gemini 2.5 Series" icon="google">
        | Model                | Best For                            |
        | -------------------- | ----------------------------------- |
        | **Gemini 2.5 Pro**   | Complex reasoning, analytical tasks |
        | **Gemini 2.5 Flash** | Fast, cost-effective general tasks  |

        <Tip>
          Gemini 2.5 Flash offers excellent performance for high-volume operations at competitive pricing.
        </Tip>
      </Accordion>
    </AccordionGroup>

    ***

    ### Mistral AI Models

    | Model                     | Best For                                      |
    | ------------------------- | --------------------------------------------- |
    | **Mistral Large Latest**  | Advanced tasks requiring high capability      |
    | **Mistral Medium Latest** | Balanced performance for general business use |

    <Info>
      Mistral models are automatically updated to the latest versions, ensuring you always have access to improvements.
    </Info>
  </Tab>

  <Tab title="By Use Case">
    ## Choose by Task Type

    <Steps>
      <Step title="Simple & High-Volume Tasks">
        **Best for:** FAQs, categorization, basic summaries, simple classifications

        **Recommended Models:**

        * GPT-4o Mini - Best overall value
        * Gemini 2.5 Flash - Very cost-effective
        * Claude Haiku 4.5 - Fast processing

        **Use when:** You need quick responses for straightforward questions or tasks that don't require deep reasoning.
      </Step>

      <Step title="General Business Tasks">
        **Best for:** Emails, reports, research, content generation, customer service

        **Recommended Models:**

        * GPT-4o - High-quality, reliable
        * Claude Sonnet 4.6 - Excellent instruction following
        * Gemini 2.5 Pro - Strong analytical performance
        * Mistral Large Latest - Advanced general capabilities

        **Use when:** You need quality responses for typical business communications and analysis.
      </Step>

      <Step title="Complex & Critical Tasks">
        **Best for:** Strategic decisions, complex analysis, sensitive communications, detailed research

        **Recommended Models:**

        * GPT-5.2 - Most advanced capabilities
        * Claude Opus 4.5 - Maximum quality and accuracy
        * Gemini 3 Pro Preview - Next-gen reasoning
        * Claude Sonnet 4.6 - Complex instruction following

        **Use when:** Quality and accuracy are paramount, and the task requires sophisticated reasoning.
      </Step>

      <Step title="Specialized Requirements">
        **Speed Priority:**

        * Claude Haiku 4.5
        * Gemini 2.5 Flash
        * GPT-4o Mini

        **Instruction Following:**

        * Claude Sonnet 4.6
        * Claude Opus 4.5

        **Latest Technology:**

        * GPT-5.2
        * Gemini 3 Pro Preview
        * Claude Sonnet 4.6
      </Step>
    </Steps>
  </Tab>

  <Tab title="Model Comparison">
    ## Quick Comparison Guide

    ### Entry Level (Economy)

    **When to use:** High volume, simple tasks, internal tools

    | Model            | Strengths                               |
    | ---------------- | --------------------------------------- |
    | GPT-4o Mini      | Best all-around value, reliable quality |
    | Gemini 2.5 Flash | Very cost-effective, fast               |
    | Claude Haiku 4.5 | Quick turnaround, efficient             |

    ***

    ### Mid-Tier (Standard)

    **When to use:** General business tasks, customer-facing content

    | Model                | Strengths                       |
    | -------------------- | ------------------------------- |
    | GPT-4o               | High quality, complex reasoning |
    | Claude Sonnet 4.6    | Excellent instruction following |
    | Gemini 2.5 Pro       | Strong analytical capabilities  |
    | Mistral Large Latest | Advanced general performance    |

    ***

    ### Premium (Advanced)

    **When to use:** Critical decisions, complex analysis, maximum quality

    | Model                | Strengths                    |
    | -------------------- | ---------------------------- |
    | GPT-5.2              | Cutting-edge AI capabilities |
    | Claude Opus 4.5      | Maximum accuracy and quality |
    | Gemini 3 Pro Preview | Next-generation reasoning    |

    ***

    ### By Workload Type

    **High-Volume Operations:**
    → GPT-4o Mini, Gemini 2.5 Flash

    **Customer-Facing:**
    → GPT-4o, Claude Sonnet 4.6, Mistral Large Latest

    **Internal Analysis:**
    → Gemini 2.5 Pro, Claude Sonnet 4.5, GPT-4o

    **Critical Tasks:**
    → GPT-5.2, Claude Opus 4.5, Claude Sonnet 4.6
  </Tab>
</Tabs>

***

## Frequently Asked Questions

<AccordionGroup>
  <Accordion title="How are tokens counted?" icon="calculator">
    Tokens are counted for both input (your prompt + agent instructions) and output (the AI's response).

    Rough estimate: 1 token ≈ 0.75 English words

    Example conversation:

    * Your question: "Summarize this document" (3 tokens)
    * Document content: 2,000 words (≈2,666 tokens)
    * AI summary: 200 words (≈267 tokens)
    * Total: \~2,936 tokens consumed
  </Accordion>

  <Accordion title="Which model should I use?" icon="circle-question">
    Start here:

    * Most teams: GPT-4o Mini - best value
    * Speed-focused: Gemini 2.5 Flash or Claude Haiku 4.5
    * Complex tasks: Claude Sonnet 4.6 or GPT-4o
    * Maximum capability: GPT-5.2 or Claude Opus 4.5

    Test with your actual use case to find the sweet spot.
  </Accordion>

  <Accordion title="Can I switch models?" icon="arrows-rotate">
    Yes! You can configure different models for different agents. Use expensive models only where quality matters most.

    Strategy:

    * Customer-facing: Premium or standard models
    * Internal tools: Economy models
    * Testing: Economy models
  </Accordion>

  <Accordion title="How is pricing calculated?" icon="coins">
    Token pricing is based on:

    1. Input tokens: Your prompt + system instructions + context
    2. Output tokens: The AI's generated response

    Different models have different pricing for input vs output. For specific pricing details, please contact our Sales Support team.
  </Accordion>

  <Accordion title="What's the difference between model versions?" icon="code-branch">
    Newer versions (like Claude Sonnet 4.6 vs 4.5, or GPT-5.2 vs GPT-5) generally offer:

    * Improved reasoning capabilities
    * Better instruction following
    * Enhanced accuracy
    * Sometimes better pricing

    The "latest" variants (like Mistral Large Latest) automatically update to the newest version.
  </Accordion>

  <Accordion title="Are preview models stable for production?" icon="flask">
    Preview models (like Gemini 3 series) are:

    * Cutting-edge but may have changes
    * Best for testing new capabilities
    * Not recommended for critical production workloads

    For production use, stick with stable releases like GPT-4o, Claude Sonnet 4.6, or Gemini 2.5 series.
  </Accordion>
</AccordionGroup>

***

## Need Help Choosing?

<Card title="Contact Our Team" icon="comments">
  Not sure which models fit your use case? Our Sales Support team can:

  * Analyze your requirements
  * Recommend the optimal model mix
  * Provide detailed pricing for your expected volume
  * Help you test different options

  [Get Model Recommendations](/en/support-resources/getting-help)
</Card>
