What Are Tokens?
Tokens are the units that measure AI model usage. Both your input (prompt) and the AIâs response (completion) consume tokens. Think of tokens as roughly 3/4 of a word in English.Example: âHello, how are you today?â â 6 tokens
Available Models in WonkaChat
The following models are currently available in WonkaChat. For specific pricing information, please contact our Sales Support team.
- By Provider
- By Use Case
- Model Comparison
OpenAI Models
GPT-5 Series (Latest)
GPT-5 Series (Latest)
GPT-5 series represents the latest in AI capabilities with enhanced reasoning and generation quality.
GPT-4o Series
GPT-4o Series
Anthropic Claude Models
Claude Sonnet Series
Claude Sonnet Series
Claude Sonnet models are renowned for following complex, nuanced instructions accurately.
Claude Specialized Models
Claude Specialized Models
Claude Haiku is optimized for speed, while Opus provides the highest quality for critical work.
Google Gemini Models
Gemini 3 Series (Preview)
Gemini 3 Series (Preview)
Gemini 2.5 Series
Gemini 2.5 Series
Mistral AI Models
Mistral models are automatically updated to the latest versions, ensuring you always have access to improvements.
Frequently Asked Questions
How are tokens counted?
How are tokens counted?
Tokens are counted for both input (your prompt + agent instructions) and output (the AIâs response).Rough estimate: 1 token â 0.75 English wordsExample conversation:
- Your question: âSummarize this documentâ (3 tokens)
- Document content: 2,000 words (â2,666 tokens)
- AI summary: 200 words (â267 tokens)
- Total: ~2,936 tokens consumed
Which model should I use?
Which model should I use?
Start here:
- Most teams: GPT-4o Mini - best value
- Speed-focused: Gemini 2.5 Flash or Claude Haiku 4.5
- Complex tasks: Claude Sonnet 4.6 or GPT-4o
- Maximum capability: GPT-5.2 or Claude Opus 4.5
Can I switch models?
Can I switch models?
Yes! You can configure different models for different agents. Use expensive models only where quality matters most.Strategy:
- Customer-facing: Premium or standard models
- Internal tools: Economy models
- Testing: Economy models
How is pricing calculated?
How is pricing calculated?
Token pricing is based on:
- Input tokens: Your prompt + system instructions + context
- Output tokens: The AIâs generated response
What's the difference between model versions?
What's the difference between model versions?
Newer versions (like Claude Sonnet 4.6 vs 4.5, or GPT-5.2 vs GPT-5) generally offer:
- Improved reasoning capabilities
- Better instruction following
- Enhanced accuracy
- Sometimes better pricing
Are preview models stable for production?
Are preview models stable for production?
Preview models (like Gemini 3 series) are:
- Cutting-edge but may have changes
- Best for testing new capabilities
- Not recommended for critical production workloads
Need Help Choosing?
Contact Our Team
Not sure which models fit your use case? Our Sales Support team can:
- Analyze your requirements
- Recommend the optimal model mix
- Provide detailed pricing for your expected volume
- Help you test different options
