Skip to content

AI models we support

Every model available on Kanbu.ai, grouped by how we recommend using it — with price and speed where we have data. Your organization's actual pricing is shown in the app.

Recommended

Best balance of price, performance and latency — our default pick for most support and shopping assistants.

  • Gemini 3.8 Flash

    Our recommendation
    Provider: Google

    A versatile assistant for helpful, accurate customer conversations.

    Connects information clearly, understands detailed product requirements and handles follow-up questions naturally. A great everyday choice for support and shopping advice.

    $0.16
    Average per answer
    7.0 s · Typical complete answer
    Typical range: $0.0296 – $0.36
  • GPT-5.6 Luna

    Provider: OpenAI

    Smart, economical support with clear answers.

    Combines approachable explanations with practical reasoning about prices and product choices. Well suited to everyday support; clear instructions help keep more complex conversations focused.

    $0.0476
    Average per answer
    6.0 s · Typical complete answer
    Typical range: $0.0077 – $0.14
  • Gemini 3.5 Flash-Lite

    Provider: Google

    Quick, concise answers for busy customer support.

    Keeps conversations moving with short response times and direct answers. Ideal for FAQs and routine requests; for intricate calculations or many simultaneous requirements, consider Gemini 3.8 Flash.

    $0.0769
    Average per answer
    3.7 s · Typical complete answer
    Typical range: $0.0236 – $0.23
  • Hunyuan HY3

    Provider: Tencent

    Thoughtful answers at an economical price.

    Good at following detailed requirements and explaining product choices without a premium price. A useful choice when answer quality matters more than immediate response speed.

    $0.0434
    Average per answer
    17.0 s · Typical complete answer
    Typical range: $0.0125 – $0.0935

Advanced

More depth for advanced, higher-stakes conversations, at a higher cost per answer.

  • GPT-5.6 Terra

    Provider: OpenAI

    More depth for demanding conversations.

    A premium option for nuanced instructions and detailed explanations. Useful when you want more thorough assistance and can allow a higher cost per answer.

    $0.45
    Average per answer
    5.0 s · Typical complete answer
    Typical range: $0.15 – $1.25
  • Claude Sonnet 5

    Provider: Anthropic

    Clear, detailed communication for complex requests.

    A premium alternative for assistants with extensive instructions and a need for thoughtful explanations. Best suited to conversations where detail is worth the additional cost.

    $0.66
    Average per answer
    13.7 s · Typical complete answer
    Typical range: $0.24 – $2.09
  • Grok 4.6

    Provider: xAI

    Premium assistance for challenging questions.

    Handles demanding questions with detailed, direct answers. A strong premium option when depth matters; longer searches and complex requests can cost considerably more.

    $0.66
    Average per answer
    28.6 s · Typical complete answer
    Typical range: $0.23 – $1.77
  • Muse Spark 1.3

    Provider: Meta

    Articulate support with a natural conversational style.

    Offers polished explanations and follows the direction of a conversation well. Suited to more involved customer interactions where a little extra time and cost are acceptable.

    $0.36
    Average per answer
    17.3 s · Typical complete answer
    Typical range: $0.0915 – $0.91
  • MiniMax M3

    Provider: MiniMax

    An affordable alternative for everyday assistance.

    Provides practical answers at a modest cost. Best for clearly defined support tasks; unusually complex or sensitive requests may need a more capable alternative.

    $0.0594
    Average per answer
    8.0 s · Typical complete answer
    Typical range: $0.0151 – $0.13

Ultra

The most capable, high-end models available. Usually overkill for a regular support assistant — reach for these only for very difficult workflows, advanced logic and tools, or complex internal use cases.

  • GPT-5.6 Sol

    Provider: OpenAI

    OpenAI's most capable model, for the hardest reasoning and tool-use tasks.

    Handles multi-step reasoning, complex tool orchestration and nuanced logic that other models struggle with. Considerable overkill for everyday support — reserve it for demanding internal workflows or advanced automation.

    ~$0.54
    Estimated — not directly measured for this model
  • Gemini 3.1 Pro

    Provider: Google

    Google's advanced reasoning model for complex, multi-step tasks.

    Strong at synthesizing large amounts of context and following intricate instructions. More capability than most support chatbots need — a good fit for sophisticated internal tools or elaborate product logic.

    ~$0.28
    Estimated — not directly measured for this model
  • Claude Opus 5

    Provider: Anthropic

    Anthropic's most capable model, for the most demanding reasoning and agentic tasks.

    Best suited to complex multi-step workflows, advanced tool use and nuanced judgment calls. Significant overkill for routine support conversations, but a strong fit for sophisticated internal use cases.

    ~$0.68
    Estimated — not directly measured for this model

Experimental

  • DeepSeek V4.1 Flash

    Provider: DeepSeek

    Detailed reasoning with an economical token price.

    A promising choice for thoughtful, detailed answers when speed is less important. Complex searches can take substantially longer or time out, so choose another model for time-sensitive support.

    $0.13
    Average per answer
    31.1 s · Typical complete answer
    Typical range: $0.0232 – $0.54
  • StepFun 3.7 Flash

    Provider: StepFun

    An economical model for exploring new assistant workflows.

    Useful for experimenting with everyday conversational tasks. More demanding instructions and tool-heavy requests can need closer supervision; use a recommended model for critical customer workflows.

    $0.0554
    Average per answer
    10.9 s · Typical complete answer
    Typical range: $0.0196 – $0.16

Other options

Older and less common models, kept available for existing assistants and specific workflows.

Missing a model? Write to us — we can add it.

  • gpt-5.5

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.69
    Estimated — not directly measured for this model
  • gpt-5.4

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.35
    Estimated — not directly measured for this model
  • GPT-5.4 Mini

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.11
    Estimated — not directly measured for this model
  • GPT-5.4 Nano

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0306
    Estimated — not directly measured for this model
  • gpt-4.1

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.27
    Estimated — not directly measured for this model
  • gpt-4.1-mini

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0558
    Estimated — not directly measured for this model
  • GPT-4.1 Nano

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0162
    Estimated — not directly measured for this model
  • gpt-5

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.18
    Estimated — not directly measured for this model
  • gpt-5-mini

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0386
    Estimated — not directly measured for this model
  • gpt-5-nano

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0101
    Estimated — not directly measured for this model
  • o3-mini-2025-01-31

    Provider: OpenAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.15
    Estimated — not directly measured for this model
  • x-ai/grok-4.5

    Provider: xAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.26
    Estimated — not directly measured for this model
  • x-ai/grok-4-fast

    Provider: xAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0286
    Estimated — not directly measured for this model
  • x-ai/grok-4.1-fast

    Provider: xAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0286
    Estimated — not directly measured for this model
  • x-ai/grok-4

    Provider: xAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.41
    Estimated — not directly measured for this model
  • x-ai/grok-4.20

    Provider: xAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.16
    Estimated — not directly measured for this model
  • x-ai/grok-4.3

    Provider: xAI

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.16
    Estimated — not directly measured for this model
  • google/gemini-2.0-flash-001

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0228
    Estimated — not directly measured for this model
  • google/gemini-2.0-flash-lite-001

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0129
    Estimated — not directly measured for this model
  • google/gemini-2.5-flash

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.046
    Estimated — not directly measured for this model
  • Gemini 2.5 Flash-Lite

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0162
    Estimated — not directly measured for this model
  • google/gemini-2.5-pro

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.18
    Estimated — not directly measured for this model
  • google/gemini-3-flash-preview

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0716
    Estimated — not directly measured for this model
  • Gemini 3.1 Flash-Lite

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0373
    Estimated — not directly measured for this model
  • google/gemini-3.1-flash-lite-preview

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0373
    Estimated — not directly measured for this model
  • google/gemini-3.7-flash

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    $0.16
    Average per answer
    8.8 s · Typical complete answer
    Typical range: $0.0567 – $0.38
  • google/gemini-3.6-flash

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.20
    Estimated — not directly measured for this model
  • google/gemini-3.5-flash

    Provider: Google

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.21
    Estimated — not directly measured for this model
  • claude-3-5-sonnet-latest

    Provider: Anthropic

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.41
    Estimated — not directly measured for this model
  • anthropic/claude-sonnet-4.6

    Provider: Anthropic

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.41
    Estimated — not directly measured for this model
  • anthropic/claude-opus-4.7

    Provider: Anthropic

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.68
    Estimated — not directly measured for this model
  • anthropic/claude-opus-4.8

    Provider: Anthropic

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.68
    Estimated — not directly measured for this model
  • anthropic/claude-opus-4.8-fast

    Provider: Anthropic

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$1.35
    Estimated — not directly measured for this model
  • anthropic/claude-haiku-4.5

    Provider: Anthropic

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    $0.17
    Average per answer
    7.6 s · Typical complete answer
    Typical range: $0.0839 – $0.20
  • z-ai/glm-5.2

    Provider: Z.ai

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.18
    Estimated — not directly measured for this model
  • deepseek/deepseek-v4-pro

    Provider: DeepSeek

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.22
    Estimated — not directly measured for this model
  • deepseek/deepseek-v4-flash

    Provider: DeepSeek

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    $0.0561
    Average per answer
    22.9 s · Typical complete answer
    Typical range: $0.0138 – $0.12
  • xiaomi/mimo-v2.5

    Provider: Xiaomi

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0165
    Estimated — not directly measured for this model
  • xiaomi/mimo-v2.5-pro

    Provider: Xiaomi

    Legacy model retained for existing configurations.

    Available for existing assistants and familiar workflows.

    ~$0.0582
    Estimated — not directly measured for this model

Approximate usage prices include our standard service margin and conversion into the selected currency. Your organization’s prices are shown in the app. Costs and response times vary with context, tools and answer length. Estimated prices for less common models are derived from typical usage patterns and the model's own published rate, not measured directly.