Skip to content

BlogGuides

Which AI Model Should Your Discord Ticket Bot Use? Sonnet 5.5 and the Token Trade

Pick the model behind your Discord ticket bot: what a heavier model costs in tokens, when it pays off, where Claude Sonnet 5.5 fits, and how to switch safely.

Dani, Founder, AI Ticket Bot

8 min read

Most admins pick a model the way they pick a phone: the newest one, the biggest one, done. For a Discord ticket bot that is usually the wrong instinct, because the model is not free. Every reply spends tokens from your plan's allowance, and a heavier model spends more of them for the same reply.

So the real question is not "which model is best", it is "which model is worth what it costs on my server". This guide answers that, using what we changed today as the example.

Does the AI model matter for a Discord ticket bot?

Less than what the AI knows, more than nothing. A model that has never been told your refund rules will guess on any model. A model that has been taught them will answer a plain question correctly on the lightest model we offer.

Where the model starts to matter is judgment. A ticket that needs two rules weighed against each other, a lookup in a channel followed by a second lookup, or a member who describes the problem badly: those are the tickets where a stronger model gets it right more often. If your tickets are mostly "where do I find X" and "how do I do Y", the default model is enough and a heavier one spends tokens for nothing.

Stay on the default

  • Most tickets are one question with one answer
  • The AI answers correctly once it has been taught
  • Your plan's allowance runs close to the limit
  • You are on the free plan

Move to a heavier model

  • Tickets often need several rules weighed at once
  • The AI was taught the answer and still got it wrong
  • You end most months with tokens to spare
  • You are on a paid plan with room to spend

What does a heavier model cost?

Tokens, measured against the lightest model. A Discord ticket bot with an AI inside it pays the AI provider per token, and your plan's allowance is set in tokens. A model that costs twice as much per token uses your allowance twice as fast. We state it that way everywhere, as a rate, so you can compare models without knowing anyone's price sheet.

Our models on September 29, 2026

Claude Haiku 4.5
The default. Normal token rate
Claude Sonnet 5.5
New today. Twice the token rate
Claude Sonnet 5
Three times the token rate at the time of writing
Claude Sonnet 4.6
Three times the token rate

The picker lists them lightest first and shows the rate on each. Rates can change when the provider's prices change, so read them there rather than from this page.

What a rate means in tickets depends on your plan. On Premium, the monthly allowance covers about 230 average tickets on the default model. On Sonnet 5.5 that becomes about 115, and on a three-times model about 75. Pro's allowance is five times Premium's, so the same drop leaves far more room. The free plan's allowance is small enough that a heavier model can run out of tokens within weeks.

Where does Claude Sonnet 5.5 fit?

It is the new middle choice: stronger than the default, cheaper to run than the older Sonnet models. Anthropic released it on September 28 and we added it the next day, after testing it against the same calls the AI makes in a real ticket: looking something up, handing the result back, and finishing a reply.

Two things are worth knowing before you pick it:

  • It thinks before it answers. The picker marks it that way. On a tangled ticket that is the point. On a simple one it is time and tokens spent on a question that did not need it, and a reply can take a little longer to arrive.
  • It is priced below the Sonnet models that came before it. So if your server already runs on an older Sonnet model, Sonnet 5.5 uses your tokens more slowly than the model you have now. The confirm screen shows you the change in tickets a month before anything switches.

If you have been on the default and are happy with the answers, there is no reason to move because a new model exists. The question is always the one from the first section: can you name answers the default got wrong?

How do you tell whether you need a stronger model?

Read the tickets the AI handed over, not the ones it closed. A ticket the AI answered and the member closed is working. The signal is in the tickets where staff had to step in, and in why.

  • The AI had the rule and applied it to the wrong case

    A stronger model is likely to help

  • The AI did not know the rule at all

    Teach it. No model knows your server's rules untaught

  • The AI handed over a request only staff can approve

    That is the AI working as designed, on every model

Most wrong answers turn out to be the middle row: something nobody taught it. Fixing that is free on every plan, and teaching the AI from a closed ticket is usually the faster route to better answers than a heavier model. Switch models when the teaching is done and the AI still stumbles on tickets that need judgment.

Then check the cost side. /ai usage shows the model you are on, its token rate, and how many tokens this month has used. If you are already near the limit on the default, a heavier model will leave members without an AI answer at the end of the month, which is worse than a slightly weaker one all month.

How do you switch the AI model?

From /ai model in Discord or the Model section on your server's AI page in the dashboard. Both need Manage Server, and both show the same confirm screen.

  1. Open the model picker

    /ai model in Discord, or the AI page in the dashboard

  2. Pick a model

    Each one shows its token rate and about how many tickets a month your plan covers on it

  3. Read the confirm screen

    A heavier model shows your tickets a month before and after the switch

  4. Confirm

    The AI uses the new model from its next reply

  5. Watch a week of tickets

    Check /ai usage and the tickets staff took over, then keep it or switch back

Nothing the AI knows is touched by a switch. Its memory, your rules and your past tickets stay as they are, and switching back is the same few presses.

What does the model not change?

Everything around the reply. The model writes the answer. It does not decide what the AI is allowed to do, and it does not change what the AI knows.

What stays the same on every model

What the AI was taught
Memory and rules belong to your server, not the model
Its guardrails
It never invents a price or a policy, and it hands requests only staff can approve to your staff
When it hands over
The same handover rules on every model
Where it answers
Tickets, the public AI channel and the website widget all use the one model
Your data
Messages go through our AI platform and Anthropic's model API, as described in the privacy policy

That last row changes only if you leave our models entirely. On Pro, a server can connect its own AI provider and pick any compatible model its account offers; its replies then stop counting against the plan. That is a different trade, covered in its own guide.

Where a model switch is the wrong fix

The honest list.

Good fit

  • Tickets that need several rules weighed together
  • Servers on a paid plan that end months with tokens to spare
  • Moving off an older Sonnet model onto a cheaper, newer one

Not the right call

  • Answers that are wrong because the AI was never taught
  • Servers already close to their monthly allowance
  • The free plan, where a heavier model runs out fastest
  • Wanting a different model per category. It is one per server

If the AI runs out of tokens mid-month, it stops answering until the allowance resets, and tickets carry on working for your staff. A token pack tops it up without changing plan, and a lighter model stretches what is left.

Frequently asked questions

Start on the default, the lightest model, and change only when you can name answers it got wrong that a stronger model would get right. On AI Ticket Bot the default is Claude Haiku 4.5. Move up to Claude Sonnet 5.5 when your tickets need judgment across several rules or several lookups, and keep an eye on how many tickets a month your plan then covers.

Yes. Every model is measured against the lightest one. On AI Ticket Bot, Haiku 4.5 uses your tokens at the normal rate, Sonnet 5.5 uses them twice as fast, and the older Sonnet models three times as fast at the time of writing. The model picker shows the rate and roughly how many tickets a month your plan covers on each model before you switch.

No. The switch applies from the AI's next reply, and everything it was taught stays: its memory, your rules and your past tickets. You can switch back at any time, and switching back costs nothing either.

Yes. Every plan can pick any of our models, Free included. The free plan's token allowance is small, though, so a model that uses tokens twice as fast halves how many tickets the AI can answer in a month. On Free, the default model usually gets more done.

Not for every ticket. A newer, stronger model helps where a ticket needs judgment, such as several rules weighed together. On a plain question the AI was already taught, it gives the same answer and spends more tokens doing it. Change models for answers you can point to, not for the release date.

No. The model is one choice per server, and the tickets, the public AI channel and the website chat widget all use it. If you only want the AI in some categories, switch it off for the others instead.

It’s not just an AI, it’s your AI.

See it on your own server.

Add the bot free, teach it a few of your most common answers, and watch it clear the repeat tickets on its own.

Free plan, no card. Your first panel starts 14 days of Premium.