AI Discord moderation · trained per server

Everything says AI now. Ask it to explain itself.

Most "AI moderation" is a toxicity score with a threshold slider. Wardn reads the message, the picture, the link and the account behind it, writes down why it acted — and when your mods overrule it, that becomes a rule you never had to write.

hey! saw you in the tournament — I'm hiring EU players, dm me for details 💰

  1. 1

    No banned words, no links, nothing for a filter to match.

  2. 2

    Identical message sent to 22 members within 4 minutes.

  3. 3

    Account is 3 days old with no channel activity.

  4. 4

    Your mods have banned this exact pattern four times before.

Removed + banned

Mass DM recruitment from a throwaway. Matches a pattern your team already approved.

The honest version

Three things people call AI moderation.

Only one of them survives contact with a mod team that has to justify a ban to a member.

A toxicity score

One number between 0 and 1, a slider, and a shrug. Set it low and you nuke sarcasm. Set it high and it catches nothing. Nobody can explain any individual decision.

A general-purpose model with a prompt

Better at language, but it doesn't know your server. It applies the average internet's standards to a community with its own in-jokes, and it forgets every correction you make.

A reader trained on your history

Starts from your bans, warnings and the messages your mods deliberately left alone. Keeps a written reason for every action. Gets corrected once, not repeatedly.

The pattern-training loop

Your mods' corrections are the configuration.

Nothing here asks you to write a rule. You moderate the way you already moderate; Wardn turns the disagreements into policy — and never applies one until a human signs it off.

12345patterntraining loop
Step 1 of 5

Wardn makes a call

Removes a message, resets a name, or holds something for review — with a written reason.

Guardrails

What it will not do.

An AI moderator that can't be stopped is a liability, not a feature.

No silent actions

Every removal, mute and ban lands in one feed with the reason attached and the mod who can undo it.

No permanent mistakes

Undo restores the message, the roles and the member — and files the correction as training.

No unapproved patterns

A learned pattern is a proposal until your team accepts it. You can reject one and it stops suggesting it.

No acting on a coin flip

Below its confidence bar it escalates to a human instead of guessing. Borderline cases are a queue, not a ban.

The same reading applies to scam links, images and slurs and self-promotion — the judgment is the same, the context is what changes.

Fair questions

What people push back on.

Are you training on our server's data for other customers?

No. Each server's tuning stays with that server. What you teach it is yours; nobody else inherits your community's standards.

What happens the first week, before it's learned anything?

It acts on nothing. It reads, and produces a report of what it would have done. You review that in a sandbox before it gets a single permission.

Can it be wrong in a way we don't notice?

Everything it touched is in one feed, filterable, with reasons. The most common first-week activity for a new team is scrolling that feed and pressing undo twice.

Our server has weird norms. Will it fight us?

That's the point of training per server. Communities where roasting each other is affection get a very different Wardn than a professional support Discord.

Get on the list

Bring us your hardest calls.

Tell us the decisions your team argues about. Those are the ones we tune against first — and they tell us more than your member count does.

We onboard 5 servers per batch · quote is built per server

Name or invite link. Add as many as you like.

1Mods are stretched5

No install, no permissions, no card. We'll quote your setup fee and ongoing fee on the call, once we know your volume and how you want it set up.