One AI assistant — as a widget, in your apps, and via APIStart your free trial
Kyros
Chat widget
Embeddable AI chat for any website
Video avatar bot
Real-time talking avatar
Knowledge base
Grounded answers with sources
Interactive cards
In-chat forms & actions
Compare models
Compare models side by side
Moderation
PII & content guardrails
Customer support
Deflect repetitive tickets
Sales & lead gen
Qualify visitors & capture leads
Internal knowledge
One place for team answers
E-commerce
Insurance
Law firms
Healthcare
Blog
Guides & best practices
Glossary
AI terms explained
Templates
Ready-made cards & assistants
Case studies
Results from real teams
Documentation
Embed, API & more
Integrations
Connect your tools
Pricing
Sign inStart free
  1. Home
  2. Product
  3. Moderation

Content moderation

Screen every message before it reaches the model.

Moderation screens incoming messages and withholds anything flagged — before the model and before storage — with a polite, on-brand refusal.

Start for freeSee how it works

What moderation screens

Protection that acts before it’s too late.

Six categories on by default

PII, sexual content, hate & discrimination, violence & threats, dangerous & criminal content and self-harm are on by default.

Opt-in categories

Health, financial and legal advice, plus jailbreak attempts, can be enabled on top.

Withheld before model & storage

Flagged messages are withheld before they reach the model or get stored.

Polite, on-brand refusal

Instead of a hard error, the user gets a polite refusal that matches your brand.

Calibrated verdict

The verdict is calibrated (provider: Mistral) — with an optional custom sensitivity threshold.

Before processing

The filter acts on the way in — not just in the reply — protecting the model, storage and your brand.

How to set up moderation.

  • The six default categories are on out of the box — nothing to do.
  • Choose opt-in categories: health, financial, legal advice and jailbreak attempts.
  • Optionally set a custom sensitivity threshold for the calibrated verdict.
  • Tune the refusal text to your brand and test live — flagged content is withheld upfront.

Frequently asked questions

Six categories are on out of the box: PII, sexual content, hate & discrimination, violence & threats, dangerous & criminal content and self-harm.

Related

Security

EU hosting, encryption and audit log.

GDPR

DPA, TOMs and data processing.

Chat widget

Moderation applies inside every embeddable chat.

Pricing

Transparent credits, Starter free forever.

Build an assistant you can trust live.

14-day free trial. No credit card. German & English.

Start for freeSee how it works
Kyros

Build AI assistants grounded in your content and deploy them anywhere — widget, apps, API. From the DACH region.

Product
  • Chat widget
  • Video avatar bot
  • Knowledge base
  • Interactive cards
  • Compare models
  • Moderation
  • Pricing
Solutions
  • Customer support
  • Sales & lead gen
  • Internal knowledge
Industries
  • E-commerce
  • Insurance
  • Law firms
  • Healthcare
Resources
  • Blog
  • Glossary
  • Templates
  • Case studies
  • Documentation
  • Integrations
Company
  • About
  • Careers
  • Partner program
  • Contact
  • Book a demo
Trust & legal
  • Security
  • GDPR
  • Imprint
  • Privacy
  • Terms
© 2026 omniratio UG
Hosted in the EU · German & English