Two engines
The scoring is rules. The writing is AI.
Treating regulation and prose as the same problem is how you end up with a number you can't defend. We keep them apart on purpose.
Before the model
What reaches Mistral, and what doesn't.
Your raw answers
email: jane@acme.eu
SIRET: 123 456 789
phone: +33 6 12 34 56 78
sector: recruitment
→
What Mistral receives
email: EMAIL_1
SIRET: SIRET_1
phone: PHONE_2
sector: recruitment
Direct identifiers are swapped for neutral tokens before anything leaves. The business context that makes generation possible- sector, use case- is the only thing the model ever sees.
Our line in the sand
Things we will never do.
A score you can defend.
Run the deterministic audit, see exactly how your classification is computed, then let the verified documents follow.
Start free audit