Un agent IA multilingue qui prévient, contextualise, contrecarre et transforme la haine en ligne — sans jamais simplement supprimer. Civic Guardian analyse le sens et le contexte d’un message sans avoir besoin de connaître l’identité de son auteur. Il protège la participation citoyenne sans transformer la modération en surveillance.
Le système qualifie le risque sans construire un outil de surveillance. Il prépare une réponse contextualisée, documentée et proportionnée.
Un agent IA multilingue qui prévient, contextualise, contrecarre et transforme la haine en ligne — sans jamais simplement supprimer.
Les chiffres sont synthétiques : ils illustrent les indicateurs à comparer pendant le prototype et dans la note de recherche.
Données agrégées et anonymisées sur sept jours.
Chaque score reste interprétable et soumis à contrôle humain.
Le système ne demande jamais plus de données que nécessaire. Chaque étape peut être inspectée, documentée et comparée dans la note de recherche.
Widget web, corpus de test ou connecteur validé. Aucun scraping incontrôlé.
Suppression ou transformation des identifiants inutiles avant toute inférence.
Transmission du minimum conversationnel nécessaire pour interpréter le message.
Classification multilingue, preuves textuelles, confiance et incertitude.
Mesure du risque d’exposition, d’inférence et de ré-identification.
Recommandation lisible et décision sensible laissée à une personne autorisée.
Le système ne bannit pas automatiquement et ne cherche pas à connaître l’auteur. Il structure un dossier de preuve explicable pour une revue humaine.
Les tendances agrégées aident à détecter les fenêtres d’escalade prioritaires sans exposer les individus.
Contenus nécessitant une revue.
Civic Guardian transforme la modération en assistance démocratique : il explique, contextualise, désamorce et documente, sans automatiser la censure.
Participer plus sereinement au débat public, reconnaître la manipulation et recevoir une contextualisation proportionnée.
Prioriser les cas urgents grâce à des signaux explicables, des preuves traçables et une revue humaine structurée.
Observer des tendances agrégées et anonymisées pour éclairer les politiques publiques sans exposer les individus.
Le MVP est ambitieux, mais livrable dans le temps du Hackathon.
Importer un fil de démonstration.
Neutraliser localement les identifiants.
Afficher haine, escalade, confiance et fuite.
Afficher signaux utiles et sources.
Valider toute décision sensible.
Le droit fixe la limite. La technologie la mesure. Le design la rend visible. La démocratie gagne en confiance.
Traduit les droits fondamentaux en règles opérationnelles et recours humains.
Conçoit le pipeline Privacy-HSD, l’orchestration et le prototype.
Rend les garanties visibles, compréhensibles et actionnables.
Protéger le débat public sans transformer la modération en surveillance ou profilage.
Pare-feu d’identité, analyse contextuelle et dossier de preuve explicable.
Droit, technologie et design coopèrent pour rendre les droits concrets.
Civic Guardian PrivHSD is a multilingual hate-speech detection system designed to analyse the meaning and context of a message without needing to know the author’s identity. It protects civic participation without turning moderation into a surveillance tool.
Privacy protection is not added at the end of the project. It determines what enters the engine, what is excluded and what remains subject to human review.
Remove unnecessary identifiers before inference and strictly limit retention.
Analyse the meaning of the message and the minimum conversational context that is genuinely necessary.
Compare HSD performance with residual leakage, inference and re-identification risks.
Display the main reasons, uncertainty and safeguards applied before any sensitive decision.
The project is designed to maximise the Privacy-HSD trade-off: maintain useful detection while reducing profiling, re-identification and malicious doxing risks.
The engine receives sanitised content and a limited Context Capsule. The author’s identity, precise location and cross-platform history are not required.
The Identity Firewall, multilingual semantic classifier, adversarial privacy tests, explainable scores and human validation are integrated from the architecture stage.
The proof of concept can be tested on synthetic or governed datasets, then deployed with NGOs, moderated communities and institutional stakeholders.
Data minimisation, short retention, visible uncertainty, human review, freedom of expression and the audit trail are not merely legal notices.
The PrivHSD challenge requires more than a high-performing engine. Law, engineering and user experience must work together so that privacy protection remains measurable, verifiable and understandable.
Advocate · Policy & Law
She translates Council of Europe principles into operational rules: authorised purposes, excluded data, human review and safeguards against surveillance creep.
Technologist · Engineering & AI
He designs the Identity Firewall, the Context Capsule, the auditable NLP pipeline and the indicators that genuinely measure the trade-off between utility and privacy.
Designer · UX & Visualisation
She makes human rights visible in the interface: anonymisation status, score explanations, uncertainty, audit trail and human validation.
This continuous dialogue prevents two frequent failures: technically brilliant but legally dangerous innovation, or abstract compliance that is impossible to understand in real-world use.
The system never requests more data than necessary. Every step can be inspected, documented and compared in the research note.
Web widget, test corpus or validated connector. No uncontrolled scraping.
Removal or transformation of unnecessary identifiers before any inference.
Transmission of the minimum conversational context required to interpret the message.
Multilingual semantic classification, textual evidence, confidence and uncertainty.
Measurement of exposure, attribute-inference and re-identification risks.
Readable recommendation, audit trail and sensitive decision left to an authorised person.
The sequence starts when this section reaches the centre of the screen. The scenario illustrates anonymised example messages and shows what the engine retains - and above all what it refuses to use.
“Thank you to the residents who took part in our civic workshop this weekend. The discussions were simple, warm and useful. This is how we build a city where everyone can find their place.”
The displayed data are synthetic. They illustrate the research dashboard used to compare engine utility before and after data sanitisation.
Civic Guardian is designed to protect dignity, privacy and freedom of expression simultaneously. Each legal principle is translated into a visible, understandable and auditable control.
Minimisation, limited purposes, proportionate processing and control mechanisms adapted to digital technologies.
Unnecessary identifiers are neutralised before inference. Remaining traces are time-limited and deletable.
The system distinguishes hate speech from lawful controversy, satire and legitimate democratic disagreement.
Transparency, accountability, a risk-based approach and human oversight structure the system lifecycle.
Democratic participation cannot be safe if citizens must choose between enduring hate or being profiled in order to be protected. Civic Guardian rejects this false dilemma.
The status confirms that the Identity Firewall worked before inference.
The remaining exposure level and the applied tests are displayed.
The user understands why an alert was produced and can discuss it.
The person knows whether an action has been proposed, reviewed or validated.
Civic Guardian transforms raw signals into proportionate, privacy-respecting information. The objective is not to control debate, but to allow everyone to participate without enduring hate or giving up their rights.
Understand hate signals, receive a clear explanation and regain a safer discussion space without being profiled.
Protected democratic participationPrioritise alerts, contextualise situations and intervene consistently through explainable scores.
Proportionate and documented actionReceive aggregated and structured trends to inform public consultations without unnecessarily exposing people.
Transparency and institutional trustCompare HSD performance with residual leakage risks and share reproducible results to improve practices.
Evidence-based decisionsExample conversation, Identity Firewall, Context Capsule, NLP classification, leakage tests, utility/privacy comparison and explainable dashboard.
Tests on synthetic or governed datasets, followed by experimentation with NGOs or authorised moderated communities under human oversight.
Modular moderation-support service, privacy-respecting aggregated observatory and contribution to a scientific publication.
A multilingual tool that detects hate speech without depending on the author's identity. Civic Guardian analyses harm, contextualises the situation and structures a civic response without turning moderation into surveillance.
Our ambition: maximise HSD utility while minimising privacy loss and re-identification risks.
Beyond deletion: act on social dynamics while protecting people.
Online hate combines harassment, ideological framing, false claims and emotional contagion. It erodes social cohesion. Yet an overly intrusive tool can also become an instrument of profiling, tracking or re-identification.
The system neutralises identifiers before inference, keeps a limited context capsule and produces an explainable evidence bundle. Civic Guardian does not automate censorship; it structures trusted response.
A multilingual AI agent that prevents, contextualises, counters and transforms online hate - without ever simply deleting it.
A modular, locally deployable and measurable Privacy-HSD pipeline.
The system never requests more data than necessary. Every step can be inspected, documented and compared in the research note.
Web widget, test corpus or validated connector. No uncontrolled scraping.
Removal or transformation of unnecessary identifiers before any inference.
Transmit the minimum conversational context needed to interpret the message.
Multilingual semantic classification, textual evidence, confidence and uncertainty.
Measure exposure, attribute-inference and re-identification risks.
Readable recommendation, audit log and sensitive decision left to an authorised person.
Inspectable multilingual models, documented rules and reproducible scoring.
Locally controlled workflow, separated from the scientific engine and sensitive decisions.
Strictly necessary anonymised traces, limited retention and verifiable deletion.
Component inventory, applicable licences and documented compliance.
Trust principle: no unlimited surveillance, no author profiling, no cross-platform citizen tracking and no opaque automated sanction. Civic Guardian measures risk before claiming to reduce it.
Identifiers removed before analysis. No author profile is created.
✓ The HSD core remains local, separable and auditable.
✓ Airtable is reserved for project management or synthetic data.
✓ A proprietary LLM may intervene only after anonymisation, for reformulation - never to decide the HSD score.
✓ A software inventory and licence register document every component.
Turn the principles of the New Democratic Pact into a practical tool for real-world use.
Participate more safely in public debate, recognise manipulation and receive proportionate contextualisation.
Prioritise urgent cases through explainable signals, traceable evidence and structured human review.
Observe aggregated and anonymised trends to inform public policy without exposing individuals.
Civic Guardian turns moderation into democratic assistance. It does not merely remove harmful content: it explains, contextualises, de-escalates and educates. Privacy protection becomes a condition for civic participation.
The dashboard can isolate forms of targeting: gender, perceived origin, religion or orientation. Views are aggregated to identify coordinated campaigns without creating individual profiles.
Demonstrate the effectiveness of the analytical hate-speech detection model without needing to ban or identify the author.
From a demonstration thread, Civic Guardian must recognise hate speech, neutralise unnecessary identifiers, display understandable signals and suggest a proportionate response subject to human review.
Three complementary profiles: rules, engine and user experience.
Translates fundamental rights into ethical rules, usage limits and human-recourse requirements.
Designs the Privacy-HSD pipeline, n8n orchestration, prototype and API integrations.
Makes safeguards visible, understandable and actionable for users and the jury.
Policy alignment | Technical innovation | Synergy
Protect public debate without turning moderation into surveillance or profiling.
Identity Firewall, contextual analysis and an explainable evidence bundle to act without identifying the author.
Law, technology and design work together to make rights concrete, visible and applicable.
No mass surveillance, no hidden cross-platform tracking, no re-identification tool, no citizen scoring and no opaque automated sanction.
Civic Guardian turns hate speech into active democracy: it detects, explains, contextualises, de-escalates and structures trusted response. It protects people without creating a new surveillance tool.
Civic Guardian helps detect, explain, contextualise and de-escalate online attacks. It protects people without requiring the author's identity and without automating censorship.