Trust, Safety & Content Integrity in Afghan Languages

Your trust & safety stack is multilingual on paper. Under the DSA and the Online Safety Act, paper is not a defense.

Native-grade trust & safety and content integrity for Afghan-language content — detecting, classifying, assessing, and mitigating harmful content across all 24 languages, aligned to DSA Articles 34–35, the UK Online Safety Act, the EU AI Act, and the OWASP Top 10 for LLM Applications. Speech-protective, reviewer-protected, evidence-governed.

Convened by Ariana Nexus · AI & Data Systems Practice · Washington, D.C.
Exhibit 01 · The coverage gap
Reviewed at scale. Unread in 24 languages.
Covered by current classifiersAfghan-language · unreviewed
24 languages in scope — 0 with native-grade classifiers.
Harmful content in Pashto, Dari and 22 more moves through DSA-scoped platforms faster than any monolingual stack can see it. Illustrative.
20.5M
CyberTipline reports to NCMEC in 2024 — generative-AI reports up 1,325% year over year
~20M/day
content-moderation decisions logged in the EU DSA Transparency Database
88/100
LLM responses that became health disinformation on instruction — Annals of Internal Medicine, 2025
9%
of adults are confident they can identify a deepfake — Ofcom, 2024
THE PROBLEM

The harm your platform answers for is in languages your systems cannot read

The harmful content your platform answers for is in languages your safety systems never learned to read.

A platform’s trust and safety stack — its classifiers, its moderators, its escalation paths — works in the languages it was built for. The terrorism recruitment, the abuse, the gender-based violence, the targeted harm in Pashto and Dari run underneath it: undetected, unclassified, and unreviewed, because nothing and no one in the pipeline reads the language. The content does not become less harmful for being unread. It becomes invisible.

The law does not accept invisibility as a defense. Under the DSA, very large platforms and search engines must assess and mitigate systemic risks across their services every year — illegal content, fundamental-rights harms, threats to civic and electoral integrity, gender-based violence, harms to minors — and the regulation states plainly that the assessment must take regional and linguistic aspects into account. The UK Online Safety Act imposes parallel duties under Ofcom. A risk you cannot read is a risk you cannot assess, and a risk you cannot assess is a finding waiting in an audit.

And the systems meant to close the gap are themselves an attack surface. The AI classifiers doing the moderation can be evaded, manipulated, and jailbroken — most easily in the low-resource languages they were trained on least. Securing them is its own discipline.

Ariana Nexus is the layer that reads what the stack cannot: native-speaker detection and classification across all 24 Afghan languages, classifier evaluation and security testing, and systemic-risk assessment and mitigation documented for the DSA and the Online Safety Act — speech-protective, reviewer-protected, and strictly within legal frameworks.

DSA Art 34(2)
— the law requires accounting for regional and linguistic aspects; most assessments do not.
24
— Afghan languages your moderation stack likely does not cover.
The Multilingual Trust & Safety Standard
— our coverage method.
40languages a typical moderation stack covers
+24Afghan languages it likely does not
The coverage gap — 40 languages a typical stack covers; 24 Afghan languages remain uncovered (gold outline): the blind spot a DSA systemic-risk assessment must account for.
Visualization: of sixty-four language tiles, forty light up as covered by a typical moderation stack; twenty-four Afghan-language tiles remain dark with a gold outline, representing the coverage gap.
Evidence ledger

The scale is documented. The stakes are not in dispute.

Eight indicators that frame the trust-and-safety problem for a regulated, healthcare-relevant platform — each traced to a primary or authoritative source. Industry-vendor figures are labeled as such.

Indicator
What it measures · source
Domain
20.5M
CyberTipline reports of suspected child sexual exploitation received by NCMEC in 2024; reports involving generative AI rose 1,325% year over year.
NCMEC CyberTipline data, 2024
CSAM · scale
88/100
Responses from instruction-tuned LLM chatbots that were health disinformation, delivered in an authoritative tone; four of five models did so 100% of the time.
Modi, Menz et al., Annals of Internal Medicine, 2025
Health AI
51%
Vaccine-related social posts carrying health misinformation (up to); up to 28.8% of COVID-19 posts and up to 60% of pandemic-related posts, per a WHO-commissioned review.
WHO Europe systematic review, 2022
Health misinfo
~20M/day
Content-moderation decisions logged by EU platforms in the DSA Transparency Database — ~3.57B over ~180 days, roughly half fully automated.
European Commission, DSA Transparency Database
Moderation scale
Increase in detected deepfakes from 2023 to 2024 in identity-verification flows; deepfakes now make up roughly 7% of all fraud attempts.
Sumsub Identity Fraud Report, 2024 · industry
Synthetic media
$12.5B
Consumer fraud losses reported to the FTC in 2024 — up 25% year over year — much of it driven by impersonation and manipulated content.
FTC Consumer Sentinel Network Data Book, 2024
Economic
6%
Maximum fine under the Digital Services Act, as a share of a provider's total worldwide annual turnover — plus periodic penalties for continued non-compliance.
Digital Services Act, Reg. (EU) 2022/2065, Art. 52
Regulatory
9%
Adults confident they can identify a deepfake; 43% reported seeing at least one in the prior six months — rising to half of children aged 8–15.
Ofcom deepfakes research, 2024
Public trust
Verified June 2026 against primary and authoritative sources. Ranges marked “up to” are reported ceilings. Figures update annually; re-pull before external citation.
DEFINITION

What is multilingual trust, safety, and content integrity?

Trust, Safety & Content Integrity is multilingual trust-and-safety and content-integrity support for Afghan-language content — helping platforms, AI developers, and institutions detect, classify, assess, and mitigate harmful content across all 24 Afghan languages, aligned to DSA Articles 34 and 35, the UK Online Safety Act, and the OWASP Top 10 for LLM Applications. It covers systemic-risk assessment and mitigation, in-language classifier evaluation, and the security of the AI systems doing moderation, performed by native-speaker experts with reviewer-wellbeing protections and a speech-protective, rights-respecting methodology. Ariana Nexus provides linguistic, cultural, and methodological expertise within legal frameworks; it does not handle, store, or process illegal content, which remains with legally authorized entities.

DOCTRINE

A platform is accountable for systemic risk across all of its content, but its safety systems see only the languages they were built for — so the harm in every other language runs underneath the dashboard, unread and unmitigated. Coverage is not a translation feature added later; it is the difference between a safety system and a safety claim.

Safety is not monolingual.
BINDING RAILS

The boundaries are firm and explicit

01
Speech-protective and viewpoint-neutral. The firm helps detect and mitigate genuinely illegal or harmful content and platform-safety risk — never lawful expression. This mirrors the DSA’s own red line: suppressing lawful information is not the goal.
02
The firm does not handle, store, view, or process illegal content. For child sexual abuse material in particular, handling is restricted to legally authorized entities (e.g., NCMEC referral pathways). The firm provides linguistic, cultural, and methodological expertise, classifier evaluation, and policy support — within legal and mandatory-reporting frameworks — and never the content itself.
03
Reviewer wellbeing is protected. Human review of distressing content is bounded, supported, and humane — never the extractive, unprotected labor the industry has been rightly criticized for.
04
Privacy and civil liberties. The firm supports platform and community safety; it does not conduct or enable mass surveillance of the diaspora.
05
Serves lawful trust & safety — not state censorship. Clients are platforms, AI developers, and institutions meeting safety and regulatory duties; the firm does not build tools for government speech-policing.
OPERATING MODEL

One practice. Three coordinated capabilities.

Three institutional capabilities, orchestrated into trust and safety that actually covers the languages your users speak.

HIC · HUMAN INTELLIGENCE COLLECTIVE

The people who read what classifiers cannot

Lived-expertise practitioners across all 24 Afghan languages; the cultural gatekeepers who keep every engagement anchored in ground truth, never extractive.

Native-speaker trust-and-safety experts across all 24 languages who read, classify, and contextualize harmful content — catching the coded and contextual harm classifiers miss — with reviewer wellbeing protected throughout.

Protocol: The Multilingual Trust & Safety Standard.
ADF · AI DATA FACTORY

The data and the security testing

Governed Afghan-language data infrastructure, evaluation benchmarks, and institutional-grade training assets meeting auditable standards.

Classifier-evaluation data and harmful-content taxonomies mapped to the DSA risk categories; in-language test and red-team sets; OWASP-LLM-aligned security testing of moderation AI; de-identified analytics.

Protocol: The ADF Pipeline.
CCB · CULTURAL COMPLIANCE BUREAU

The governance layer

An audit-grade review regime translating cultural intelligence into compliance-ready practice — the governance layer threading through every engagement.

Risk methodology and the mapping to DSA Articles 34/35 and the Online Safety Act; cultural and contextual validation of classifications; reviewer-wellbeing, rights, and speech-protection governance; the CCB Sign-Off Mark on risk assessments.

Protocol: The CCB Sign-Off Mark.
Three capabilities. One safety system that sees what it is accountable for.
THE PATH

How Ariana Nexus closes the coverage gap: the Multilingual Trust & Safety Standard

The Multilingual Trust & Safety Standard closes the language-coverage gap in content moderation; the Five-Gate Validation Protocol governs the rigor, the rights, and the reviewer.

The Five Gates

GATE 1

Linguistic Accuracy

Harmful-content detection and classification linguistically accurate across all 24 languages and dialects.

GATE 2

Cultural Validity

Contextual and cultural validity of classifications, since harm is context-dependent and often coded; the Cultural Hallucination Audit applied to moderation classifiers; cleared by the CCB Sign-Off Mark.

GATE 3

Standards Conformance

DSA Articles 34 and 35, the UK Online Safety Act, and the OWASP Top 10 for LLM Applications; DSA transparency and independent-audit obligations.

GATE 4

Population Risk

Speech-protective and viewpoint-neutral, targeting illegal and harmful content rather than lawful expression; reviewer wellbeing protected; privacy and civil liberties preserved; no handling of illegal content, which remains with legally authorized entities.

GATE 5

Institutional Sign-Off

Risk assessments, mitigation measures, and classifier evaluations documented, traceable, and ready for DSA and Online Safety Act regulators and independent auditors.

The Four-Phase Orchestration Cycle

Diagram: the four-phase orchestration cycle traced as a continuous path from Phase one Situation through Phase four Measured Outcome.

I · Situation — Understand.

The platform or system, its Afghan-language exposure, and its DSA and Online Safety Act obligations mapped; the in-language systemic-risk picture assessed.

Cultural mapping · stakeholder calibration · constraint discovery.

II · Complication — Architect.

The multilingual coverage plan, classifier-evaluation framework, DSA-aligned harmful-content taxonomies, and mitigation measures designed.

Program scaffolding · compliance baseline · governance charter.

III · Resolution — Deploy.

In-language detection and classification evaluated and improved; mitigation implemented; harmful content addressed within legal frameworks; reviewers supported.

In-context execution · data infrastructure.

IV · Measured Outcome — Govern.

Systemic-risk assessment and mitigation documented for the DSA and Online Safety Act; classifier performance monitored; coverage maintained as content and threats evolve.

Continuous documentation · red-team validation · multi-decade horizon.
Active throughout: all three capabilities at full intensity — the harm in low-resource languages is exactly where human judgment, governed data, and cultural validation each matter most.
Why coverage holds

Coverage is a discipline, not a feature

Behind every classification is a native speaker, a documented standard, and a reviewer whose wellbeing is protected. That is what makes a safety system defensible under audit — and what a translated dashboard alone can never show.

DSA ARTICLES 34 & 35

The systemic-risk categories the law names — mapped to the languages it requires you to account for

Illegal contentFundamental-rights harmsCivic & electoral integrityGender-based violenceHarms to minorsCoverage across24 Afghan languagesASSESSED · MITIGATED · DOCUMENTED
Illegal content→ coverage across 24 languages
Fundamental-rights harms→ coverage across 24 languages
Civic & electoral integrity→ coverage across 24 languages
Gender-based violence→ coverage across 24 languages
Harms to minors→ coverage across 24 languages
Mandate register

The obligations you answer to — with their status, as of June 2026

The instruments that govern trust, safety, and content integrity for a healthcare-relevant platform. Status is tracked, not assumed: a proposal is not a law, and a contested rule is flagged as such.

Instrument
Scope for trust & safety
Status
EU Digital Services Act
Reg. (EU) 2022/2065 · Arts. 34–35, 37
Annual systemic-risk assessment, mitigation, and independent audit for very large platforms. Public health is a named systemic-risk category. First fine — €120M against X — adopted Dec 2025, under appeal.
In force
Applicable 17 Feb 2024
EU AI Act
Reg. (EU) 2024/1689 · Art. 50, Ch. V, Annex III
Labeling of AI-generated and manipulated content, AI-interaction disclosure, and high-risk duties. Health chatbots fall under Art. 50; clinical AI is high-risk.
Phased
Art. 50 + high-risk: 2 Aug 2026
UK Online Safety Act 2023
c. 50 · Ofcom Illegal Harms & Children Codes
Duty to assess and mitigate illegal content and protect children, with highly effective age assurance. Suicide, self-harm, and eating-disorder content is Primary Priority Content.
In force
Children's duties from 25 Jul 2025
US TAKE IT DOWN Act
Pub. L. 119-12 · FTC enforcement
Criminalizes non-consensual intimate imagery, including AI deepfakes, and requires platforms to remove reported material within 48 hours.
In force
Platform duty live 19 May 2026
Kids Online Safety Act
S. 1748 (119th Congress)
Would impose a minor-protection duty of care and safe-by-default design. It is not law: it died in the House in 2024 and is pending in committee.
Proposed
No floor vote
Section 1557 — Nondiscrimination
45 CFR 92.210 · 89 Fed. Reg. 37522
Bars discrimination through clinical decision-support tools and AI in patient communications. The decision-support provision took effect, but portions are vacated or enjoined.
Litigated
DSI duty effective 1 May 2025
ONC / ASTP HTI-1 Final Rule
45 CFR 170.315(b)(11)
Source-attribute transparency for predictive and AI decision-support in certified health IT — thirty-one disclosed attributes for predictive interventions.
In force
Compliance 31 Dec 2024
C2PA / Content Credentials
C2PA spec v2.x · Content Authenticity
Cryptographically signed media provenance to counter synthetic and misleading content. An open standard with broad adoption — no legal mandate.
Voluntary
Conformance program, 2025
NIST AI RMF + ISO/IEC 42001
NIST AI-600-1 · ISO/IEC 42001:2023
Voluntary frameworks for managing AI risk — including information integrity and confabulation — and a certifiable AI management system.
Voluntary
GenAI Profile, Jul 2024
Status as of 14 June 2026. Proposals are marked as such and carry no legal force until adopted; contested rules are flagged “litigated.” Not legal advice — confirm against the current text before relying on any item.
THE RECORD

What happens when safety stops at English

Platforms accountable for systemic risk under the DSA and the Online Safety Act carried that accountability into every language they served — including the ones their safety systems could not read. The terrorism recruitment, the abuse, the gender-based violence, the harm to minors in Afghan languages ran underneath the dashboard: undetected because no classifier was trained for it, unreviewed because no moderator could read it.

The content did not become less harmful for being invisible. And when the regulator asked how the risk had been assessed, the answer — in those languages — was that it had not been. A clean transparency report is no defense against a systemic risk the platform was never equipped to see.

Safety that stops at English is not safety — it is exposure.
Readiness ladder

Where your safety posture stands today

Multilingual trust and safety is a maturity curve, not a switch. A defensible position under the DSA and the Online Safety Act sits at the top of it — and most platforms are not there yet.

L1
Monolingual by default
The safety stack operates in a handful of high-resource languages. Harm in every other language is unmeasured — not absent.
Unmeasured exposure
L2
Translation-patched
Machine translation is bolted on, but classifiers remain trained on English-centric data. Coded, cultural, and contextual harm still slips through.
Most platforms are here
L3
Coverage-aware
In-language detection is piloted for priority languages and the gaps are documented — but not yet closed across the long tail you answer for.
Gaps documented
L4
Natively covered
Native-speaker detection and classification span every in-scope language, with reviewer wellbeing and speech protection governed throughout.
Coverage closed
L5
Governed & audit-ready
Systemic-risk assessment, mitigation, and evidence are documented for the DSA and the Online Safety Act — and coverage is maintained as content and threats evolve.
Defensible
Ariana Nexus moves a platform from L2 to L5 — and keeps it there — through the Multilingual Trust & Safety Standard and the Five-Gate Validation Protocol.
PARTNERSHIP

Your safety systems, extended to the languages they miss

From foundations to continuous stewardship.

1/4

Foundations

Scoped, mapped, assessed.

The platform, its Afghan-language exposure, and its DSA and Online Safety Act obligations understood.

2/4

Activation

Built to standard.

The coverage plan, classifier-evaluation framework, and DSA-aligned taxonomies designed.

3/4

Operating Rhythm

The active state.

Detection and classification evaluated and improved; mitigation implemented; reviewers supported.

4/4

Continuous Stewardship

As threats evolve.

Risk assessment and mitigation documented and maintained; coverage held over time.

THE RECEIVABLES

What an engagement returns

A multilingual trust-and-safety coverage assessment.

Where your systemic-risk obligations meet your language blind spots — DSA Article 34/35 aligned.

In-language harmful-content detection and classification evaluation.

How your classifiers actually perform in Pashto, Dari, and 22 more.

A Cultural Hallucination Audit of your moderation models.

The fluent-but-wrong classifications that pass automated review.

OWASP-LLM-Top-10-aligned security review of your moderation AI.

The model doing the moderation, tested as an attack surface.

Systemic-risk assessment and mitigation documentation.

Ready for the DSA, the Online Safety Act, and independent audit.

Native-speaker human review, with reviewer wellbeing protected.

Real judgment, humanely sourced.

Within legal frameworks, always.

The firm provides expertise and evaluation; it does not handle illegal content, which stays with authorized entities.

Speech-protective and rights-respecting throughout.

Harmful content addressed; lawful expression protected.

What you receive is not a promise that nothing slips through. It is the truth about what your safety systems can and cannot see in the languages your users actually speak — and the evidence that you looked.
PROOF
24
Afghan languages
0
Security incidents
100%
Senior-led engagements
41+
Trust Center documents
[classifiers evaluated / coverage metric — operator to supply]
GLOBAL REACH

The platform is global. The harm is in a language — and you answer for it.

The DSA reaches anyone serving the EU market, the Online Safety Act reaches services with UK users, and platform-safety duties are multiplying worldwide — while harmful content crosses every border in languages most safety systems do not cover. Ariana Nexus extends trust and safety across all 24 Afghan languages, worldwide, within the law of each jurisdiction.

ONLINE SAFETY ACT · UKDSA · EUROPEAN UNIONGULF & ARAB STATESUnited StatesUnited KingdomSwedenNetherlandsGermanyFranceAustriaItalyUAEQatarSaudi Arabia
United States · anchor
United Kingdom · Online Safety Act (Ofcom)
European Union · DSA Articles 34/35
Gulf & Arab states with significant Afghan communities
United StatesAnchor · Washington, D.C.
United KingdomOnline Safety Act · Ofcom
GermanyDSA · European Union
FranceDSA · European Union
ItalyDSA · European Union
SwedenDSA · European Union
NetherlandsDSA · European Union
AustriaDSA · European Union
United Arab EmiratesSignificant Afghan community
Saudi ArabiaSignificant Afghan community
QatarSignificant Afghan community
The regime differs by border. The blind spot is the same language everywhere.
THE DOOR

Request a Trust & Safety Coverage Review.

For platforms, including very large online platforms and search engines; AI developers building moderation systems; trust-and-safety teams; and content-integrity functions. Speech-protective and within legal frameworks. Briefings are conducted under NDA, in Washington, D.C. or virtually.

Request a confidential briefingHave a specific scenario you would like assessed — a language, a content category, a regulatory deadline? Bring it to the briefing.
The harm you cannot read is the harm you still answer for. Read it.
The Multilingual Trust & Safety Standard · Standards adherence (DSA Articles 34/35, the UK Online Safety Act, the OWASP Top 10 for LLM Applications) · Five-Gate Validation Protocol · The Cultural Hallucination Audit · Reviewer-wellbeing and speech-protection commitments · CCB Sign-Off Mark. Full index →
Updated June 2026 · living document, reviewed quarterly.