Skip to main content

Pashto and Dari Content Moderation and Policy Localization for Platforms

Ariana Nexus provides Pashto and Dari content moderation and trust and safety policy localization for social platforms, messaging apps, marketplaces and AI labs. We rewrite your community guidelines and enforcement rules for Afghan context, build the lexicons, review high-risk and ambiguous content, calibrate your moderators, and test your classifiers and LLM moderation against the same rules.

Languages
Pashto, psپښتو
Dari, prsدری
Content types
Text
Posts
Images
Memes
Audio
Voice notes
Video
Short video
DSA Article 42
24
official EU languages in which the Digital Services Act requires very large platforms to report their moderators. Pashto and Dari are not on the list.7
EU TCO Regulation
1 hour
to remove terrorist content after an EU removal order under Regulation (EU) 2021/784.8
TAKE IT DOWN Act
48 hours
to remove non-consensual intimate images after a valid request under the U.S. TAKE IT DOWN Act, enforced by the FTC since 19 May 2026.10
Facebook, 2021
0
Pashto hate-speech classifiers Facebook reported having in October 2021, when it said it had them for Urdu.2
Coverage

Pashto and Dari in Perso-Arabic script and Romanized form, mixed with Urdu, Iranian Persian and English, in text, images, audio and video — from Afghanistan, Pakistan and the Afghan diaspora.

Led by scholars and alumni of
  • Cornell University
  • University of Chicago
  • University of British Columbia
  • Otto-von-Guericke University Magdeburg

Pashto and Dari content moderation at a glance

Languages
Pashto (Southern and Northern varieties, in Afghanistan and Pakistan) and Dari (Kabuli, Herati and Hazaragi, a variety of Dari)
Scripts and forms
Perso-Arabic script, Romanized Pashto and Dari, keyboard variants, and code-switching with Urdu, Iranian Persian (Farsi) and English
Content types
Text and comments, images and memes, audio and voice notes, video and livestream clips
Harm areas
12 Afghan-context harm areas, from violent extremism and hate speech to scams, doxxing and graphic war footage
Services
Policy localization, lexicons, expert review and escalations, moderator calibration, classifier and LLM evaluation, crisis surge support
Regulations mapped
EU Digital Services Act, EU Terrorist Content Online Regulation, UK Online Safety Act, U.S. TAKE IT DOWN Act, Australia’s Online Safety Act
Delivery
In-house team led from Washington, D.C., working inside your tools or a secure environment you control; no subcontractors
Engagement
Scoped after a confidential review of your policies and queues; no per-item pricing

Why Pashto and Dari content moderation fails on most platforms

Most platforms moderate Pashto and Dari with rules written in English, classifiers trained on other languages and reviewers hired for fluency rather than judgment. The result fails in both directions: harmful content stays up, and news, documentation and criticism come down.

Jan 2021

Rules that were translated, not localized

A January 2021 internal Facebook note found that only a few hate-speech reporting categories had been translated into Pashto, that the Pashto translation of “hate speech” did not seem accurate, and that the Dari instructions had the same problems. Reviewers and users cannot apply rules they cannot read correctly.1
Oct 2021

Automated detection that does not cover Pashto or Dari

In October 2021 Facebook said it had hate-speech classifiers for Urdu but not for Pashto. Large language models used for moderation still perform best in English and lose accuracy in low-resource languages, especially with dialect, slang and mixed scripts.2,6
2022–2023

Over-removal of news, documentation and criticism

Meta’s Oversight Board overturned the removal of a news report about the Taliban and girls’ education in 2022, and in 2023 Meta reinstated a post condemning the Taliban’s ban on girls’ education that it had removed under its dangerous organizations rules. Researchers in Pakistan documented the same pattern across Pashto and Dari content after August 2021.3,4,5
Ongoing

Evasion that exploits script and dialect

The same Pashto word can arrive with Persian, Urdu or Arabic keyboard letters, in Latin letters, or spoken in a voice note or a chanted tarana. Lexicons that do not normalize these variants miss them, and filters built for Iranian Persian misread Dari.
2024–2026

Regulators now ask about language

Under the Digital Services Act, very large platforms must assess systemic risks with regional and linguistic aspects in mind, and the UK Online Safety Act’s illegal content duties have applied since 17 March 2025. A Pashto or Dari blind spot is now a compliance question, not only a quality one.7,9
Also relevant
48 hours
Nationwide communications blackout in Afghanistan, 29 September to 1 October 2025.12

What our Pashto and Dari content moderation service includes

Six service lines, built into one program around your own policies.

SVC-01

Trust and safety policy localization

We localize your community guidelines and internal enforcement guidelines into Pashto and Dari, with Afghan-context definitions, worked examples, edge cases and the known questions your reviewers ask. User notices, statement-of-reasons text and appeal templates are written to be understood, not only translated.
  • Localized enforcement guidelines in Pashto and Dari
  • Definitions and worked examples for each policy
  • Known-questions log
  • Notice, statement-of-reasons and appeal templates
SVC-02

Lexicons and coded-language glossaries

Slurs, coded terms, extremist vocabulary, chants and meme conventions, each with Romanized spellings, keyboard variants and context notes, versioned and updated monthly.
  • Pashto and Dari lexicon with variants
  • Normalization rules for script and keyboard variants
  • Monthly change log
  • Emerging-term alerts
SVC-03

Expert review and escalations

First-language Pashto and Dari reviewers handle high-severity, ambiguous and appealed content inside your tools, with a second read on every critical decision and a written rationale tied to your policy.
  • Tier-2 and tier-3 review
  • Appeals and overturn review
  • Newsworthiness and documentation decisions
  • Decision records with policy citations
SVC-04

Moderator calibration and quality assurance

Golden sets, calibration sessions and blind QA sampling for your in-house team or your outsourcing vendor, so every Pashto and Dari seat applies the same decision to the same post.
  • Golden sets for each harm area
  • Calibration sessions
  • Blind QA sampling
  • Agreement reporting
SVC-05

Classifier and LLM moderation evaluation

We measure your automated moderation against the localized policy: precision and recall by harm area, language, script and dialect, and LLM-as-moderator prompts tested on Afghan-context edge cases.
  • Labeled evaluation sets
  • Error analysis by dialect and script
  • LLM policy-prompt testing
  • Retest after each model change
SVC-06

Crisis response and surge support

Rapid guidance and review capacity when events move faster than policy: attacks, elections, deportation waves, communications blackouts and viral misinformation.
  • Crisis guidance within hours
  • Surge review capacity
  • Emerging-narrative briefings
  • Post-incident review

Afghan-context harms we moderate in Pashto and Dari

A gold top rule marks harms that are critical by default. Final tiers are set with you.

VET1 Critical

Violent extremism and terrorist content

Taliban and ISKP (Islamic State Khorasan Province) propaganda, recruitment, glorification of attacks, and chanted taranas and nasheeds.

Maps to dangerous organizations; terrorism

HST2 High

Ethnic and sectarian hate speech

Slurs and dehumanizing language targeting Pashtun, Tajik, Hazara, Uzbek and other communities, anti-Shia rhetoric, and anti-Afghan hate in Pakistan, Iran, Türkiye and Europe.

Maps to hate speech

GVT1 Critical

Gender-based abuse

Threats and sexualized abuse aimed at Afghan women journalists, activists and athletes, honor-based threats, and non-consensual intimate images.

Maps to harassment; intimate image abuse

THT1 Critical

Threats and incitement

Calls for violence, blasphemy and apostasy accusations used to incite, and threats against former soldiers, officials and interpreters.

Maps to violence and incitement

PDT1 Critical

Doxxing and privacy exposure

Publishing the addresses, phone numbers, family names or identity documents of people at risk.

Maps to privacy; doxxing

GCT2 High

Graphic violence and war footage

Footage from four decades of war in Afghanistan, where the same video can be documentation, news or glorification.

Maps to violent and graphic content

MST2 High

Misinformation and scams

Fake evacuation, visa and humanitarian-parole offers, fake aid appeals, and rumors that spread during shutdowns.

Maps to fraud; misinformation

HTT1 Critical

Human smuggling and trafficking

Smuggling routes and services, offers of forged documents, and exploitative job ads aimed at migrants.

Maps to human exploitation

CST1 Critical

Child safety

Recognition of grooming and exploitation indicators in Pashto and Dari, escalated immediately under your legal reporting protocol.

Maps to child sexual exploitation

RGT2 High

Drugs, weapons and regulated goods

Sale of opium, heroin and methamphetamine, and of weapons and ammunition.

Maps to regulated goods

SHT1 Critical

Suicide and self-harm

Recognition of self-harm content expressed in Pashto and Dari idiom and poetry, routed to your safety resources.

Maps to suicide and self-injury

CIT2 High

Coordinated inauthentic behavior

Sockpuppet networks, hashtag campaigns and cross-border influence operations in Pashto and Dari.

Maps to platform manipulation

Pashto and Dari dialects, scripts and formats we cover

Pashto on the right, Dari on the left, wherever they sit side by side.

Pashto

پښتو
Varieties
Southern (Kandahari) and Northern (Yusufzai) Pashto, in Afghan and Pakistani usage across Kandahar, Nangarhar, Khyber Pakhtunkhwa, Balochistan and Karachi
Scripts
The Pashto alphabet in Perso-Arabic script, including the letters ټ ډ ړ ږ ښ ځ څ ګ ڼ ې ۍ, and Romanized Pashto
Note
Pashto and Pakhto are the same language, and Pashto content moderation has to read both: the ښ sound shifts from sh in the south to kh in the north, and Romanized spellings follow it.

Dari

دری
Varieties
Kabuli Dari, Herati, and Hazaragi, a variety of Dari
Scripts
Perso-Arabic script with Afghan spelling and vocabulary, and Romanized Dari
Note
Dari is not interchangeable with Iranian Persian (Farsi) for moderation: vocabulary, slang and political references differ. Farsi is read where it is mixed into Dari content; standalone Farsi or Urdu queues are scoped by engagement.

Why Pashto and Dari need their own lexicons

ټ
U+067C
ډ
U+0689
ړ
U+0693
ږ
U+0696
ښ
U+069A
ځ
U+0681
څ
U+0685
ګ
U+06AB
ڼ
U+06BC
ې
U+06D0
ۍ
U+06CD

Eleven letters of the Pashto alphabet that standard Persian and Arabic keyboards do not carry, with their Unicode code points. Text typed without them arrives in variant spellings.

Keyboard and encoding variants
PairLetters and code pointsWhere it comes from
PairPashto gaf and Persian gafLetters and code pointsګU+06ABandگU+06AFWhere it comes fromPashto typed on a Persian keyboard
PairPashto ṭe and Urdu ṭe, two encodingsLetters and code pointsټU+067CandٹU+0679Where it comes fromPashto typed on an Urdu keyboard in Pakistan
PairPersian ye and Arabic yeLetters and code pointsیU+06CCandيU+064AWhere it comes fromKeyboard variants in Dari; different letters in Pashto
PairPersian kaf and Arabic kafLetters and code pointsکU+06A9andكU+0643Where it comes fromArabic keyboards and copied text
PairTwo forms of the digit fourLetters and code points۴U+06F4and٤U+0664Where it comes fromDigit variants slip phone numbers past filters

Normalization has to know which language it is reading: in Dari, ی and ي are keyboard variants of one letter; in Pashto they are different letters and must not be merged.

Romanizationپښتو
  • Pashto
  • Pakhto
  • Pukhto
  • Pushtu

One word, four common Latin spellings. Southern speakers say Pashto; northern speakers say Pakhto.

The zero-width non-joiner (U+200C) belongs inside many Dari words. It is invisible, and it is also used to split a banned term so a filter no longer matches it.

په پښتو او دري ژبو کې د منځپانګې څارنه

نظارت بر محتوا به زبان‌های پشتو و دری

Content moderation in Pashto and Dari

Also read
Mixed with Urdu, Iranian Persian (Farsi), English and Arabic religious phrases. Uzbeki, Turkmeni, Balochi, Pashayi and the other Afghan languages we work in are read on request, with response times set by language.

How we deliver Pashto and Dari moderation and policy localization

  1. 01Weeks 1–3

    Review your policies, queues and risks

    We read your policies, tooling and a sample of Pashto and Dari decisions, map them to the 12 harm areas and your regulatory duties, and measure where decisions go wrong today.

    Output: Baseline error sample and gap log

  2. 02Weeks 3–6

    Localize the policy

    Enforcement guidelines, definitions, examples and edge cases are rewritten for Afghan context in Pashto and Dari, then reviewed by our legal and linguistic leads before release.

    Output: Localized enforcement guidelines

  3. 03Weeks 4–7

    Build the lexicon and golden sets

    Lexicons with script, keyboard and Romanized variants, and labeled golden sets for each harm area, language and dialect.

    Output: Versioned lexicon and golden sets

  4. 04Weeks 6–8

    Calibrate the reviewers

    Training and calibration for your in-house or vendor moderators until agreement on the golden sets meets the target we set with you.

    Output: Calibration record

  5. 05Ongoing

    Review, escalate and report

    Expert review of high-severity and ambiguous content, appeals and crisis surges, with a written rationale for every critical decision.

    Output: Decision records and monthly quality report

  6. 06Monthly

    Measure and update

    Monthly measurement of accuracy, agreement, appeal overturns and classifier precision and recall; lexicon and policy updates as language and events change.

    Output: Monthly report and change logs

How we measure moderation quality

Reported monthly by harm area, language, script and dialect.

What we report every month

  • Decision accuracy

    correct / reviewed

    Share of decisions that match the localized policy on blind golden-set items.
  • Reviewer agreement

    κ · α

    Inter-rater agreement (Cohen’s kappa or Krippendorff’s alpha) by harm area and language.
  • Appeal overturn rate

    overturned / appealed

    How often appealed decisions are reversed, and why.
  • Classifier precision and recall

    TP / (TP+FP) · TP / (TP+FN)

    By harm area, language, script and dialect, so a model that works for Dari but not Pashto shows up.
  • Time to action

    t(action) − t(receipt)

    By severity tier, against the targets in your contract.
  • Lexicon change

    Δ terms / month

    New and retired terms each month, with the events that produced them.

Severity tiers and response targets

TierWhat it coversResponse target
TierT1 CriticalWhat it coversCredible threats, terrorist content, child safety, intimate image abuse, doxxing of people at riskResponse targetAction within 1 hour of receipt
TierT2 HighWhat it coversHate speech, graphic violence, scams, smuggling, regulated goodsResponse targetAction within 4 hours
TierT3 StandardWhat it coversRoutine policy review and appealsResponse targetAction within 1 business day

How fast each obligation makes you act

1 hour

EU TCO removal order

4 hours

Our tier-2 target

48 hours

TAKE IT DOWN Act request

1 year

DSA risk assessment cycle

Logarithmic scale, one hour to one year. Gold marks our tier-2 target.

Regulations this work supports: DSA, TCO, Online Safety Act and TAKE IT DOWN Act

RegulationJurisdictionWhat it requiresWhere Pashto and Dari come inStatus
RegulationDigital Services Act7JurisdictionEuropean UnionWhat it requiresNotice and action, statements of reasons and internal appeals; for very large platforms, annual systemic risk assessments that account for regional and linguistic aspects (Article 34) and moderator reporting by official EU language (Article 42).Where Pashto and Dari come inPashto and Dari risk inputs, localized statements of reasons and appeal texts, and capacity evidence for languages Article 42 does not break out.StatusAll platforms since 17 February 2024; very large platforms since August 2023
RegulationTerrorist Content Online Regulation (EU) 2021/7848JurisdictionEuropean UnionWhat it requiresRemove or disable terrorist content within one hour of a removal order.Where Pashto and Dari come inPashto and Dari identification of designated-group propaganda and one-hour escalation paths.StatusApplies since 7 June 2022
RegulationOnline Safety Act 20239JurisdictionUnited KingdomWhat it requiresIllegal harms risk assessments and duties to remove illegal content quickly; children’s safety duties.Where Pashto and Dari come inPashto and Dari inputs to illegal harms risk assessments; review of terrorism, threats and intimate image abuse.StatusIllegal content duties since 17 March 2025
RegulationTAKE IT DOWN Act (Pub. L. 119-12)10JurisdictionUnited StatesWhat it requiresRemove non-consensual intimate images, including AI-generated ones, within 48 hours of a valid request.Where Pashto and Dari come inPashto and Dari request handling and victim-aware review.StatusFTC enforcement since 19 May 2026
RegulationOnline Safety Act 202111JurisdictionAustraliaWhat it requiresBasic Online Safety Expectations and removal notices from the eSafety Commissioner.Where Pashto and Dari come inPashto and Dari review for Australia’s Afghan communities.StatusIn force since January 2022

We work alongside your counsel and do not give legal advice. Regulatory dates on this page are reviewed every quarter.

Who uses our Pashto and Dari trust and safety services

Platforms

Social media and video platforms

Pashto and Dari queues, appeals and regulatory duties across feeds, comments, short video and livestreams.
Messaging

Messaging and voice apps

Channels, groups and forwarded voice notes, where much Pashto and Dari audio and misinformation travels.
AI labs

Generative AI labs

Safety classifiers and LLM moderation for Pashto and Dari prompts and outputs, tested against a localized policy.
Commerce

Marketplaces, classifieds and payments

Scams, smuggling, forged documents, weapons and drugs listed in Pashto and Dari.
Communities

Gaming, dating and community platforms

Harassment, grooming indicators, extremist symbols and doxxing in chat and profiles.
Vendors

Trust and safety outsourcing vendors

An Afghan-context policy, calibration and escalation layer for your Pashto and Dari seats.

Where platforms bring us in

Illustrative engagement types, not client case studies.

Before a DSA risk assessment

A four-week Pashto and Dari policy and risk review that feeds the linguistic section of a very large platform’s systemic risk assessment.

Calibrating a vendor’s Pashto and Dari seats

Golden sets and monthly blind QA for an outsourcing partner’s moderators, with agreement reported to the platform.

Before an LLM moderation launch

Evaluation of an LLM moderation prompt for Pashto and Dari before rollout in Afghanistan and Pakistan, with a retest after each model update.

During a communications blackout

Surge guidance and review during events like Afghanistan’s 48-hour nationwide shutdown in September 2025, when rumor and scams move to diaspora channels.

The team behind our Pashto and Dari trust and safety work

The people accountable for this work are scholars and alumni of Cornell University, the University of Chicago, the University of British Columbia and Otto-von-Guericke University Magdeburg who come from the Afghan community. They write the policy your moderators apply, measure the results, and answer for every critical decision.

Portrait of Hassan Ukasha
Program oversight

Hassan Ukasha

Managing Partner, Ariana Nexus

  • B.S.Cornell University
  • M.P.H.Cornell University
Hassan Ukasha oversees the firm’s operations and this program. He approves the scope of every engagement, sets the review standard your moderators are measured against, and signs off each localized policy before it reaches your queues. He is the executive point of escalation for critical incidents and is based in Washington, D.C.

Languages: Pashto and Dari (native); English, Urdu and Hindi; Arabic (working).

Zeba Haqbani

Zeba Haqbani

Senior Partner

  • B.Sc.
    University of British Columbia
Hussain Ahmad

Hussain Ahmad

Principal

  • M.Eng.
    Cornell University
  • Ph.D.
    University of Chicago
Wasil Peroz

Wasil Peroz

Principal

  • B.A.
    Milli University
  • M.Sc.
    Otto-von-Guericke University Magdeburg
Maryam Safi

Maryam Safi

Principal

  • B.A.
    Cornell University

Who owns what on your program

PersonOwns on your programSigns off
PersonHassan UkashaOwns on your programScope, review standard and executive escalationSigns offEvery localized policy before release
PersonZeba HaqbaniOwns on your programSecure review environment, tool integration and data flowSigns offEnvironment access and the deletion record at close
PersonHussain AhmadOwns on your programGolden sets, sampling plans, and classifier and LLM moderation evaluationSigns offEach evaluation report
PersonWasil PerozOwns on your programRegulatory mapping and designated-entity rulesSigns offDangerous-organization guidance before release
PersonMaryam SafiOwns on your programPashto and Dari lexicons, worked examples and the gap logSigns offEach monthly lexicon release
The review bench
Reviewers on the bench read Pashto or Dari as a first language and are not named on this page. Moderating extremist and gender-based abuse in Afghan languages carries personal risk for reviewers and their families, so we identify them by function, never by name.

Security, privacy and reviewer wellbeing

NDA first

We sign your NDA before we see any policy, queue or user content.

Your environment

Work happens inside your moderation tools or a secure virtual environment you control.

No retention

No copies of user content are kept after review, and you receive a written deletion record at close.

No training use

Client content is never used to train any model.

Reviewer location

Reviewer location is set by contract, including U.S.-only teams.

European work

GDPR and UK GDPR terms for European engagements.
Reviewer wellbeing

Diana Ayubi, Psy.D., Engagement Manager, sets and reviews the reviewer wellbeing protocol: exposure limits, rotation off graphic queues, and scheduled debriefs. Reviewers who come from the communities in Afghan war footage face a higher risk of trauma and moral injury, and the protocol is built for that risk.

Why Ariana Nexus for Pashto and Dari content moderation

No certification exists for Pashto or Dari trust and safety work in the United States.

We wrote the review standard, we train our reviewers and your vendors’ moderators to it, and every client sees it.

01

Scholars, not a pool of bilinguals

Leads are alumni and scholars of leading universities, and reviewers read Pashto or Dari as a first language. Fluency is the entry requirement, not the qualification.
02

Policy before queue

We localize the rules before we review a single post, so your staff, your vendors and your models apply the same decision.
03

Both kinds of error count

We measure wrongful removals of news, documentation and criticism as seriously as missed harm, because Afghan users lose in both directions.
04

Built for AI moderation

Golden sets, precision and recall by dialect and script, and LLM policy prompts tested against the localized rules: evidence your ML team can act on.
05

One accountable team

Every part is produced by our own people: one engagement, one point of accountability, no subcontractors, and no reuse of your data to train anything.
06

Care for the reviewers

A clinical wellbeing protocol for people who review violence from their own communities.

How our approach compares

CriterionMachine translation and generic classifiersOutsourced bilingual seatsAriana Nexus
Pashto and Dari policyMachine translation and generic classifiersWord-for-word translation of English rulesOutsourced bilingual seatsThe vendor’s generic guidelinesAriana NexusEnforcement rules rewritten for Afghan context, with examples and edge cases
Coded language and script variantsMachine translation and generic classifiersMissed unless spelled as trainedOutsourced bilingual seatsDepends on the individual moderatorAriana NexusVersioned lexicon with Romanized and keyboard variants, updated monthly
Taliban-related news and criticismMachine translation and generic classifiersOver-removal is commonOutsourced bilingual seatsInconsistentAriana NexusDesignated-entity guidance with tested news and condemnation rules
What gets measuredMachine translation and generic classifiersAggregate accuracyOutsourced bilingual seatsHandle time and volumeAriana NexusPrecision, recall and agreement by harm, language, script and dialect
Regulatory evidenceMachine translation and generic classifiersNoneOutsourced bilingual seatsLimitedAriana NexusInputs for DSA, TCO, Online Safety Act and TAKE IT DOWN duties
Reviewer wellbeingMachine translation and generic classifiersNot applicableOutsourced bilingual seatsVaries by vendorAriana NexusClinical protocol with exposure limits and debriefs
Abstract dark building geometry against a night sky
Answers

Pashto and Dari content moderation: frequently asked questions

What is Pashto and Dari content moderation?

It is the review of user-generated content written or spoken in Pashto and Dari against a platform’s rules: removing, labeling, restricting or keeping posts, comments, images, audio and video. Done well, it uses rules localized for Afghan context, lexicons that cover script and dialect variants, and reviewers who read the languages as first languages.

What is policy localization in trust and safety?

Policy localization adapts a platform’s community guidelines and internal enforcement rules to a language and culture. It goes beyond translation: definitions, examples, edge cases and exceptions are rewritten so a Pashto or Dari reviewer, or a model, reaches the decision the policy intends.

Can AI moderate Pashto and Dari content accurately?

Not on its own today. Automated classifiers and large language models perform best in English and lose accuracy in low-resource languages, especially with dialect, slang, Romanized text and audio. We measure your models against a localized policy and design the human review that covers what they miss.

Is Dari the same as Farsi?

Dari and Iranian Persian (Farsi) are closely related but not interchangeable for moderation. Vocabulary, slang, political references and spelling conventions differ, and a reviewer trained on Iranian content will misread Afghan context. We moderate Afghanistan Dari, including Hazaragi, a variety of Dari, and read Farsi where it is mixed into Dari content.

How do you handle Taliban-related content?

We apply your dangerous organizations or designated-entity policy with Afghan-context guidance that separates praise, support and representation from news reporting, neutral discussion and condemnation. Two Oversight Board cases involving posts about the Taliban and girls’ education show how often the second group is removed by mistake.

Do you moderate Romanized Pashto and Pashto from Pakistan?

Yes. Our Pashto content moderation covers Pashto in Perso-Arabic script and in Latin letters, Southern and Northern varieties, and usage in Khyber Pakhtunkhwa, Balochistan and Karachi as well as Afghanistan, including code-switching with Urdu and English.

Can you provide Pashto and Dari moderators for our queues?

We staff expert review for high-severity, ambiguous and appealed Pashto and Dari content, and we calibrate the in-house or outsourced moderators who handle volume. We do not sell per-item seats; engagements are scoped after a confidential review of your policies and queues.

Which regulations does this work support?

The EU Digital Services Act, including risk assessments that must account for linguistic aspects; the EU Terrorist Content Online Regulation’s one-hour removal duty; the UK Online Safety Act’s illegal harms duties; the U.S. TAKE IT DOWN Act’s 48-hour removal rule; and Australia’s Online Safety Act. We work alongside your counsel and do not give legal advice.

How do you protect moderator wellbeing?

Our protocol, set by a clinical psychologist, limits exposure to graphic content, rotates reviewers off the hardest queues and schedules debriefs. Many reviewers come from the communities shown in Afghan war and violence footage, which raises the risk of trauma and moral injury; the protocol is designed for that risk.

How do you protect our data and our users’ data?

We sign your NDA before we see anything, work inside your moderation tools or a secure environment you control, keep no copies of user content after review, provide a written deletion record at close, and never use client content to train models.

Do you cover other Afghan languages?

Content in Uzbeki, Turkmeni, Balochi, Pashayi and the other Afghan languages we work in can be read on request, with response times set by language. This service is scoped to Pashto and Dari, where platform volume and risk are highest.

Where is Ariana Nexus based, and how do we start?

Ariana Nexus is headquartered in Washington, D.C. Engagements start with a confidential briefing and a review of your policies and a sample of your Pashto and Dari queues. We sign your NDA first.

Are you hiring Pashto and Dari content reviewers?

We recruit reviewers directly and train them to our standard. Reviewers looking for work can apply through Ariana Nexus careers.

Pashto and Dari trust and safety terms, defined

Terms as we use them in policies, lexicons and reports.

Content moderation
Reviewing user-generated content against a platform’s rules and deciding whether to keep, label, restrict or remove it.
Policy localization
Adapting community guidelines and enforcement rules to a language and culture, including definitions, examples, edge cases and exceptions.
Enforcement guidelines
The internal rules reviewers apply, with definitions, examples and edge cases; more detailed than public community guidelines.
Lexicon
A maintained list of slurs, coded terms and variants for a language, with context notes, used by reviewers and detection systems.
Golden set
A labeled set of examples with agreed correct decisions, used to calibrate reviewers and measure accuracy.
Inter-rater agreement
A statistic, such as Cohen’s kappa or Krippendorff’s alpha, that measures how consistently different reviewers reach the same decision.
Designated entity
An organization or person restricted by a platform’s dangerous organizations policy, such as a group on a terrorism list.
Statement of reasons
The explanation a platform gives a user when it restricts their content, required under Article 17 of the Digital Services Act.
Tarana
A chanted, unaccompanied song in Pashto; the Taliban use taranas as a propaganda format.
Romanized Pashto
Pashto written in Latin letters, with no standard spelling.
Zero-width non-joiner
An invisible Unicode character (U+200C) used inside Persian-script words; it can also be used to split a banned term.
Hazaragi
A variety of Dari spoken by Hazaras.
Moral injury
Psychological harm that can follow exposure to acts that violate a person’s moral beliefs; a recognized risk in content review work.

Request a confidential briefing on your Pashto and Dari moderation

We sign your NDA before we see a single policy or queue item.

Request a confidential briefing