Maalamalama All articles
Arts & Culture

Flagged, Filtered, and Forgotten: How Social Media's Moderation Machine Is Muting the World's Loudest Voices

Maalamalama
Flagged, Filtered, and Forgotten: How Social Media's Moderation Machine Is Muting the World's Loudest Voices

Imagine spending three weeks crafting a video — writing, filming, editing, translating captions — only to wake up and find it gone. No explanation. No appeal that goes anywhere. Just a notification that your content "violated community guidelines," and a silence so loud it costs you sponsorships, followers, and sometimes your entire livelihood.

This isn't a hypothetical. For creators outside the US and Western Europe, it's practically a Tuesday.

Social media platforms have spent years marketing themselves as democratizing forces — tools that give anyone with a phone and a story the power to reach millions. And in many ways, they have. But the fine print of that promise contains a massive asterisk: the systems deciding what's acceptable, what's amplified, and what quietly disappears were built by teams that largely don't look like, sound like, or share the cultural context of the majority of the world's internet users.

The Algorithm Has a Hometown

Content moderation at scale is a two-headed beast. There's the automated layer — AI systems trained on datasets to detect nudity, violence, hate speech, and misinformation — and the human layer, teams of reviewers making judgment calls when the machines flag something ambiguous. Both layers have the same problem: they reflect the cultural assumptions of whoever designed and trained them.

Researchers at the University of Toronto's Citizen Lab and organizations like Article 19 have documented this pattern for years. Languages spoken by hundreds of millions of people — Amharic, Tagalog, Yoruba, Burmese — are dramatically underrepresented in moderation training data. That means automated systems are essentially guessing when they encounter content in those languages, and guessing wrong at alarming rates.

A 2022 report from Global Voices found that Arabic-language content was being removed at disproportionate rates compared to equivalent English content — even when the Arabic posts contained no policy violations whatsoever. The same study noted that content discussing Palestinian cultural identity was being suppressed at rates that Palestinian journalists described as "digital ethnic cleansing" of their online presence.

That's not a bug report. That's a structural failure.

When Culture Looks Like a Crime

Here's where it gets personal. Ask Priya, a classical Indian dancer based in Chennai with nearly 400,000 Instagram followers, about her content moderation experience, and she'll tell you about the saree.

She posted a traditional Bharatanatyam performance — a dance form that dates back over 2,000 years, performed in temples across South India, and recognized by UNESCO as a significant cultural heritage practice. Instagram's system flagged it for "sexual content" because the costume, like all authentic Bharatanatyam attire, includes a fitted blouse and exposes the midriff. The video was removed. Her account was temporarily restricted. A Western influencer in a crop top and leggings posting a workout video the same day? Still up.

"I've been doing this since I was six years old," she told us. "It's sacred to me. And some algorithm decided it was pornographic."

Similar stories echo across communities. West African creators report traditional ceremonial content — scarification rituals, ancestral practices, healing ceremonies — being flagged for "graphic violence" or "dangerous activities." Middle Eastern creators posting about Ramadan have had religious content mistakenly caught in filters designed to catch extremist material because certain Arabic phrases appear in both contexts. Mexican creators have had Día de los Muertos imagery removed for violating policies on "morbid content."

Each removal, on its own, might seem like a fixable mistake. Collectively, they form a pattern: non-Western cultural expression is being systematically misread by systems that weren't built to understand it.

The Economic Toll Nobody's Counting

For creators in the Global South, the stakes are higher than bruised feelings. A demonetized video or a restricted account isn't just an inconvenience — it can be the difference between paying rent and not.

The creator economy has opened real income pathways in countries where traditional employment opportunities are limited or unstable. In Nigeria, Kenya, Indonesia, and the Philippines, content creation has become a legitimate profession for thousands of young people. Brand partnerships, ad revenue, and platform monetization programs have given creators in these markets access to income that rivals — and sometimes exceeds — local professional salaries.

But platform monetization programs themselves carry built-in inequities. YouTube's Partner Program, for instance, has historically paid out lower CPM (cost per thousand views) rates for content in non-English languages, even when that content reaches massive audiences. And if your account gets flagged repeatedly — even incorrectly — you can lose monetization eligibility entirely. The appeals process, when it exists, is slow, opaque, and often conducted in English only.

A creator in Lagos who loses a brand deal because their account was temporarily restricted doesn't have the same safety net as a creator in Los Angeles facing the same situation. The financial consequences compound in ways that platform executives in Menlo Park rarely have to consider.

The Human Reviewers Aren't a Fix

It's tempting to think the solution is simple: hire more human moderators who speak the relevant languages and understand the relevant cultures. And yes, that would help. But the human review layer has its own deep problems.

Content moderation work is notoriously traumatic — reviewers are exposed to the worst of what humans post online, day after day, with inadequate psychological support. Outsourced moderation teams, often located in countries like Kenya, the Philippines, and India, work under grueling conditions for wages that don't reflect the psychological toll of the job. Investigative reporting by TIME and others has exposed how little support these workers receive, and how high the rates of PTSD and burnout are among them.

And even culturally informed human reviewers are operating within policy frameworks written by people who may not share their context. If the rulebook says "no nudity" without nuance for traditional dress, a reviewer who personally understands the cultural significance of Bharatanatyam costuming may still have to apply the rule as written.

What Accountability Could Actually Look Like

Creators and digital rights advocates aren't asking for platforms to lower their standards. They're asking for standards that were built with the whole world in mind from the start.

That means investing seriously in multilingual moderation infrastructure — not as an afterthought, but as a core product requirement. It means building appeals processes that are accessible, fast, and available in languages other than English. It means publishing transparency reports that break down removals and restrictions by language and region, so patterns of disproportionate enforcement become visible and undeniable.

It also means actually listening to the creator communities being impacted. Several advocacy organizations — including the Global Network Initiative and the Digital Rights Foundation — have been pushing platforms to create formal consultation mechanisms with civil society groups in the Global South. Progress has been slow.

In the meantime, creators are developing their own workarounds. Code-switching their captions into English to avoid language-based misclassification. Avoiding certain visual elements even when they're culturally significant. Self-censoring in ways that dilute the very cultural authenticity that made their content worth watching in the first place.

The Story Platforms Are Telling About Whose Culture Matters

Here at Maalamalama, we believe every culture has a story to tell. But believing that isn't enough if the infrastructure of modern storytelling is built to favor some stories over others.

Social media platforms have become the new public square — the place where culture lives, spreads, and survives. When those platforms systematically misread, flag, and erase non-Western cultural expression, they're not just making technical errors. They're making a statement about whose art is legitimate, whose traditions are acceptable, and whose voices deserve to be heard.

That's a statement worth pushing back on — loudly, persistently, and in every language the algorithm hasn't figured out how to silence yet.

All Articles

Keep Reading

Same Country, Six Different Accents: Hollywood's Accent Problem Is Bigger Than You Think

Same Country, Six Different Accents: Hollywood's Accent Problem Is Bigger Than You Think

Gloss, Pigment, and Power: How Beauty Creators of Color Are Dismantling Instagram's Unwritten Rules

Gloss, Pigment, and Power: How Beauty Creators of Color Are Dismantling Instagram's Unwritten Rules

Your Meme Doesn't Travel: The Hilarious, Complicated World of Humor That Refuses to Cross Borders

Your Meme Doesn't Travel: The Hilarious, Complicated World of Humor That Refuses to Cross Borders