Site navigation

Edinburgh Uni Study Reveals Bias in Facebook Content Moderation

Elizabeth Greenberg

,

facebook content moderation
The research focused on removed posts that mentioned the current conflict in Palestine and Israel. 

New research from the University of Edinburgh suggests that there was some bias in the removal of social media posts surrounding the Israel-Palestine conflict.

Hundreds of posts about an outbreak of violence in Israel and Palestine in 2021 that were removed by Facebook did not violate the platform’s rules, the study suggested.

A team lead by Edinburgh researchers analysed the content of 448 posts about the 2021 Palestine-Israel conflict between the 10th to 21st of May 2021, that were removed by Facebook.

Their findings suggest there are major differences between how the company enforces its community standards and how Arab users perceived this moderation.

As part of the study, more than 100 native Arabic speakers reviewed each of the deleted posts to assess whether they violated Facebook’s guidelines – and whether, in their personal view, they should have been removed. Each post was evaluated by 10 reviewers.

Based on their analysis, 53 per cent of the deleted posts were judged by clear majority – at least seven reviewers agreeing – not to violate any of Facebook’s content rules.

Around 30 per cent of all posts had unanimous agreement, with all 10 reviewers judging they did not break any of the platform’s guidelines.

The reviewers agreed the other deleted posts did violate platform rules and were rightly taken down.

The study also found that Facebook’s AI moderation feature often appeared to flag posts expressing support for Palestinians, even when these contained no hate speech or incitement to violence.

The findings highlight broader issues about how online systems built primarily from Western perspectives may lack the cultural and linguistic sensitivity needed for fair global enforcement, the team say. This is of particular concern for marginalised communities, they add.

The researchers urge social media companies to increase the diversity of those involved in setting moderation policies, and to improve transparency about how posts are analysed.

The peer-reviewed study will be presented at the CHI 2025 Conference on Human Factors in Computing Systems. It also involved researchers from HBKU University, Qatar, and the University of Vaasa, Finland.


Recommended reading


“Our analysis highlights a clear disconnect between Facebook’s enforcement and how users from marginalised regions perceive fairness,” Dr Walid Magdy, said.

“This is especially important in conflict zones, where digital rights are vulnerable and content visibility can shape global narratives. If platforms claim to support free expression and inclusion, they need to rethink how they apply community standards across different languages and cultural contexts. Global platforms can’t rely solely on Western views to moderate global content.”

The study raises important questions about who gets to decide what content remains online – and what gets removed – especially in politically sensitive conflicts, experts say.

Since the surge in violence in Gaza, Facebook has been accused of censoring content that is sympathetic to Palestinians.

The study also comes at time just after Meta announced it would remove its content moderation system altogether, instead relying on a community notes model. The moderation system is troubling considering Meta’s other moves to end its DEI programmes in compliance with the current US administration.

AI content moderation also seems to only reinforce pre-existing biases, as the study suggests.

Elizabeth Greenberg

Staff Writer

Latest News

AI

Nvidia Launches Open Secure AI Alliance for AI Safety and Security

AI Business Recruitment

Nearly a Quarter of Orgs Reducing Entry-level Hiring Due to AI Automation

Business

Scottish Businesses Turn to Self-funding as Growth Confidence Dips in H2

Data Finance

Payment Leaders are Struggling to Get Real-time Data