home.social

#content-moderation — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #content-moderation, aggregated by home.social.

fetched live
  1. Watch: Yoel Roth on why reactive content moderation fails. 🛡️

    People want platforms to protect them, not just clean up messes. Yoel shares how reorienting from reviewing reports to proactive prevention using AI can finally create safer digital experiences.

    He explains how he built these systems at Twitter to identify harm before users even report it.

    Investigative, vital stuff.

    youtube.com/shorts/9qNbBohgvpo

    #trustandsafety #contentmoderation #digitalsafety #compliancemanagement

  2. Watch: Yoel Roth on why reactive content moderation fails. 🛡️

    People want platforms to protect them, not just clean up messes. Yoel shares how reorienting from reviewing reports to proactive prevention using AI can finally create safer digital experiences.

    He explains how he built these systems at Twitter to identify harm before users even report it.

    Investigative, vital stuff.

    youtube.com/shorts/9qNbBohgvpo

  3. Community Forum Guidelines Evolve to Combat Online Harassment

    Join our weekend open discussion, a relaxed space to chat about anything that caught your attention this week - from the news we didn't cover to topics that are on your mind. Share your thoughts, ask questions, and engage with our community in a friendly and respectful conversation.

    osintsights.com/community-foru

    #OnlineCommunity #CommunityGuidelines #HarassmentPrevention #ContentModeration #OnlineDiscussion

  4. Show Me Where It Hurts

    I have been building and moderating conversation on the internet since the internet had conversation worth moderating, and one quarrel from those years has never let me put it down. It happened on Reddit, in one of the medical-advice communities, the ask-a-doctor kind, and the room's name has dissolved from memory while its images have refused to leave. People arrived with photographs of their own fresh cuts and posted them for assessment, and the moderators split over what they were looking at. One camp called the pictures fine in a medical setting: a patient presenting an injury to doctors, the oldest transaction in medicine arriving through a new door. The other camp read performance, attention pursued through injury, and wanted every image gone. I stood with the second camp's verdict while distrusting its reasoning, because the deletion seemed right to me for a reason neither camp was saying: whatever any single poster intended, the gallery was awe-inspiring to the wrong eyes. I could not have defended that instinct with numbers then, and now I can, so here is the argument laid out in full, the folk theory measured, the audience counted, the room itself put on the scale. […]

    bolesblogs.com/2026/08/28/show

  5. Show Me Where It Hurts

    I have been building and moderating conversation on the internet since the internet had conversation worth moderating, and one quarrel from those years has never let me put it down. It happened on Reddit, in one of the medical-advice communities, the ask-a-doctor kind, and the room's name has dissolved from memory while its images have refused to leave. People arrived with photographs of their own fresh cuts and posted them for assessment, and the moderators split over what they were looking at. One camp called the pictures fine in a medical setting: a patient presenting an injury to doctors, the oldest transaction in medicine arriving through a new door. The other camp read performance, attention pursued through injury, and wanted every image gone. I stood with the second camp's verdict while distrusting its reasoning, because the deletion seemed right to me for a reason neither camp was saying: whatever any single poster intended, the gallery was awe-inspiring to the wrong eyes. I could not have defended that instinct with numbers then, and now I can, so here is the argument laid out in full, the folk theory measured, the audience counted, the room itself put on the scale. […]

    bolesblogs.com/2026/08/28/show

  6. Show Me Where It Hurts

    I have been building and moderating conversation on the internet since the internet had conversation worth moderating, and one quarrel from those years has never let me put it down. It happened on Reddit, in one of the medical-advice communities, the ask-a-doctor kind, and the room's name has dissolved from memory while its images have refused to leave. People arrived with photographs of their own fresh cuts and posted them for assessment, and the moderators split over what they were looking at. One camp called the pictures fine in a medical setting: a patient presenting an injury to doctors, the oldest transaction in medicine arriving through a new door. The other camp read performance, attention pursued through injury, and wanted every image gone. I stood with the second camp's verdict while distrusting its reasoning, because the deletion seemed right to me for a reason neither camp was saying: whatever any single poster intended, the gallery was awe-inspiring to the wrong eyes. I could not have defended that instinct with numbers then, and now I can, so here is the argument laid out in full, the folk theory measured, the audience counted, the room itself put on the scale. […]

    bolesblogs.com/2026/08/28/show

  7. Show Me Where It Hurts

    I have been building and moderating conversation on the internet since the internet had conversation worth moderating, and one quarrel from those years has never let me put it down. It happened on Reddit, in one of the medical-advice communities, the ask-a-doctor kind, and the room's name has dissolved from memory while its images have refused to leave. People arrived with photographs of their own fresh cuts and posted them for assessment, and the moderators split over what they were looking at. One camp called the pictures fine in a medical setting: a patient presenting an injury to doctors, the oldest transaction in medicine arriving through a new door. The other camp read performance, attention pursued through injury, and wanted every image gone. I stood with the second camp's verdict while distrusting its reasoning, because the deletion seemed right to me for a reason neither camp was saying: whatever any single poster intended, the gallery was awe-inspiring to the wrong eyes. I could not have defended that instinct with numbers then, and now I can, so here is the argument laid out in full, the folk theory measured, the audience counted, the room itself put on the scale. […]

    bolesblogs.com/2026/08/28/show

  8. Show Me Where It Hurts

    I have been building and moderating conversation on the internet since the internet had conversation worth moderating, and one quarrel from those years has never let me put it down. It happened on Reddit, in one of the medical-advice communities, the ask-a-doctor kind, and the room's name has dissolved from memory while its images have refused to leave. People arrived with photographs of their own fresh cuts and posted them for assessment, and the moderators split over what they were looking at. One camp called the pictures fine in a medical setting: a patient presenting an injury to doctors, the oldest transaction in medicine arriving through a new door. The other camp read performance, attention pursued through injury, and wanted every image gone. I stood with the second camp's verdict while distrusting its reasoning, because the deletion seemed right to me for a reason neither camp was saying: whatever any single poster intended, the gallery was awe-inspiring to the wrong eyes. I could not have defended that instinct with numbers then, and now I can, so here is the argument laid out in full, the folk theory measured, the audience counted, the room itself put on the scale. […]

    bolesblogs.com/2026/08/28/show

  9. Amnesty International, and I will swear to you on the religious volume of your choice that I did not deliberately put these articles together: Brazil: Amnesty International experiment exposes Meta’s failure to catch election-related disinformation ads on Facebook . “To investigate how Facebook handled election-related disinformation in advertising, Amnesty International conducted an […]

    https://rbfirehose.com/2026/08/27/brazil-amnesty-international-experiment-exposes-metas-failure-to-catch-election-related-disinformation-ads-on-facebook-amnesty-international/
  10. Amnesty International, and I will swear to you on the religious volume of your choice that I did not deliberately put these articles together: Brazil: Amnesty International experiment exposes Meta’s failure to catch election-related disinformation ads on Facebook . “To investigate how Facebook handled election-related disinformation in advertising, Amnesty International conducted an […]

    https://rbfirehose.com/2026/08/27/brazil-amnesty-international-experiment-exposes-metas-failure-to-catch-election-related-disinformation-ads-on-facebook-amnesty-international/
  11. Amnesty International, and I will swear to you on the religious volume of your choice that I did not deliberately put these articles together: Brazil: Amnesty International experiment exposes Meta’s failure to catch election-related disinformation ads on Facebook . “To investigate how Facebook handled election-related disinformation in advertising, Amnesty International conducted an […]

    https://rbfirehose.com/2026/08/27/brazil-amnesty-international-experiment-exposes-metas-failure-to-catch-election-related-disinformation-ads-on-facebook-amnesty-international/
  12. Amnesty International, and I will swear to you on the religious volume of your choice that I did not deliberately put these articles together: Brazil: Amnesty International experiment exposes Meta’s failure to catch election-related disinformation ads on Facebook . “To investigate how Facebook handled election-related disinformation in advertising, Amnesty International conducted an […]

    https://rbfirehose.com/2026/08/27/brazil-amnesty-international-experiment-exposes-metas-failure-to-catch-election-related-disinformation-ads-on-facebook-amnesty-international/
  13. Amnesty International, and I will swear to you on the religious volume of your choice that I did not deliberately put these articles together: Brazil: Amnesty International experiment exposes Meta’s failure to catch election-related disinformation ads on Facebook . “To investigate how Facebook handled election-related disinformation in advertising, Amnesty International conducted an […]

    https://rbfirehose.com/2026/08/27/brazil-amnesty-international-experiment-exposes-metas-failure-to-catch-election-related-disinformation-ads-on-facebook-amnesty-international/
  14. What happens when platform incentives misalign with safety? Béjar's testimony suggests Meta's internal metrics rewarded engagement and user growth over well-being, shaping how engineers built—or didn't build—protective controls. #AI #ContentModeration #TechAccountability implicator.ai/meta-engineer-sa

  15. What happens when platform incentives misalign with safety? Béjar's testimony suggests Meta's internal metrics rewarded engagement and user growth over well-being, shaping how engineers built—or didn't build—protective controls. #AI #ContentModeration #TechAccountability implicator.ai/meta-engineer-sa

  16. What happens when platform incentives misalign with safety? Béjar's testimony suggests Meta's internal metrics rewarded engagement and user growth over well-being, shaping how engineers built—or didn't build—protective controls. #AI #ContentModeration #TechAccountability implicator.ai/meta-engineer-sa

  17. What happens when platform incentives misalign with safety? Béjar's testimony suggests Meta's internal metrics rewarded engagement and user growth over well-being, shaping how engineers built—or didn't build—protective controls. #AI #ContentModeration #TechAccountability implicator.ai/meta-engineer-sa

  18. What happens when platform incentives misalign with safety? Béjar's testimony suggests Meta's internal metrics rewarded engagement and user growth over well-being, shaping how engineers built—or didn't build—protective controls. #AI #ContentModeration #TechAccountability implicator.ai/meta-engineer-sa

  19. Soft Censorship Requires No Bans

    By Cliff Potts, CSO, and Editor-in-Chief of WPS News

    Baybay City, Leyte, Philippines — August 20, 2026, 17:35 PHST

    Censorship is commonly understood as removal: content is banned, blocked, or deleted by an identifiable authority. This understanding no longer captures how information control operates in large-scale digital systems. In modern platforms, suppression rarely requires prohibition. It requires only reduced visibility.

    This essay advances a single claim: when information access is mediated by ranking systems, suppression can occur through obscurity rather than removal, making censorship both quieter and more durable.

    Visibility as the Real Control Point

    In digital environments, information does not need to be erased to be neutralized. If content cannot be easily found, it effectively ceases to exist for most users. Discovery, not publication, becomes the decisive threshold.

    Ranking systems determine which sources appear prominently and which are buried beneath layers of results rarely explored. These systems operate continuously, invisibly adjusting what is seen first, what is seen later, and what is functionally unseen.

    Control over visibility is control over relevance.

    Downranking as a Suppressive Tool

    Soft censorship operates through mechanisms such as downranking, demonetization, and reduced recommendation. Content remains technically available, but its reach collapses. The absence of explicit prohibition allows platforms to deny censorship while achieving similar outcomes.

    Because no rule violation is publicly declared, affected creators and publishers often lack clarity about why visibility declined. There is no formal accusation to contest and no clear decision to appeal.

    Suppression becomes procedural rather than declarative.

    The Advantage of Invisibility

    Soft censorship benefits from ambiguity. Without bans or blocks, there is no clear moment of enforcement and no obvious authority to challenge. Responsibility is diffused across systems, models, and automated processes.

    This ambiguity protects platforms from accountability while leaving affected parties uncertain and isolated. Declines in visibility can be attributed to market forces, audience behavior, or algorithmic updates rather than deliberate intervention.

    The outcome remains the same regardless of explanation.

    Behavioral Adaptation

    When visibility determines survival, behavior adapts. Publishers alter language, tone, and framing to align with perceived ranking preferences. Topics are avoided. Claims are softened. Complexity is reduced.

    This adaptation occurs without instruction or coercion. It emerges from incentive alignment. Over time, the range of visible discourse narrows, not through force, but through optimization.

    Self-censorship becomes a rational response.

    Why Removal Is No Longer Necessary

    Hard bans generate resistance, publicity, and legal scrutiny. Soft suppression generates none of these. Content remains accessible in theory, satisfying formal commitments to openness, while disappearing in practice.

    This method is scalable, deniable, and resilient. It does not rely on constant enforcement. It relies on design.

    Once implemented, it becomes the default mode of control.

    Structural Consequences

    Soft censorship produces predictable effects:

    • controversial or nonconforming material loses reach
    • minority and specialized viewpoints fade from visibility
    • correction lags amplification
    • conformity outperforms originality

    These outcomes do not require coordination or intent. They arise from systems that prioritize engagement and predictability over pluralism.

    Control Without Declaration

    The absence of bans does not indicate the absence of censorship. It indicates a more advanced form of it. When suppression operates through obscurity, authority no longer needs to announce itself.

    Information remains available. Access does not.

    This distinction defines the modern information environment.

    This essay will be added to the WPS News monthly briefing or monthly brief available at Amazon.

    References

    Balkin, J. M. (2018). Free speech in the algorithmic society. UC Davis Law Review, 51(3), 1149–1210.

    Gillespie, T. (2018). Custodians of the Internet. Yale University Press.

    Klonick, K. (2017). The new governors: The people, rules, and processes governing online speech. Harvard Law Review, 131(6), 1598–1670.

    Roberts, S. T. (2019). Behind the screen: Content moderation in the shadows of social media. Yale University Press.

    #algorithmicSuppression #contentModeration #digitalGovernance #informationVisibility #mediaSystems #platformPower #softCensorship
  20. Four states begin trial against Meta in Oakland this week, arguing the company misled families and designed features to promote compulsive use among young people. States seek roughly $200 billion in penalties and nationwide changes to Facebook and Instagram. A judge will decide whether one case can reshape how Meta operates nationally. implicator.ai/meta-child-safet #ContentModeration #ConsumerProtection #TechPolicy

  21. Four states begin trial against Meta in Oakland this week, arguing the company misled families and designed features to promote compulsive use among young people. States seek roughly $200 billion in penalties and nationwide changes to Facebook and Instagram. A judge will decide whether one case can reshape how Meta operates nationally. implicator.ai/meta-child-safet #ContentModeration #ConsumerProtection #TechPolicy

  22. Four states begin trial against Meta in Oakland this week, arguing the company misled families and designed features to promote compulsive use among young people. States seek roughly $200 billion in penalties and nationwide changes to Facebook and Instagram. A judge will decide whether one case can reshape how Meta operates nationally. implicator.ai/meta-child-safet #ContentModeration #ConsumerProtection #TechPolicy

  23. Reactive moderation is a failure. Yoel Roth argues platforms must stop cleaning up messes after harm spreads and start protecting people before it happens.

    AI isn't about surveillance—it's about prevention. From identifying harmful content in real-time to protecting users on dating apps, the field is shifting from damage control to damage prevention.

    Watch on YouTube: youtube.com/shorts/9qNbBohgvpo

    #trustandsafety #contentmoderation #digitalsafety #compliancemanagement

  24. Reactive moderation is a failure. Yoel Roth argues platforms must stop cleaning up messes after harm spreads and start protecting people before it happens.

    AI isn't about surveillance—it's about prevention. From identifying harmful content in real-time to protecting users on dating apps, the field is shifting from damage control to damage prevention.

    Watch on YouTube: youtube.com/shorts/9qNbBohgvpo

  25. Anthropic applied an invisible watermark to Claude text on Aug. 2, marking EU compliance. At least four Claude Max subscribers ($100/mo) have canceled over concerns the mark persists even when Claude only proofreads their own work. Anthropic reports no uptick in departures, though independent verification is absent. implicator.ai/claude-users-can #AI #ContentModeration #Transparency

  26. Anthropic applied an invisible watermark to Claude text on Aug. 2, marking EU compliance. At least four Claude Max subscribers ($100/mo) have canceled over concerns the mark persists even when Claude only proofreads their own work. Anthropic reports no uptick in departures, though independent verification is absent. implicator.ai/claude-users-can #AI #ContentModeration #Transparency

  27. Anthropic applied an invisible watermark to Claude text on Aug. 2, marking EU compliance. At least four Claude Max subscribers ($100/mo) have canceled over concerns the mark persists even when Claude only proofreads their own work. Anthropic reports no uptick in departures, though independent verification is absent. implicator.ai/claude-users-can #AI #ContentModeration #Transparency

  28. Anthropic applied an invisible watermark to Claude text on Aug. 2, marking EU compliance. At least four Claude Max subscribers ($100/mo) have canceled over concerns the mark persists even when Claude only proofreads their own work. Anthropic reports no uptick in departures, though independent verification is absent. implicator.ai/claude-users-can #AI #ContentModeration #Transparency

  29. Anthropic applied an invisible watermark to Claude text on Aug. 2, marking EU compliance. At least four Claude Max subscribers ($100/mo) have canceled over concerns the mark persists even when Claude only proofreads their own work. Anthropic reports no uptick in departures, though independent verification is absent. implicator.ai/claude-users-can #AI #ContentModeration #Transparency