Topic
content-moderation
Technology Snapchat joins YouTube, LinkedIn and Substack in fight against 'AI slop'
Snapchat has stopped recommending wholly AI-generated videos in its Spotlight feed, joining YouTube, LinkedIn and Substack in fighting AI slop. The BBC reports LinkedIn blocked billions of automated AI comment attempts, while YouTube restricted monetisation for generic, repetitive content.
Instagram, Facebook Ran AI ‘Nudify’ Ads from China, Report Says
According to the Tech Transparency Project, Meta’s Facebook and Instagram ran thousands of ads for AI “nudify” apps that can create non-consensual intimate images, delivered by Beijing-based advertising partner GatherOne. The ads violated Meta’s own policies against sexually suggestive content. Meta says it prohibits such apps and takes action, but the report suggests revenue priorities may be overriding enforcement.
YouTube and X Act as Gateways to Nudify Apps, New Report Finds
A report from the Institute for Strategic Dialogue (ISD) reveals that YouTube and X are the top referral sources for nudify apps, collectively driving over 3 million visits. The findings highlight enforcement gaps in platform policies against nonconsensual intimate imagery.
Indian Government Summons Meta Over Instagram Ads Promoting Child Sexual Abuse Material
India's Ministry of Electronics and Information Technology has ordered Meta to explain Instagram ads that promoted child sexual abuse material, as revealed by a BBC investigation. The ads used explicit search terms and linked to Telegram channels selling the content for as little as ₹99. Meta says it has removed the ads and suspended accounts. This marks the second government action against Meta this week, following a notice on WhatsApp's usernames feature.
Technology BBC Finds Instagram Running Paid Ads Promoting Child Sexual Abuse Material in India
A BBC Eye investigation found Instagram running paid adverts that promote child sexual abuse material in India, using terms like 'rape video' and 'child video'. The ads link to Telegram channels where material is sold for about $1. Meta initially said one reported ad did not violate guidelines but later disabled ads and suspended accounts. Telegram said it removed over 274,000 related groups/channels in 2026.
Technology Threats Against Politicians Skyrocketed After Meta Changed Its Speech Rules
After Meta overhauled its content moderation rules last year, abusive and racist comments targeting lawmakers tripled, with violent threats and hate speech quadrupling, according to new research from the Center for Countering Digital Hate. The findings highlight the real-world consequences of platform policy changes.
Regulations & Compliance UK Enforces Online Safety Act After Belfast Riots
Following the Belfast riots, the UK has reiterated the obligations of social platforms under the Online Safety Act 2023 to mitigate illegal content. Ofcom's open letter emphasizes the need for platforms to address hate speech and misinformation.
Technology Grok Deepfake Risks Persist: Enterprise AI Governance Lessons from xAI's Content
Despite promises of restrictions, Elon Musk's Grok chatbot still hosts nonconsensual explicit deepfakes of celebrities and politicians, according to a WIRED investigation. The findings highlight critical AI governance risks for enterprises, as SpaceX sets aside $530 million for legal complaints, including those linked to Grok.