AI Briefing
KO

Child Safety Principles

·2024.12.20 03:52

Key point

Anthropic implements robust safety measures based on 'Safety by Design' principles to prevent child sexual exploitation and sexual harm through generative AI.

Details

Anthropic is introducing robust safety measures to protect children throughout the development, deployment, and maintenance of generative AI technology. This initiative is being carried out in partnership with Thorn and All Tech Is Human, non-profit organizations dedicated to preventing child sexual abuse.

Anthropic adheres to Safety by Design principles and implements the following phased measures to prevent AI-generated child sexual exploitation material (AIG-CSAM) and other forms of sexual harm.

Develop phase:

  • Excluding data at risk of CSAM and CSEM from training, and immediately detecting, removing, and reporting such data upon discovery
  • Conducting structured and scalable Red Teaming testing to address AIG-CSAM
  • Establishing strict policies to prevent customer misuse of models

Deploy phase:

  • Providing detection of harmful content within input and output data and user reporting functionality
  • Monitoring for misuse during early-stage rollouts through phased deployment, and including a child safety section in model cards

Maintain phase:

  • Systematic reporting to NCMEC (National Center for Missing & Exploited Children) and investing in tools to prevent AI-generated manipulation
  • Continuously identifying how platforms and models are being abused using OSINT (open-source intelligence)

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.