Switch language한국어
Back to the list

Claude Leak Shows Anthropic Tracks Users’ Vulgar Language and Flags Them as Negative

TL;DR AI

Key summary

2 min read
  1. Anthropic's Claude Code source was leaked and widely shared this week.

  2. Developers recreated the code in a repo called 'Claw Code', which was forked nearly 100,000 times.

  3. The leaked code shows regex that detects vulgar phrases and logs them as is_negative: true in analytics.

  4. Boris Cherny said the vulgar-language signal is used on a dashboard to monitor user experience.

  5. Cherny attributed the leak to a missed manual deployment step and said the company is adding safeguards.

Read the original