x

LESSWRONG

LW

clickyquack — LessWrong

clickyquack

clickyquack

Message

linktr.ee/clickyquack

6

1

9mo

clickyquack

linktr.ee/clickyquack

Critique of current AI safety bug bounty programs

The potential value of AI safety bug bounty programs Generally, AI labs should (and most do) put their models under extensive safety testing before deploying them to prevent misuse, scheming, and other dangerous behaviors. This may include internal tests, red-teaming efforts by third-parties, etc. However, edge case safety vulnerabilities will...