Critique of current AI safety bug bounty programs
The potential value of AI safety bug bounty programs Generally, AI labs should (and most do) put their models under extensive safety testing before deploying them to prevent misuse, scheming, and other dangerous behaviors. This may include internal tests, red-teaming efforts by third-parties, etc. However, edge case safety vulnerabilities will...
Jun 17