Between June and September 2024, we ran the third iteration of the PIBBSS Summer Research Fellowship. Here are our reflections on how the program went and what we learned. Apply for the 2025 program here! TLDR: The 2024 fellowship demonstrated continued success in attracting and developing senior academic talent while...
Tl;dr We are pleased to invite you to the second PIBBSS Symposium, where the fellows from the ‘24 fellowship program present their work. The symposium is taking place online, over several days in the week of September 9th. * Check out the full program, including brief descriptions for each talk....
-1. Motivation As AI systems become increasingly powerful, the chance of them developing dangerous features becomes increasingly likely. A key concern of AI alignment researchers is that there are dangerous features, such as power-seeking or situational awareness, which are convergent.[1] This would mean that these capabilities would likely arise regardless...
When we're trying to do AI alignment, we're often studying systems which don't yet exist. This is a pretty weird epistemic activity, and seems really hard to get right. This post offers one frame for thinking about what we're actually doing when we're thinking about AI alignment: using parts of...