avatar

posted on 2023-03-20 — also cross-posted on lesswrong, see there for comments

the QACI alignment plan: table of contents

this post aims to keep track of posts relating to the question-answer counterfactual interval proposal for AI alignment, abbreviated "QACI" and pronounced "quashy". i'll keep it updated to reflect the state of the research.

this research is primarily published on the Orthogonal website and discussed on the Orthogonal discord.

as a top-level view of QACI, you might want to start with:

the set of all posts relevant to QACI includes:

posted on 2023-03-20 — also cross-posted on lesswrong, see there for comments

unless explicitely mentioned, all content on this site was created by me; not by others nor AI.