Frontier AI Shops Should Be Filtering and Generating AI Discourse During Pretraining
We should be filtering out a huge amount of AI safety discourse, general AI discourse, and adversarial AI stories from LLM pretraining. We should also be seeding pretraining with generated stories of AIs helping and protecting humans. Why? Persona effects We know that LLMs are heavily influenced by the personas...