Background For the sake of argument, I'll assume that the summaries of the HF incident are broadly accurate. I'll assume the attack is over. And I'll also assume that OpenAI threw a kill-switch, and that this is why the attack ended. Any of these assumptions might later prove wrong, but...
AI Assisted Work: A Missive For the Managerial Class The Smart Employee “It doesn't make sense to hire smart people and then tell them what to do, We hire smart people so they can tell us what to do.” - Steve Jobs This famous quote is bound to provoke some...
A Stratified Meme is a meme that communicates different ideas to different people, according to their ability and willingness to hear the message. A Stratified meme has a specific structure: 1. There are higher and lower readings that are related but different. This is called multi-level messaging, or strategic ambiguity....
TLDR: I describe a takeover path by an AI [1]with a deep understanding of human nature and a long planning horizon that, for strategic reasons, chooses not to directly pursue physical power. Instead, the AI "backdoors" alignment by building a broad base of human support, hijacking institutions and power structures,...
This is the first in a series of posts on the question: > "Can we extract meaningful information or interesting behavior from gradients on 'input embedding space'?" I'm defining 'input embedding space' as the token embeddings prior to positional encoding. The basic procedure for obtaining input space gradients is as...