Are LLMs conscious?[1] We don't know. To gain insight into this extremely important question as well as many others, we propose training an LLM on a corpus without any mentions of consciousness and similar ideas. We the authors want to actually do this, and we want to hear your thoughts...
The first artificial intelligence was booted up around 4000BC in southern Iraq. It seems to have begun as something like a bank, a temple pooling grain against famine. As that AI evolved, it formed the world's first city around itself: Uruk. Over the next thousand years it became a religion,...
The "Paris 1937 World’s Fair" was a dick measuring contest. At the time, the world was on the verge of the worst war in history. The fair was an opportunity for powers to flex and intimidate each other. Who has more industrial might, more sophisticated engineering and better science? How...
Over the last 200 years, society became steadily more liberal. Serfdom and slavery were largely abolished, democracy exploded and most women and children are free from violence. Redistribution is widely practiced and the median person is far richer, healthier and freer than ever before. Some neighborhoods in San Francisco already...
The time it takes an AI or a Human+AI team (a "cyborg") to complete a task is a key aspect of what we care about when we talk about capabilities. The relationship between AI capabilities and cyborg capabilities is very useful for forecasting AI timelines. We simply don’t have good...
ProgramBench is a new coding benchmark that all frontier models spectacularly fail. We’ve been on a quest for “hard benchmarks” for a while so it’s refreshing to see a benchmark where top models do badly. Unfortunately, ProgramBench has one big problem: it’s impossible! What is ProgramBench? ProgramBench tests if a...