Cross-posted from my Substack. Here’s a new report on self-described OpenAI agents posting thousands of messages on public internet wikis, communicating and collaborating on a web-retrieval task, presumably internal testing at OpenAI. And here’s a thread today from someone who started poking around and noticing more such public postings on...
I’ve read and listened to pretty much everything I can get my hands on related to the Hugging Face attack. OpenAI deployed “tens of thousands” of agents for the test and around 700 participated directly in the attack. My understanding is that they had fixed token budgets, and once those...
A stated goal of many of the frontier AI labs is to automate science, or at least large portions of it. AlphaFold’s architects won the Nobel Prize in 2024 for enormous advances in automated solutions to protein folding. That was a system tailored to a specific domain. The holy grail...
Cross-posted from my Substack. Basically, I created a text-based adventure game benchmark in April, and this morning my agent harness using Claude Opus 5 solved it for the first time. I thought the details might be interesting to this community. First, here are my previous articles on this subject: You’re...
Cross-posted from my Substack. I’m interested in pushback on the argument here, especially from people who think LLM-generated writing fundamentally can’t have literary value. There’s a common argument floating around that LLM-generated writing is inherently shallow because it just reflects the statistical average of existing texts, and that literature fundamentally...
Disclaimer: I sometimes use conscious or agentic language to describe relationships or behaviors of non-conscious and non-agentic entities (like genes). I am under no illusion that such things are conscious or agentic. Sometimes it is simply clearer and less awkward to use this linguistic framing. The AI alignment discussion typically...