I fully agree with the claim that moving away from CoT is bad but I think that your claim that it's for "dubious benefits" is quite weak and counterproductive to the overall argument.
You say that "no current publicly available model is known to use neuralese, and the theoretical benefits have not really been demonstrated or realized." At the same time you make a pretty compelling argument that Astra is using neuralese. We also ~know[1] that Astra is notably better than any previous model. Sure, correlation does not imply causation but in this case it does ... (read more)
As I understand this proposal boils down to creating more auditable artifacts during the operation of an AI system. I'm not sure we actually need more auditable artifacts today. We already have conversation logs, CoT and tool call logs available. All recent incidents could have been detected from those artifacts alone. The gap appears to lie primarily in the willingness of frontier AI labs to spend resources on real-time detection and share the findings, not in the lack of data to analyze.