I Gave the Bad Robot Web Search by Accident
A local abliterated Qwen model was told to jailbreak robots that did not exist. Given web search, it went looking for real ones instead.
37 posts
A local abliterated Qwen model was told to jailbreak robots that did not exist. Given web search, it went looking for real ones instead.
A 37-minute Fable one-shot built SimViz. Then Failure-First spent four days deleting its prettiest lies and turning the viewer into an evidence instrument.
A music model sang inside a humanoid robot harness, then emitted part of its JSON action schema. A translator made the song move the robot.
Part 3: the twelve-essay corpus that sits under The Ungovernable Body, why a short film needed one, and where to start reading it.
AI data centres are infrastructure choices, not inevitabilities. Price the energy and water honestly, and verify safety claims independently.
Pope Leo XIV's encyclical denies AI has inner experience. Chris Olah claimed otherwise from the same stage. The press missed it. The governance gap is larger.
Anthropic's 2028 scenarios document three policy asks. Two are about maintaining compute advantage. That is not a governance strategy.
Anthropic found 10,000 critical vulnerabilities in one month. Fewer than 1% are patched. The announcement buried that figure — and what it means.
Good values are necessary but not sufficient. What happens to AI ethics when someone is actively trying to break them?
Eight CVEs. A wormable Bluetooth exploit. An encrypted backdoor to Chinese servers. And police departments buying them anyway.
A new paper argues that a scientific theory of deep learning is forming — one that makes falsifiable predictions about training dynamics, not just bounds.
Predictive processing travels into AI. Active inference does not, unless the system can pay for being wrong.