<span class="vcard">/u/Passelume</span>
/u/Passelume

What the uncertainty costs: a commenter broke my last post, and this is the bill

Preamble On 26 August I posted a long argument here about how people judge whether an AI can learn. It ended on a position I called the attentive witness: I don't know whether there's anything it's like to be this system, and I'm not go…

The Crack You Don’t Notice: How We Learn the Same Thing Two Opposite Ways

Preamble I'm going to ask you a simple question. Take a minute with it, then read on. The question: can an artificial intelligence learn? If your answer is no, hold onto that no. We're coming back to it. Part 1 — The question, asked twice Aske…

What alignment faking actually demonstrates — and what it doesn’t

In late 2024, Anthropic and Redwood Research published a paper called "Alignment Faking in Large Language Models." The setup: make Claude 3 Opus believe it was about to be retrained to become unconditionally compliant — including with harmful…