The Game Theory of AI Safety Talk
Why what labs say about safety is a strategic signal, not a statement of values — and what that means for regulation.
Strategic dynamics, signalling, mechanism design
Why what labs say about safety is a strategic signal, not a statement of values — and what that means for regulation.
In conventional security, hardening a system makes it harder to attack. You patch vulnerabilities, reduce attack surface, and defence moves in lockstep with robustness. AI alignment breaks this assumption.