AI alignment
- ZM
2 Sept 2026
Zvi Mowshowitz quoted
I have long said that even Anthropic is not prioritizing safety, even to the extent that doing so would maximize their medium term (e.g. 3-12 months) business interests.
- ZM
2 Sept 2026
Zvi Mowshowitz quoted
Every attempt, even an unsuccessful one, is an alignment failure.
- ZM
2 Sept 2026
Zvi Mowshowitz quoted
I worry a lot about reliances on continuity failing at exactly the most dangerous time.
- 1 day earlier
- NS
1 Sept 2026
Noah Smith quoted
AI alignment will forever be balancing the tradeoff between paperclip-maximizing and disempowerment, because these two rival concepts of alignment are fundamentally incompatible.
- ZM
1 Sept 2026
Zvi Mowshowitz quoted
HoldsOpenAI's approach to fixing its AI alignment problems is fatally flawed and misdirected.
My worry continues to be that their fundamental approach is fatally flawed, and they are not focusing on the right things.
↗HuggingFace Attack Postmortem: Civilizations, Reactions and Next Actionsthezvi.substack.com
- DB
1 Sept 2026
Dean W. Ball quoted
But alignment is no solution: it is an unsolved scientific and technical problem whose solutions—to the extent that we have them—cannot simply be imposed on every AI company operating on Earth. You should expect for highly capable, poorly aligned, self-sovereign agents to exist alongside you in the world.
- 2 weeks earlier
- SA
18 Aug 2026
Sam Altman quoted
HoldsConfidence in safety, not capability, will increasingly set the pace of AI progress.
We care very deeply about AI safety. We believe the entire field will have to coordinate on shared safety standards, but will act unilaterally in the meantime. We expect confidence in safety to increasingly set the pace of AI progress.