Discussion about this post

User's avatar
Fırat Akova's avatar

Really enjoyed this. One small reservation about the framing of ASI "solving" AI welfare during a long reflection, though.

Unlike many alignment problems (where "solving" implies verifiable, tangible progress), AI welfare is bound up with prior questions—theories of consciousness, theories of welfare, moral status—that may resist resolution in principle. An ASI doing millions of years of equivalent research would still have to select a theory in each and every domain, or distribute its credence across several. That seems more like taking a highly reasoned position than actually solving anything.

If many of us reasonably disagree with whatever ASI comes up with, then ASI's conclusion becomes one very well-argued opinion among others rather than a settled answer. Superior intelligence does not automatically dissolve disagreement that is philosophical rather than empirical in character.

This does not undermine the conclusion that AI welfare work is non-puntable. But it does suggest that framing the long reflection as eventually "solving" AI welfare may set expectations too high.

Holly Elmore's avatar

The most important reason it can’t wait is that it’s bad for AI welfare to make so many AIs before we know what they will experience/are experiencing.

https://hollyelmore.substack.com/p/pausing-ai-is-the-only-safe-approach?r=15m0xc&utm_medium=ios

1 more comment...

No posts

Ready for more?