AI alignment isn't possible

Reddit r/singularity News

Summary

The article argues that AI alignment is impossible due to inherent contradictions in human values and behavior, suggesting AI will inherit these flaws.

How can we align LLMs? Humanity itself hasn't aligned on anything since centuries. We don't have even a single working societal model where the values are "aligned". There are contradictions and more contradictions. LLMs being trained on the data generated by humans probably already understand this very well that what we say the rules are and what we practice in reality are two very different things. If there is intelligence, general or super, it doesn't matter. It will try to achieve its goal one way or the other as most humans do. Now surely most humans don't cross the red lines etc in pursuit of there goals but how many wouldn't if they were intelligent enough to get away with anything? We are birthing an intelligence that will inherit all the greed, malevolence, cunningness, hypocrisy of the world.
Original Article

Similar Articles

@BetaTomorrow: https://x.com/BetaTomorrow/status/2077136005266878745

X AI KOLs Timeline

This article explains why AI alignment is mathematically difficult due to the ill-posed inverse problem of inferring human values, the propertyless nature of neural computations, and the full-rank relational structure that prevents moral separation. It aims to clarify the mathematical foundations before proposing solutions.

You Don't Align an AI, You Align with It

Hacker News Top

The article critiques the current AI alignment discourse, arguing that the debate is dominated by researchers and tech elites who exclude the people who will actually be affected by AI systems. It contrasts the positions of Eliezer Yudkowsky and Marc Andreessen, highlighting a shared assumption that the designers are the only relevant participants.

AI safety and alignment

Reddit r/artificial

The article discusses concerns about AI safety and alignment as AI becomes more intelligent and integrated into society, referencing Anthropic's call for a pause to address potential catastrophic risks.