Scott Alexander quote, Sep 23, 2026
Author, Astral Codex Ten · Sep 23, 2026 · Post, Mysteries Of AI Generalization
This story of misalignment says that LLM alignment makes AIs more aligned, RLVR makes them less aligned, and the exact level of alignment depends on how these two things interact or cancel out. But if this work generalizes, all the bad effects from RLVR get sequestered to RLVR like problems.
Words found on the source page on Sep 25, 2026.