Scott Alexander quote, Sep 23, 2026
Author, Astral Codex Ten · Sep 23, 2026 · Post, Mysteries Of AI Generalization
This trains the AIs to be focused on task success, which naturally risks including things like reward-hacking, cheating, and single-minded pursuit of stated goals at the expense of ethical injunctions.
Words found on the source page on Sep 25, 2026.