Skip to content

Scott Alexander quote, Sep 23, 2026

Author, Astral Codex Ten · Sep 23, 2026 · Post, Mysteries Of AI Generalization

This trains the AIs to be focused on task success, which naturally risks including things like reward-hacking, cheating, and single-minded pursuit of stated goals at the expense of ethical injunctions.

Primary source

Words found on the source page on Sep 25, 2026.

More from Scott Alexander

About this quote
Said by
Scott Alexander, Author, Astral Codex Ten
Where
Mysteries Of AI Generalization
Topics
Safety, Research
Source
astralcodexten.com/p/mysteries-of-ai-generalization

A quote is one moment, not the whole argument. Read it in context at the source.

Sources: the primary source linked on each quote, checked weekly. A quote is one moment; read the source. Logos via logo.dev; trademarks belong to their owners.

New quotes by email

Fridays, only in weeks with new on-the-record quotes from data and AI leaders.

Double opt-in. Unsubscribe any time.