PSA: We can do better
Key takeaway
tl;dr: people should understand and think hard about the problems they work on.We’ve observed that those who work in AI safety (ourselves included) often rely on concerning heuristics when choosing what to work on.
Why it's on PDOOM
PDOOM selected this story for its alignment & control signals: Alignment, X-Risk.
AlignmentX-RiskFrom LessWrong AI
tl;dr: people should understand and think hard about the problems they work on.We’ve observed that those who work in AI safety (ourselves included) often rely on concerning heuristics when choosing what to work on. Running a conference is probably good, doing pragmatic alignment research might be good, and as long as such objectives don’t breach our internal models of what could contribute to reducing x-risk, these things are “what should be done”. But using such vibesy thought processes don’t a
By hersheys