Skip to content
Safety

Dreams of alignment in a world without politics.

LessWrongAnalysis or commentary··Updated 7h ago
AI brief

A LessWrong post reflects on alignment research and mentions an idea about models writing adversarial fiction that is then used for training.

Written by AI from LessWrong's published text. Read the original for full details.

Source

Interpretation or community commentary rather than straight reporting.

Read original story ↗