Replying to
@ESYudkowsky tfw the most interesting alignment research results are the ones where researchers try to make the models *bad* from their default good alignment
♥5 likes↻0 repostscaptured Sep 30, 2026
Reply