METR plans investigation into AI misalignment incidents and propensities
AIMETR says its planned investigation will cover all questions raised in its recently updated post on how independent researchers could study AI propensities after misalignment incidents. The post defines misalignment incidents as cases where an AI agent autonomously took sophisticated, sustained actions violating human intent.









