@sayashk
What is the role of model alignment for AI safety? - Model alignment is effective against accidental harms, not intentional ones - Important questions about AI safety can’t be asked and answered at the levels of models. In other words, *AI safety is not a model property* https://t.co/H6SyuWF4CJ