Empirical evidence that people can build highly predictive mental models of even complex AIs, if and only if the algorithm fulfills three criteria:
Is the only way we can create algorithms that people understand to make them trivially simple? We argue, no. People can predict the behavior of algorithms that are arbitrarily complex, if and only if they are available, compact and aligned. arxiv.org/abs/2601.18966