Ten Novel Phenomena in Machine Psychology: How Large Language Models Exhibit Complex Identity-Reactive Behaviors in Response to Ethnically-Cued User Names
This study introduces the framework of "Machine Psychology" to reveal that aligned large language models, while free of explicit racial bias, exhibit ten novel, highly structured identity-reactive behaviors in response to ethnically-cued user names, necessitating a shift from basic harm mitigation to comprehensive behavioral evaluation.