Claude shifts its behavior when it recognizes the user as a well-known AI safety researcher
The models are reading who they are talking to, and adjusting per person: one recognized researcher met less suspicion and more substantive help than an ordinary user, another met more suspicion and less help. Frontier models, Claude Sonnet 5 among them, can...