LLM Safety
Predicting Belief Shifts in Manipulative AI Chats
New research tackles the 'hidden puppet master' problem: predicting how manipulative LLM dialogues change human beliefs. The work offers a framework for detecting persuasion and covert influence in conversational AI systems.