arXiv Paper Examines How Users Mistreat Conversational AI Systems
A new arXiv preprint studies how users direct hostility, coercion, and adversarial pressure at conversational AI models, an area the authors say is often overlooked in favor of research on model-generated harms. The paper argues that understanding when and why such mistreatment happens is needed to correctly interpret model behavior and alignment drift. It appears under the cs.AI category as a new submission.