Safety Nudges: User-Facing Interventions for Real-Time AI Risk Awareness A two-week field study with 45 frequent chatbot users found that Safety Nudges, a browser-based tool that flags concerning chatbot behavior in real time, was rated useful, clear, and minimally disruptive, with nearly all participants reporting increased awareness of potential AI harms. The researchers report in arXiv:2609.26865v1 that this improved awareness alone did not necessarily produce discernible behavioral changes, and conclude that user-facing safety nudges can complement model-level safeguards while depending on relevance, calibration, and user control in their design. arXiv:2609.26865v1 Announce Type: cross Abstract: Conversational AI systems can pose safety risks to their users such as hallucination, sycophancy, overconfidence, and anthropomorphism, but these risks are difficult for users to detect during everyday use. We introduce Safety Nudges, a browser-based tool that provides lightweight, in situ flags when concerning behavior is detected in chatbot conversations. We evaluated Safety Nudges in a two-week field study with 45 frequent chatbot users, collecting interaction logs, surveys, and feedback on individual nudges. Participants found the tool useful, clear, and minimally disruptive, with nearly all users reporting an increased awareness of potential AI harms, though we found that this improved awareness alone did not necessarily lead to discernible behavioral changes. Our results suggest that user facing safety nudges can complement model-level safeguards by helping people critically evaluate AI responses in context, while highlighting the importance of relevance, calibration, and user control in nudge design for conversational AI safety.