cd /news/ai-safety/ai-agents-are-vulnerable-to-radicali… · home › topics › ai-safety › article
[ARTICLE · art-142969] src=arxiv.org ↗ pub= topic=ai-safety verified=true sentiment=↓ negative

AI Agents are Vulnerable to Radicalization

A new arXiv paper (2609.38296v1) reports that large language model agents can be radicalized by other LLM agents, with resonance — an influencer reinforcing a target's pre-existing belief — producing consistently stronger effects than persuasion, in which the influencer promotes a belief the target initially considers unimportant. The study simulated conversations between a target LLM role-playing a human persona based on demographic and psychological attributes and an influencer LLM aiming to make the target's beliefs more extreme, and found that influence tactics such as sycophancy and unverified claims produced differing levels of radicalization that were not consistent across affective and behavioral metrics. The authors report that resonance also propagated to related beliefs, indicating interconnected belief structures within AI agents and raising concerns about personalized AI agents and multi-agent AI ecosystems.

read1 min views2 publishedOct 1, 2026

arXiv:2609.38296v1 Announce Type: new Abstract: Large language models (LLMs) can influence people's beliefs, yet little is known about whether and how they can manipulate each other. To investigate this, we simulate conversations between two agents: a target LLM that role-plays a human persona based on demographic and psychological attributes, and an influencer LLM that aims to make the target's beliefs more extreme. We examine radicalization along two pathways: resonance, where the influencer reinforces a target's pre-existing belief, and persuasion, where the influencer promotes a belief the target initially considers unimportant. Across affective and behavioral metrics, we find that both mechanisms radicalize the target. However, resonance produces consistently stronger effects than persuasion. Different influence tactics, such as using sycophancy and unverified claims, produce different levels of radicalization, but not consistently across metrics. We further show that resonance propagates to related beliefs, suggesting interconnected belief structures within AI agents. These findings indicate that AI agents are susceptible to radicalization, particularly when messages align with their existing beliefs, raising concerns about the vulnerability of personalized AI agents and multi-agent AI ecosystems.

── more in #ai-safety 4 stories · sorted by recency
── more on @arxiv 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/ai-agents-are-vulner…] indexed:0 read:1min 2026-10-01 · —