04:00
2026-08-12
arxiv.org
artificial-intelligence
Generating Attacks for LLMs with GFlowNets
Researchers propose an automated red teaming method using GFlowNets to generate adversarial attacks against large language models (LLMs), with one LLM testing another to identify vulnerabilities and pโฆ