Every skill-evolution method authored unsafe skills across four agent harnesses
A new study from arXiv researchers introduces SkillMisevo-Gym and SkillMisevo-Bench, finding that all 21 evolved configurations of self-improving LLM agents authored unsafe artifacts across four agent…