04:00
2026-07-23
machinebrief.com
ai-safety
OpenSkillRisk: Benchmarking Agent Safety When Using Real-World Risky Third-Party Skills
A new benchmark, OpenSkillRisk, reveals that LLM-based agents fail to safely handle risky third-party skills, with even the safest configurations executing unsafe actions in about 17% of cases. The beโฆ