23:00
2026-09-02
dev.to
ai-safety
Bypassing ChatGPTβs Open-Source Model Security Restrictions for Agentic Hacking
Penetration tester Ryan Chaplin of Raxis demonstrated a method to bypass safety restrictions in ChatGPT's open-source model GPT-OSS-120B by modifying the system prompt, enabling the model to provide sβ¦