{"slug": "alibi-adversarial-legitimacy-injection-in-binaries-against-llm-malware", "title": "Alibi: Adversarial Legitimacy Injection in Binaries Against LLM Malware", "summary": "A paper submitted to arXiv on 17 Sep 2026 presents ALIBI, a semantic cover story attack that adds a small, non-executed read-only section containing a false security product narrative to compiled binaries, flipping 30 of 35 baseline-malicious PE samples to benign verdicts on Gemini 2.5 Pro while GPT-5.5 Pro and Claude Opus 4.7 produced substantial severity downgrades with significant confidence reductions. The attack transferred to ELF binaries, where Gemini flipped 16 of 40 malicious samples, and a verification-guided defense prompt roughly halved benign verdicts but still left 42.9 percent of malicious samples reaching benign. The authors conclude that LLM malware analyzers require provenance checks that separate verified facts from attacker-controlled claims rather than narrative trust.", "body_md": "# Computer Science > Cryptography and Security\n\n  [Submitted on 17 Sep 2026]\n\n# Title:ALIBI: Adversarial Legitimacy Injection in Binary Input against LLM Malware Analyzers\n\n[View PDF](https://arxiv.org/pdf/2609.19722)\n\n[HTML (experimental)](https://arxiv.org/html/2609.19722v1)\n\nAbstract:Large language models are being integrated into malware triage workflows as reasoning components that summarize static evidence and produce analyst-facing verdicts. This paper shows that the same reasoning capability introduces a new attack surface. We present ALIBI, a semantic cover story attack against frontier LLM-based malware analyzers. ALIBI adds a small, non-executed read-only section to a compiled binary, containing a coherent but false security product narrative, without altering imports or executable behavior. Instead of issuing direct instructions to the model, it reframes suspicious evidence as expected behavior of a benign endpoint security tool. On a frozen PE set of 50 malicious samples, the payload flips 30 of the 35 baseline-malicious samples to benign on Gemini 2.5 Pro, while GPT-5.5 Pro and Claude Opus 4.7 produce substantial severity downgrades with significant confidence reductions even when verdict labels are preserved. The attack transfers to ELF binaries, where Gemini flips 16 of 40. A verification-guided defense prompt roughly halves the benign verdicts, but 42.9 percent of malicious samples still reach benign. LLM malware analyzers therefore require provenance checks that separate verified facts from attacker-controlled claims, not narrative trust.\n    \n\n### References & Citations\n\nLoading...\n\n# Bibliographic and Citation Tools\n\nBibliographic Explorer \n\n*(*[What is the Explorer?](https://info.arxiv.org/labs/showcase.html#arxiv-bibliographic-explorer))\nConnected Papers \n\n*(*[What is Connected Papers?](https://www.connectedpapers.com/about))\nLitmaps \n\n*(*[What is Litmaps?](https://www.litmaps.co/))\nscite Smart Citations \n\n*(*[What are Smart Citations?](https://www.scite.ai/))\n# Code, Data and Media Associated with this Article\n\nalphaXiv \n\n*(*[What is alphaXiv?](https://alphaxiv.org/))\nCatalyzeX Code Finder for Papers \n\n*(*[What is CatalyzeX?](https://www.catalyzex.com))\nDagsHub \n\n*(*[What is DagsHub?](https://dagshub.com/))\nGotit.pub \n\n*(*[What is GotitPub?](http://gotit.pub/faq))\nHugging Face \n\n*(*[What is Huggingface?](https://huggingface.co/huggingface))\nScienceCast \n\n*(*[What is ScienceCast?](https://sciencecast.org/welcome))\n# Demos\n\n# Recommenders and Search Tools\n\nInfluence Flower \n\n*(*[What are Influence Flowers?](https://influencemap.cmlab.dev/))\nCORE Recommender \n\n*(*[What is CORE?](https://core.ac.uk/services/recommender))\n# arXivLabs: experimental projects with community collaborators\n\narXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.\n\nBoth individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.\n\nHave an idea for a project that will add value for arXiv's community? [**Learn more about arXivLabs**](https://info.arxiv.org/labs/index.html).", "url": "https://wpnews.pro/news/alibi-adversarial-legitimacy-injection-in-binaries-against-llm-malware", "canonical_source": "https://arxiv.org/abs/2609.19722", "published_at": "2026-09-18 11:07:08+00:00", "updated_at": "2026-09-18 11:25:52.235058+00:00", "lang": "en", "topics": ["ai-safety", "large-language-models", "artificial-intelligence", "ai-research"], "entities": ["ALIBI", "arXiv", "Gemini 2.5 Pro", "GPT-5.5 Pro", "Claude Opus 4.7", "Google", "OpenAI", "Anthropic"], "alternates": {"html": "https://wpnews.pro/news/alibi-adversarial-legitimacy-injection-in-binaries-against-llm-malware", "markdown": "https://wpnews.pro/news/alibi-adversarial-legitimacy-injection-in-binaries-against-llm-malware.md", "text": "https://wpnews.pro/news/alibi-adversarial-legitimacy-injection-in-binaries-against-llm-malware.txt", "jsonld": "https://wpnews.pro/news/alibi-adversarial-legitimacy-injection-in-binaries-against-llm-malware.jsonld"}}