00:00
2026-07-31
aclanthology.org
large-language-models
Code Without Context: Can We Trust LLMs to Test Software from Informal Descriptions?
A study presented at the Second International Conference on Natural Language Processing and Artificial Intelligence for Cyber Security (NLPAICS 2026) found that GPT-5-mini outperformed Qwen2.5-Coder-7โฆ