Jul 2026
A systematic evaluation of large language models for autonomous cyber defense
Overall, the results show that LLMs can produce competitive defense policies without fine-tuning but require manually engineered prompts, and their higher variance and slower response times could render them unsuitable for some real-world scenarios.
Thibaut Jacques
· Applied intelligence (Boston... · 0 citations