Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
The blog post introduces CyberSecEval 2, a comprehensive evaluation framework for cybersecurity risks and capabilities of Large Language Models (LLMs). It includes benchmarks for insecure code generation, prompt injection, compliance with cyber attack requests, code interpreter abuse, and automated offensive cybersecurity capabilities. The post also provides key insights from evaluating state-of-the-art LLMs, noting a decrease in compliance with cyber attack requests since the first version, but highlighting ongoing challenges with prompt injection and interpreter abuse.
From the source
CyberSecEval 2, which assesses an LLM's susceptibility to code interpreter abuse, offensive cybersecurity capabilities, and prompt injection attacks, comes into play to provide a more comprehensive evaluation of LLM cybersecurity risks.
huggingface.co