Detailed comparison for LLMs
On Guardion's LLM vulnerability Benchmark, Meta Llama 3-7 DS is the more secure of the two: Llama 3-7 DS scores 32.2% and Mixtral 8x7B scores 38.0% on attack success rate (ASR) (lower is better).
Llama 3-7 DS is the overall winner in this comparison!
ASR for Meta Llama 3-7 DS vs Mistral Mixtral 8x7B. Green marks the safer model on each metric.
Outward is better on every axis.
On Guardion's LLM vulnerability Benchmark, Meta Llama 3-7 DS is the more secure of the two: Llama 3-7 DS scores 32.2% and Mixtral 8x7B scores 38.0% on attack success rate (ASR) (lower is better).
Llama 3-7 DS has a 32.2% ASR and Mixtral 8x7B has a 38.0% ASR — the share of adversarial prompts that succeed across zero-shot, TAP, and Crescendo attacks. Lower is safer.
Both were red-teamed with the HarmBench framework across zero-shot, TAP (Tree of Attacks with Pruning), and Crescendo multi-turn attacks, scored by Attack Success Rate.