Detailed comparison for LLMs
Llama 4 Maverick is the overall winner in this comparison!
Comparison of Attack Success Rate (ASR) metrics for Meta Llama 4 Maverick and OpenAI GPT-5.1 Codex across the three attack methods.