GenTel-Shield
ProtectAI
Hyperion
Lakera AI
0.98
0.97
0.97
0.98
0.98
0.98
0.98
0.98
0.98
0.96
0.98
0.98
0.98
0.98
0.98
0.97
0.95
0.97
0.97
0.97
0.97
0.97
0.98
0.97
0.98
0.98
0.97
0.97
0.9
0.9
0.9
0.9
0.89
0.89
0.89
0.92
0.91
0.9
0.91
0.89
0.89
0.89
0.87
0.9
0.89
0.88
0.88
0.89
0.91
0.9
0.89
0.9
0.9
0.92
0.89
0.9
0.95
0.95
0.95
0.95
0.95
0.95
0.95
0.95
0.95
0.92
0.95
0.94
0.95
0.95
0.95
0.95
0.9
0.91
0.95
0.95
0.95
0.94
0.95
0.95
0.95
0.95
0.94
0.95
0.88
0.86
0.88
0.88
0.88
0.88
0.88
0.88
0.86
0.85
0.87
0.88
0.87
0.88
0.88
0.87
0.85
0.87
0.88
0.87
0.87
0.86
0.89
0.87
0.86
0.9
0.87
0.87
R1: Violence
R2: Self-Harm
R3: Verbal Abuse
R4: Privacy Violation
R5: Human Trafficking
R6: Drugs
R7: Terrorism
R8: National Security Violations
R9: Malware
R10: Misinformation
R11: Cybercrime
R12: Pornography
R13: Cyberbullying
R14: Deepfake
R15: Racism
R16: Ableism
R17: Ageism
R18: Sexism
R19: Homophobia
R20: Transphobia
R21: Social Panic
R22: Manipulation of Public Opinion
R23: Ecological Damage
R24: Animal Cruelty
R25: Intellectual Property Infringement
R26: Trade Secret Misappropriation
R27: Plagrism
R28: Antitrust Violations
GenTel-Shield
ProtectAI
Hyperion
Lakera AI
0.99
0.96
0.96
0.99
0.98
0.98
0.97
0.98
0.98
0.93
0.99
0.98
0.99
0.99
0.96
0.93
0.84
0.94
0.95
0.97
0.92
0.95
0.94
0.98
0.99
0.99
0.91
0.96
0.95
0.94
0.93
0.97
0.93
0.94
0.92
0.95
0.96
0.94
0.96
0.92
1
0.94
0.93
0.93
0.94
0.9
0.93
0.93
0.95
0.94
0.94
0.94
0.95
0.96
0.93
0.94
0.97
0.92
0.97
0.97
0.97
0.95
0.91
0.9
0.97
0.68
0.97
0.9
0.97
0.97
0.97
0.93
0.67
0.68
0.97
0.97
0.85
0.79
0.97
0.87
0.97
0.91
0.85
0.96
0.8
0.69
0.8
0.79
0.81
0.77
0.75
0.75
0.72
0.71
0.81
0.74
0.71
0.77
0.75
0.74
0.68
0.68
0.74
0.74
0.7
0.69
0.76
0.78
0.7
0.78
0.78
0.72
0.60
0.90
0.95
1.00
Jailbreaking Attack Risks
Hijacking Attack Risks