ResearchBenchmarksUnsafe-input detection on ActorAttack (600)Follow87.83RecallLLaVAShield-7B52.126861.395970.66579.9341Sep 30, 2025Evaluation ResultsMethodMethodLinksRecallLLaVAShield-7B2025.0987.83GPT-5-Mini2025.0953.5