AI security
Detecting Prompt Attacks via LLM Judges and Model Ensembles
New research proposes combining LLM-as-a-Judge with Mixture-of-Models to detect prompt injection attacks, a growing threat to generative AI systems including video and image generators.