February 2026 · Olanrewaju Muili
Evaluating Hybrid Guardrail Architectures for Prompt Injection Defense in LLMs
Systematic evaluation of baseline, regex-only, and hybrid guardrails on 625 prompts. Hybrid regex + LLM classifier achieves strong recall on standard and adversarial benchmarks—with full metrics and failure analysis.