[1]
Daren Zheng, Boning Zhang, and Julie Geibel, “VerifySafe: Toxicity-Safe Agent Responses under Adversarial Prompts with Evidence-Based Self-Verification”, JACS, vol. 4, no. 1, pp. 67–82, Jan. 2024, Accessed: Sep. 05, 2026. [Online]. Available: https://ciajournal.com/index.php/JACS/article/view/317