Notable AIINT arXiv cs.AI

UniGuardian: A Unified Defense for Detecting Prompt Injection, Backdoor Attacks and Adversarial Attacks in Large Language Models

arXiv:2502.13141v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are vulnerable to attacks like prompt injection, backdoor attacks, and adversarial attacks, which manipulate prompts or models to generate harmful outputs. In…

Read the full story at arXiv cs.AI ↗

ImpactNotable 31/100
Why it mattersRule-based estimate: event keywords (+4); trust 6/10.
RegionsGlobal
Published4 h ago (Fri, 02 Oct 2026 04:00:00 GMT)
RetrievedFri, 02 Oct 2026 07:00:28 GMT via rss
ClassifiedFri, 02 Oct 2026 07:00:40 GMT by heuristic
AuthorHuawei Lin, Yingjie Lao, Tony Geng, Tan Yu, Weijie Zhao