AI Security Research

2,077+ academic papers on AI security, attacks, and defenses

Total
2,077
Attack
809
Benchmark
603
Defense
272
Tool
226
Survey
113

Showing 1921–1940 of 2,077 papers

Attack HIGH

Dynamic Target Attack

Kedong Xiu, Churui Zeng, Tianhang Zheng +6 more

Existing gradient-based jailbreak attacks typically optimize an adversarial suffix to induce a fixed affirmative response, e.g., ``Sure, here...

5 months ago cs.CR cs.AI PDF
Survey MEDIUM

Position: Privacy Is Not Just Memorization!

Niloofar Mireshghallah, Tianshi Li

The discourse on privacy risks in Large Language Models (LLMs) has disproportionately focused on verbatim memorization of training data, while a...

5 months ago cs.CR cs.AI cs.CL PDF

Track AI security vulnerabilities in real time

Get breaking CVE alerts, compliance reports (ISO 42001, EU AI Act), and CISO risk assessments for your AI/ML stack.

Start 14-Day Free Trial