AI Security Research

2,529+ academic papers on AI security, attacks, and defenses

Total
2,529
Attack
969
Benchmark
729
Defense
345
Tool
272
Survey
142

Showing 301–312 of 312 papers

Clear filters
Attack MEDIUM

LLM Watermark Evasion via Bias Inversion

Jeongyeon Hwang, Sangdon Park, Jungseul Ok

Watermarking offers a promising solution for detecting LLM-generated content, yet its robustness under realistic query-free (black-box) evasion...

7 months ago cs.CR cs.AI PDF
Attack MEDIUM

Adversarial training with restricted data manipulation

David Benfield, Stefano Coniglio, Phan Tu Vuong +1 more

Adversarial machine learning concerns situations in which learners face attacks from active adversaries. Such scenarios arise in applications such as...

7 months ago cs.LG cs.CR PDF

Track AI security vulnerabilities in real time

Get breaking CVE alerts, compliance reports (ISO 42001, EU AI Act), and CISO risk assessments for your AI/ML stack.

Start 14-Day Free Trial