AI Security Research

2,529+ academic papers on AI security, attacks, and defenses

Total
2,529
Attack
969
Benchmark
729
Defense
345
Tool
272
Survey
142

Showing 461–480 of 551 papers

Clear filters
Benchmark MEDIUM

Auditing Games for Sandbagging

Jordan Taylor, Sid Black, Dillon Bowen +10 more

Future AI systems could conceal their capabilities ('sandbagging') during evaluations, potentially misleading developers and auditors. We...

5 months ago cs.AI PDF
Benchmark LOW

Privacy Practices of Browser Agents

Alisha Ukani, Hamed Haddadi, Ali Shahin Shamsabadi +1 more

This paper presents a systematic evaluation of the privacy behaviors and attributes of eight recent, popular browser agents. Browser agents are...

5 months ago cs.CR PDF

Track AI security vulnerabilities in real time

Get breaking CVE alerts, compliance reports (ISO 42001, EU AI Act), and CISO risk assessments for your AI/ML stack.

Start 14-Day Free Trial