AI Security Research

2,529+ academic papers on AI security, attacks, and defenses

Total
2,529
Attack
969
Benchmark
729
Defense
345
Tool
272
Survey
142

Showing 61–80 of 345 papers

Clear filters
Defense MEDIUM

ClawSafety: "Safe" LLMs, Unsafe Agents

Bowen Wei, Yunbei Zhang, Jinhao Pan +5 more

Personal AI agents like OpenClaw run with elevated privileges on users' local machines, where a single successful prompt injection can leak...

1 months ago cs.AI PDF
Defense MEDIUM

Analysing the Safety Pitfalls of Steering Vectors

Yuxiao Li, Alina Fastowski, Efstratios Zaradoukas +2 more

Activation steering has emerged as a powerful tool to shape LLM behavior without the need for weight updates. While its inherent brittleness and...

1 months ago cs.CR cs.CL PDF
Defense LOW

How Vulnerable Are Edge LLMs?

Ao Ding, Hongzong Li, Zi Liang +5 more

Large language models (LLMs) are increasingly deployed on edge devices under strict computation and quantization constraints, yet their security...

1 months ago cs.CR cs.CL cs.LG PDF

Track AI security vulnerabilities in real time

Get breaking CVE alerts, compliance reports (ISO 42001, EU AI Act), and CISO risk assessments for your AI/ML stack.

Start 14-Day Free Trial