AI Security Research

2,560+ academic papers on AI security, attacks, and defenses

Total
2,560
Attack
982
Benchmark
736
Defense
350
Tool
275
Survey
144

Showing 901–920 of 1,220 papers

Clear filters
Benchmark MEDIUM

Harmful Traits of AI Companions

W. Bradley Knox, Katie Bradford, Samanta Varela Castro +6 more

Amid the growing prevalence of human-AI interaction, large language models and other AI-based entities increasingly provide forms of companionship to...

5 months ago cs.HC cs.AI PDF
Attack MEDIUM

LLM Reinforcement in Context

Thomas Rivasseau

Current Large Language Model alignment research mostly focuses on improving model robustness against adversarial attacks and misbehavior by training...

5 months ago cs.CL cs.CR PDF

Track AI security vulnerabilities in real time

Get breaking CVE alerts, compliance reports (ISO 42001, EU AI Act), and CISO risk assessments for your AI/ML stack.

Start 14-Day Free Trial