AI Security Research

2,077+ academic papers on AI security, attacks, and defenses

Total
2,077
Attack
809
Benchmark
603
Defense
272
Tool
226
Survey
113

Showing 421–440 of 2,031 papers

Clear filters
Benchmark MEDIUM

Backdooring Bias in Large Language Models

Anudeep Das, Prach Chantasantitam, Gurjot Singh +3 more

Large language models (LLMs) are increasingly deployed in settings where inducing a bias toward a certain topic can have significant consequences,...

1 months ago cs.CR cs.AI PDF
Defense MEDIUM

GPTZero: Robust Detection of LLM-Generated Texts

George Alexandru Adam, Alexander Cui, Edwin Thomas +7 more

While historical considerations surrounding text authenticity revolved primarily around plagiarism, the advent of large language models (LLMs) has...

1 months ago cs.LG PDF

Track AI security vulnerabilities in real time

Get breaking CVE alerts, compliance reports (ISO 42001, EU AI Act), and CISO risk assessments for your AI/ML stack.

Start 14-Day Free Trial