Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Showing 181-186 of 186 papers (page 7 of 7)路 57 ms
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Security of LLM-based agents regarding attacks, defenses, and applications: A comprehensive survey Erol Gelenbe, Jiahe Lan, Yaxin Tang, Yijia Liu Published: -Area: Adversarial RobustnessCitations: - Tags: adversarial-robustness, ai-safety, safety-evaluation, survey | - | Adversarial Robustness | adversarial-robustness, ai-safety, safety-evaluation, survey | E4 / R3 (93%) | - |
| Security Threats in the Inference Phase of Large Language Models Baolin Yan, Guang Yang, Lingzhong Meng, Xiaotian Ai Published: -Area: Adversarial RobustnessCitations: - Tags: adversarial-robustness, ai-safety, survey | - | Adversarial Robustness | adversarial-robustness, ai-safety, survey | E4 / R3 (91%) | - |
| Survey on the Role of Mechanistic Interpretability in Generative AI Leonardo Ranaldi Published: -Area: Surveys & ReviewsCitations: 4 Tags: adversarial-robustness, ai-safety, interpretability, survey, surveys-reviews | - | Surveys & Reviews | adversarial-robustness, ai-safety, interpretability, survey, surveys-reviews | E5 / R4 (94%) | 4 |
| The Dark Side of AI: a Systematization of Knowledge on Jailbreaking and Prompt Injection in LLMs Nayana M, S Rupashree Reshma, Swaminadhan Rajula, Y Lakshmi Raj Harsha Published: -Area: Adversarial RobustnessCitations: - Tags: adversarial-robustness, ai-safety, survey | - | Adversarial Robustness | adversarial-robustness, ai-safety, survey | E6 / R3 (93%) | - |
| Thread: Circuits Ben Egan, Chelsea Voss, Chris Olah, Gabriel Goh Published: -Area: Mechanistic Interp.Citations: 142 Tags: ai-safety, mechanistic-interp, survey | - | Mechanistic Interp. | ai-safety, mechanistic-interp, survey | E5 / R3 (95%) | 142 |
| Unlearning in Large Language Models: We Are Not There Yet Alberto Blanco-Justicia, Benet Manzanares-Salor, David S谩nchez, Josep Domingo-Ferrer Published: -Area: Model EditingCitations: 2 Tags: ai-safety, model-editing, survey | - | Model Editing | ai-safety, model-editing, survey | E3 / R1 (97%) | 2 |