Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Showing 31-41 of 41 papers (page 2 of 2)路 23 ms
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Provably safe systems: the only path to controllable AGI Max Tegmark, Steve Omohundro Published: 2023-09-05Area: Formal/TheoreticalCitations: 38 Tags: ai-safety, formaltheoretical, interpretability, position | 2023-09-05 | Formal/Theoretical | ai-safety, formaltheoretical, interpretability, position | E7 / R4 (96%) | 38 |
| Intent-aligned AI systems deplete human agency: the need for agency foundations research in AI safety Ben Smith, Catalin Mitelut, Peter Vamplew Published: 2023-05-30Area: Formal/TheoreticalCitations: 7 Tags: ai-safety, formaltheoretical, theoretical | 2023-05-30 | Formal/Theoretical | ai-safety, formaltheoretical, theoretical | E5 / R3 (96%) | 7 |
| Fundamental Limitations of Alignment in Large Language Models Amnon Shashua, Noam Wies, Oshri Avnery, Yoav Levine Published: 2023-04-19Area: Formal/TheoreticalCitations: 178 Tags: adversarial-robustness, ai-safety, alignment-training, formaltheoretical, theoretical | 2023-04-19 | Formal/Theoretical | adversarial-robustness, ai-safety, alignment-training, formaltheoretical, theoretical | E4 / R2 (94%) | 178 |
| Power-seeking can be probable and predictive for trained agents Janos Kramar, Victoria Krakovna Published: 2023-04-13Area: Formal/TheoreticalCitations: 22 Tags: ai-safety, formaltheoretical, theoretical | 2023-04-13 | Formal/Theoretical | ai-safety, formaltheoretical, theoretical | E6 / R3 (96%) | 22 |
| On the Impossible Safety of Large AI Models El-Mahdi El-Mhamdi, John Stephan, L锚-Nguy锚n Hoang, Nirupam Gupta Published: 2022-09-30Area: Formal/TheoreticalCitations: 37 Tags: ai-safety, formaltheoretical, theoretical | 2022-09-30 | Formal/Theoretical | ai-safety, formaltheoretical, theoretical | E5 / R4 (94%) | 37 |
| Parametrically Retargetable Decision-Makers Tend to Seek Power Alexander Matt Turner, Prasad Tadepalli Published: 2022-06-27Area: Formal/TheoreticalCitations: 21 Tags: ai-safety, formaltheoretical, theoretical | 2022-06-27 | Formal/Theoretical | ai-safety, formaltheoretical, theoretical | E6 / R3 (96%) | 21 |
| Consequences of Misaligned AI Dylan Hadfield-Menell, Simon Zhuang Published: 2021-02-07Area: Formal/TheoreticalCitations: 95 Tags: ai-safety, formaltheoretical, theoretical | 2021-02-07 | Formal/Theoretical | ai-safety, formaltheoretical, theoretical | E6 / R3 (96%) | 95 |
| Agent Incentives: A Causal Perspective Eric D. Langlois, Pedro A. Ortega, Ryan Carey, Shane Legg Published: 2021-02-02Area: Formal/TheoreticalCitations: 61 Tags: ai-safety, formaltheoretical, theoretical | 2021-02-02 | Formal/Theoretical | ai-safety, formaltheoretical, theoretical | E6 / R3 (96%) | 61 |
| Deep Learning is Singular, and That's Good Daniel Murfet, Hui Li, Jesse Gell-Redman, Mingming Gong Published: 2020-10-22Area: Formal/TheoreticalCitations: 38 Tags: ai-safety, formaltheoretical, theoretical | 2020-10-22 | Formal/Theoretical | ai-safety, formaltheoretical, theoretical | E5 / R3 (95%) | 38 |
| Optimal Policies Tend to Seek Power Alexander Matt Turner, Andrew Critch, Logan Smith, Prasad Tadepalli Published: 2019-12-03Area: Formal/TheoreticalCitations: 97 Tags: ai-safety, formaltheoretical, theoretical | 2019-12-03 | Formal/Theoretical | ai-safety, formaltheoretical, theoretical | E5 / R3 (93%) | 97 |
| Inaugural Workshop on Provably Safe and Beneficial AI (PSBAI) Stuart Russell Published: -Area: Formal/TheoreticalCitations: - Tags: ai-safety, formaltheoretical, survey | - | Formal/Theoretical | ai-safety, formaltheoretical, survey | E6 / R3 (96%) | - |