Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Showing 61-66 of 66 papers (page 3 of 3)路 20 ms
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Two-Turn Debate Doesn't Help Humans Answer Hard Reading Comprehension Questions Alicia Parrish, Amanpreet Singh Saimbhi, Harsh Trivedi, Jason Phang Published: 2022-10-19Area: Scalable OversightCitations: 15 Tags: ai-safety, empirical, scalable-oversight | 2022-10-19 | Scalable Oversight | ai-safety, empirical, scalable-oversight | E5 / R3 (95%) | 15 |
| Self-critiquing models for assisting human evaluators Catherine Yeh, Jan Leike, Jeff Wu, Jonathan Ward Published: 2022-06-12Area: Scalable OversightCitations: 380 Tags: ai-safety, empirical, scalable-oversight | 2022-06-12 | Scalable Oversight | ai-safety, empirical, scalable-oversight | E5 / R3 (95%) | 380 |
| Single-Turn Debate Does Not Help Humans Answer Hard Reading-Comprehension Questions Alicia Parrish, Angelica Chen, Ethan Perez, Harsh Trivedi Published: 2022-04-11Area: Scalable OversightCitations: 19 Tags: ai-safety, empirical, scalable-oversight | 2022-04-11 | Scalable Oversight | ai-safety, empirical, scalable-oversight | E5 / R3 (97%) | 19 |
| Recursively Summarizing Books with Human Feedback Daniel M. Ziegler, Jan Leike, Jeff Wu, Long Ouyang Published: 2021-09-22Area: Scalable OversightCitations: 349 Tags: ai-safety, empirical, scalable-oversight | 2021-09-22 | Scalable Oversight | ai-safety, empirical, scalable-oversight | E6 / R4 (95%) | 349 |
| Eliciting Latent Knowledge (ELK) Ajeya Cotra, Mark Xu, Paul Christiano Published: -Area: Scalable OversightCitations: - Tags: ai-safety, scalable-oversight, theoretical | - | Scalable Oversight | ai-safety, scalable-oversight, theoretical | E6 / R2 (96%) | - |
| Scalable Human Oversight for Aligned Large Language Models: A Hybrid Framework for Intent Fidelity Folasade Y. Ayankoya, Olubukola D. Adekola, Oluwasefunmi B. Famodimu, Shade O. Kuyoro Published: -Area: Scalable OversightCitations: - Tags: ai-safety, empirical, scalable-oversight | - | Scalable Oversight | ai-safety, empirical, scalable-oversight | E4 / R2 (95%) | - |