Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Showing 61-66 of 66 papers (page 3 of 3)路 136 ms
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Monitoring Monitorability Annie Y. Wei, Benjamin Arnav, Bowen Baker, Ian Kivlichan Published: 2025-12-20Area: Scalable OversightCitations: 6 Tags: ai-safety, benchmark, safety-evaluation, scalable-oversight | 2025-12-20 | Scalable Oversight | ai-safety, benchmark, safety-evaluation, scalable-oversight | E5 / R3 (95%) | 6 |
| Truthfulness Despite Weak Supervision: Evaluating and Training LLMs Using Peer Prediction Cameron Allen, Micah Carroll, Tianyi Alex Qiu Published: 2026-01-28Area: Scalable OversightCitations: - Tags: ai-safety, empirical, safety-evaluation, scalable-oversight | 2026-01-28 | Scalable Oversight | ai-safety, empirical, safety-evaluation, scalable-oversight | E5 / R3 (95%) | - |
| Steering LLMs via Scalable Interactive Oversight Enyu Zhou, Guoteng Wang, Hang Yan, Long Ma Published: 2026-02-04Area: Scalable OversightCitations: - Tags: ai-safety, alignment-training, empirical, scalable-oversight | 2026-02-04 | Scalable Oversight | ai-safety, alignment-training, empirical, scalable-oversight | E5 / R3 (94%) | - |
| Basic Legibility Protocols Improve Trusted Monitoring Ashwin Sreevatsa, Cody Rushing, Sebastian Prasanna Published: 2026-02-09Area: Scalable OversightCitations: - Tags: ai-safety, empirical, scalable-oversight | 2026-02-09 | Scalable Oversight | ai-safety, empirical, scalable-oversight | E6 / R4 (96%) | - |
| Eliciting Latent Knowledge (ELK) Ajeya Cotra, Mark Xu, Paul Christiano Published: -Area: Scalable OversightCitations: - Tags: ai-safety, scalable-oversight, theoretical | - | Scalable Oversight | ai-safety, scalable-oversight, theoretical | E6 / R2 (96%) | - |
| Scalable Human Oversight for Aligned Large Language Models: A Hybrid Framework for Intent Fidelity Folasade Y. Ayankoya, Olubukola D. Adekola, Oluwasefunmi B. Famodimu, Shade O. Kuyoro Published: -Area: Scalable OversightCitations: - Tags: ai-safety, empirical, scalable-oversight | - | Scalable Oversight | ai-safety, empirical, scalable-oversight | E4 / R2 (95%) | - |