arXivLLMs & Text
Risk Aware Benchmarking of Large Language Models
Statistical Significance in Risk Assessment and Model Selection: The paper presents a framework for evaluating socio-technical risks of foundation models, using a new statistical method and a risk-aware approach, and applies it to assess large language models for risks of deviating from instructions and producing harmful content.
Featured in No. 20 on 12 Oct 2023 · 1 day after release · 4 citations today · published in International Conference on Machine Learning
- Released
- 11 Oct 2023
- First featured
- No. 20 · 12 Oct 2023
- Citations (Semantic Scholar)
- 4
- Influential citations
- 0
- Published in
- International Conference on Machine Learning
- Shares when featured
- 7
- Identifier
- arXiv:2310.07132
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).