ML-QuantSubscribe

arXivLLMs & Text

Risk Aware Benchmarking of Large Language Models

Statistical Significance in Risk Assessment and Model Selection: The paper presents a framework for evaluating socio-technical risks of foundation models, using a new statistical method and a risk-aware approach, and applies it to assess large language models for risks of deviating from instructions and producing harmful content.

Featured in No. 20 on 12 Oct 2023 · 1 day after release · 4 citations today · published in International Conference on Machine Learning

Released
11 Oct 2023
First featured
No. 20 · 12 Oct 2023
Citations (Semantic Scholar)
4
Influential citations
0
Published in
International Conference on Machine Learning
Shares when featured
7
Identifier
arXiv:2310.07132

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page