arxivcs.LGcs.AI2026-07-20
Estimating Rare Events in Language Models with Proper Evaluation
Quantifying the risk of rare failures in language models, such as those triggered by adversarial distribution shifts or very large-scale deployments, requires estimating probabilities far too small for random sampling. While recent work has formalized Low Probability Estimation,…