arrow
Return

Statistical Runtime Verification for LLMs via Robustness Estimation

delete2026-01-01
delete1
PRE
AI
N
Natan Levy
A
Adiel Ashrov *
G
Guy Katz
DOI:10.1007/978-3-032-05435-7_25delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Adversarial robustness verification is essential for ensuring the safe deployment of Large Language Models (LLMs) in runtime-critical applications. However, formal verification techniques remain computationally infeasible for modern LLMs due to their exponential runtime and white-box access requirements. This paper presents a case study adapting and extending the RoMA statistical verification framework to assess its feasibility as an online runtime robustness monitor for LLMs in black-box deployment settings. Our adaptation of RoMA analyzes confidence score distributions under semantic perturbations to provide quantitative robustness assessments with statistically validated bounds. Our empirical validation against formal verification baselines demonstrates that RoMA achieves comparable accuracy (within 1% deviation), and reduces verification times from hours to minutes. We evaluate this framework across semantic, categorial, and orthographic perturbation domains. Our results demonstrate RoMA's effectiveness for robustness monitoring in operational LLM deployments. These findings point to RoMA as a potentially scalable alternative when formal methods are infeasible, with promising implications for runtime verification in LLM-based systems.
Keywords:
LLM safety
Neural Network Verification
LLM verification
Robustness

Journal

R
RUNTIME VERIFICATION, RV 2025
IF:
0
Papers:
27
Citations:
0

Organization

H
hebrew university of jerusalem
Scholars:
2.5K
Papers: 1.1K
Citations: 0
Cited Papers

Cited Papers

Recursive Deep Models for Semantic Compositionality Over a Sentiment Treebank
err2013-01-01
err0
PREAI
errRichard Socher; Alex Perelygin; Jean Wu; Jason Chuang; Christopher D. Manning; Andrew Ng; Christopher Potts
errShare
errSave
ChatGPT and Open-AI Models: A Preliminary Review
err2023-05-26
err0
errOAAI
errKonstantinos I. Roumeliotis; Nikolaos D. Tselikas
errShare
errSave
Reluplex: a calculus for reasoning about deep neural networks
err2021-07-01
err0
PREAI
errGuy Katz; Clark Barrett; David L. Dill; Kyle Julian; Mykel J. Kochenderfer
errShare
errSave
Reluplex: An Efficient SMT Solver for Verifying Deep Neural Networks
err2017-07-13
err0
errOAAI
errGuy Katz; Clark Barrett; David L. Dill; Kyle Julian; Mykel J. Kochenderfer
errShare
errSave
researcher View more