arrow
Return

Improving Selective Classification with Pairwise Queries for Binary Classification

delete2026-06-16
delete0
PRE
AI
H
Harsh Vardhan *
S
Sunav Choudhary
N
Natwar Modani
A
Arya Mazumdar
DOI:10.1007/s10994-026-07078-ydelete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In selective classification, a model predicts the labels of data samples where it is confident, and abstains from predicting labels for samples on which it is not confident. The rejected samples are often labeled by an expert, which is expensive. The budget for the expert is best utilized when the model has low error on non-rejected samples. However, the estimate of a model’s confidence might be inconsistent with the model’s predictions, which can lead to high error on non-rejected points. Such situations can readily occur in in-context binary classification by LLMs. To remedy this, we propose making additional pairwise queries to the same model. These pairwise queries can detect high-error samples and be incorporated into selective classification techniques to reduce the error on non-rejected samples. Theoretically, we establish the conditions under which a simple algorithm using pairwise queries outperforms an inconsistent confidence estimate. We support this insight through extensive experiments for 1 synthetic and 4 in-context learning-based real binary classification datasets. In all these cases, we show that our algorithms, using pairwise queries, obtain a better accuracy-cost tradeoff than using only the raw confidence estimates, for instance, the LLM’s next-token logits.
Keywords:
Selective classification
Pairwise queries
Learning theory
In-context learning

Journal

Machine Learning cover
Machine Learning
IF:
2.9
Papers:
2.7K
Citations:
3.4W

Organization

No organization information available
Cited Papers

Cited Papers

A Heterogeneous Graph to Abstract Syntax Tree Framework for Text-to-SQL
err
IF0
err2023-01-01
err0
PREAI
errRuisheng Cao; Lu Chen; Jieyu Li; Hanchong Zhang; Hongshen Xu; Wangyou Zhang; Kai Yu
errShare
errSave
A Survey of Deep Active Learning
err2021-10-08
err579
errOAAI
errRen, Pengzhen; Xiao, Yun; Chang, Xiaojun; Huang, Po-Yao; Li, Zhihui; Gupta, Brij B.; Chen, Xiaojiang; Wang, Xin
errShare
errSave
Fairness in Criminal Justice Risk Assessments: The State of the Art
err2018-07-02
err507
errOAAI
errBerk, Richard; Heidari, Hoda; Jabbari, Shahin; Kearns, Michael; Roth, Aaron
errShare
errSave