Natural Language Processing Faculty Publications

Hybrid Uncertainty Quantification for Selective Text Classification in Ambiguous Tasks

Artem Vazhentsev, AIRI
Gleb Kuzmin, AIRI
Akim Tsvigun, National University of Science & Technology (MISIS)
Alexander Panchenko, AIRI
Maxim Panov, TII
Mikhail Burtsev, London Institute for Mathematical Sciences
Artem Shelmanov, Mohamed bin Zayed University of Artificial IntelligenceFollow

Document Type

Conference Proceeding

Publication Title

Proceedings of the Annual Meeting of the Association for Computational Linguistics

Abstract

Many text classification tasks are inherently ambiguous, which results in automatic systems having a high risk of making mistakes, in spite of using advanced machine learning models. For example, toxicity detection in user-generated content is a subjective task, and notions of toxicity can be annotated according to a variety of definitions that can be in conflict with one another. Instead of relying solely on automatic solutions, moderation of the most difficult and ambiguous cases can be delegated to human workers. Potential mistakes in automated classification can be identified by using uncertainty estimation (UE) techniques. Although UE is a rapidly growing field within natural language processing, we find that state-of-the-art UE methods estimate only epistemic uncertainty and show poor performance, or under-perform trivial methods for ambiguous tasks such as toxicity detection. We argue that in order to create robust uncertainty estimation methods for ambiguous tasks it is necessary to account also for aleatoric uncertainty. In this paper, we propose a new uncertainty estimation method that combines epistemic and aleatoric UE methods. We show that by using our hybrid method, we can outperform state-of-the-art UE methods for toxicity detection and other ambiguous text classification tasks.

First Page

11659

Last Page

11681

Publication Date

1-1-2023

Recommended Citation

A. Vazhentsev et al., "Hybrid Uncertainty Quantification for Selective Text Classification in Ambiguous Tasks," Proceedings of the Annual Meeting of the Association for Computational Linguistics, vol. 1, pp. 11659 - 11681, Jan 2023.

This document is currently not available here.

COinS

Natural Language Processing Faculty Publications

Hybrid Uncertainty Quantification for Selective Text Classification in Ambiguous Tasks

Document Type

Publication Title

Abstract

First Page

Last Page

Publication Date

Recommended Citation

Browse

Contribute

Links

Natural Language Processing Faculty Publications

Hybrid Uncertainty Quantification for Selective Text Classification in Ambiguous Tasks

Authors

Document Type

Publication Title

Abstract

First Page

Last Page

Publication Date

Recommended Citation

Share

Browse

Contribute

Links