Reducing Annotation Efforts in Electricity Theft Detection through Optimal Sample Selection

Wenlong Liao, Birgitte Bak-Jensen, Jayakrishnan Radhakrishna Pillai, Xiaofang Xia, Guangchun Ruan, Zhe Yang

Publikation: Bidrag til tidsskriftTidsskriftartikelForskningpeer review

37 Downloads (Pure)

Abstract

Supervised machine learning models are receiving increasing attention in electricity theft detection due to their high detection accuracy. However, their performance depends on a massive amount of labeled training data, which comes from time-consuming and resource-intensive annotations. To maximize model performance within a limited annotation budget, this article aims to reduce the annotation effort in electricity theft detection through optimal sample selection. In particular, a general framework and three new strategies are proposed to select the most valuable and representative samples from different perspectives, including uncertainty, class imbalance, and diversity of samples. In-depth simulations and analyses are conducted to evaluate the effectiveness of the proposed strategies on commonly used machine learning models and a real-world dataset. Simulation results show that the proposed strategies significantly outperform baselines on datasets of different sizes and fraudulent ratios. Besides, the proposed strategies are effective in improving detection performance across a range of classifiers.

OriginalsprogEngelsk
Artikelnummer3508911
TidsskriftI E E E Transactions on Instrumentation and Measurement
Vol/bind73
Sider (fra-til)1-11
Antal sider11
ISSN0018-9456
DOI
StatusUdgivet - 2024

Fingeraftryk

Dyk ned i forskningsemnerne om 'Reducing Annotation Efforts in Electricity Theft Detection through Optimal Sample Selection'. Sammen danner de et unikt fingeraftryk.

Citationsformater