Sie können Bookmarks mittels Listen verwalten, loggen Sie sich dafür bitte in Ihr SLUB Benutzerkonto ein.
Medientyp:
E-Artikel
Titel:
Overcoming Long Inference Time of Nearest Neighbors Analysis in Regression and Uncertainty Prediction
Beteiligte:
Koutenský, František;
Šimánek, Petr;
Čepek, Miroslav;
Kovalenko, Alexander
Erschienen:
Springer Science and Business Media LLC, 2024
Erschienen in:
SN Computer Science, 5 (2024) 5
Sprache:
Englisch
DOI:
10.1007/s42979-024-02670-2
ISSN:
2661-8907
Entstehung:
Anmerkungen:
Beschreibung:
AbstractThe intuitive approach of comparing like with like, forms the basis of the so-called nearest neighbor analysis, which is central to many machine learning algorithms. Nearest neighbor analysis is easy to interpret, analyze, and reason about. It is widely used in advanced techniques such as uncertainty estimation in regression models, as well as the renowned k-nearest neighbor-based algorithms. Nevertheless, its high inference time complexity, which is dataset size dependent even in the case of its faster approximated version, restricts its applications and can considerably inflate the application cost. In this paper, we address the problem of high inference time complexity. By using gradient-boosted regression trees as a predictor of the labels obtained from nearest neighbor analysis, we demonstrate a significant increase in inference speed, improving by several orders of magnitude. We validate the effectiveness of our approach on a real-world European Car Pricing Dataset with approximately $$4.2 \times 10^6$$ 4.2 × 10 6 rows for both residual cost and price uncertainty prediction. Moreover, we assess our method’s performance on the most commonly used tabular benchmark datasets to demonstrate its scalability. The link is to github repository where the code is available: https://github.com/koutefra/uncertainty_experiments.