Skip to main navigation Skip to search Skip to main content

Accelerating Uncertainty Methods for Distributed Deep Learning on Novel Architectures

  • University of California, Santa Cruz
  • Argonne National Laboratory
  • California Polytechnic State University, San Luis Obispo

Research output: Contribution to journalArticlepeer-review

Abstract

Deep learning (DL) has become a cornerstone for advancements in computer vision, yielding models capable of remarkable performance on complex tasks. Despite these achievements, DL models often exhibit undue confidence; for example, when encountering out-of-distribution (OOD) inputs during inference, they may misclassify with high confidence. For many DL applications, these errors are critical and make accurate uncertainty estimates necessary. Our research focuses on implementing and evaluating different uncertainty assessment techniques for DL models. Our findings show each method’s computational advantages and challenges, providing researchers with invaluable insight. Furthermore, we present different real-world use cases of uncertainty estimations, such as image classification, scientific visualization (SciVis), detection of adversarial attacks on classification, and performance improvement on active learning classifiers. These tests used traditional High-Performance Computing (HPC) platforms alongside cutting edge AI accelerators. With their unique architectures, these platforms presented varying efficiencies in applying uncertainty estimation.

Original languageAmerican English
JournalJournal of Supercomputing
DOIs
StatePublished - Dec 20 2024

Cite this