DEVIATE: A Deep Learning Variance Testing Framework (ASE 2021 - Tool Demonstrations)

Who

Hung Viet Pham, Mijung Kim, Lin Tan, Yaoliang Yu, Nachiappan Nagappan

Track

ASE 2021 Tool Demonstrations

Time Zone

The program is currently displayed in (GMT+11:00) Hobart.

Use conference time zone: (GMT+11:00) HobartSelect other time zone

The GMT offsets shown reflect the offsets at the moment of the conference.

Time Band

By setting a time band, the program will dim events that are outside this time window. This is useful for (virtual) conferences with a continuous program (with repeated sessions).
The time band will also limit the events that are included in the personal iCalendar subscription service.

Display full programSpecify a time band

Save

When

Wed 17 Nov 2021 09:50 - 09:55 at Kangaroo - Learning I Chair(s): Denys Poshyvanyk
Wed 17 Nov 2021 10:06 - 10:08 at Kangaroo - Tool Demo (2) Chair(s): Mattia Fazzini

Abstract

Deep learning (DL) training is nondeterministic and such nondeterminism was shown to cause significant variance of model accuracy (up to 10.8%). Such variance may affect the validity of the comparison of newly proposed DL techniques with baselines. To ensure such validity, DL researchers and practitioners must replicate their experiments multiple times with identical settings to quantify the variance of the proposed approaches and baselines. Replicating and measuring DL variances reliably and efficiently is challenging and understudied. We propose a ready-to-deploy framework DEVIATE that (1)measures DL training variance of a DL model with minimal manual efforts, and (2) provides statistical tests of both accuracy and variance. Specifically, DEVIATEautomaticallyanalyzes the DL training code and extracts monitored important metrics (such as accuracy and loss). In addition, DEVIATE performs popular statistical tests and provides users with a report of statistical p-values and effect sizes along with various confidence levels when comparing to selected baselines. We demonstrate the effectiveness of DEVIATE by performing case studies with adversarial training. Specifically, for an adversarial training process that uses the Fast Gradient Signed Method to generate adversarial examples as the training data, DEVIATEmeasures a max difference of accuracy among 8 identical training runs with fixed random seeds to be up to 5.1%.

Hung Viet Pham

University of Waterloo

Canada

Mijung Kim

Purdue University

United States

Lin Tan

Purdue University

United States

Yaoliang Yu

University of Waterloo

Nachiappan Nagappan

Microsoft Research